返回
Reinforcement learning marine predators algorithm for global optimization
DOI:10.1007/s10586-024-04381-y.png)
摘要
En 中文
Given the weak convergence, limited balance capacity, and optimization limitations observed in the Marine Predators Algorithm (MPA), which draws inspiration from the predatory behavior of marine organisms during evolutionary processes, this study introduces a Reinforcement Learning Marine Predators Algorithm (RLMPA). Firstly, based on the predatory characteristics at different stages, we have designed three location update strategies for search agents aimed at creating high-quality candidate solutions from three perspectives. In particular, ranking paired mutually beneficial learning is specifically designed to expand the scope of exploration to generate as many high-quality candidate solutions as possible for future generations. The Gaussian random walk learning is specifically designed to achieve better optimization in the transitional phase by adjusting the step-size control parameters, successfully completing the transition from exploration to local exploitation phase. Additionally, modified somersault foraging strategy is introduced to accelerate local convergence and perform more extensive local exploitation. Secondly, we integrate reinforcement learning into MPA and use Q-learning mechanism to adaptively select location update strategies. Agents fully utilize the collected information to evaluate the next action of the agents, coordinate the exploration phase and exploitation phase, and enhance the global optimization ability. Finally, compared with 10 competitive algorithms, RLMPA achieves better comprehensive performance in global optimization ability, search efficiency and convergence speed on 41 test functions and 5 practical engineering problems. In the Friedman rank sum tests, RLMPA achieves a preferable overall ranking, and has certain ascendant preponderances in solving practical problems with stability, effectiveness and robustness.
Keyword:
Global optimization
Reinforcement learning
Q-learning
Marine predators algorithm
Practical engineering problems
期刊
C
IF:
4.1
论文数:
5.0K
被引数:
7.5K
机构
引用论文
Entry Guidance Command Generation for Hypersonic Glide Vehicles Under Threats and Multiple Constraints
IEEE ACCESS
IF3.6
Manta ray foraging and Gaussian mutation-based elephant herding optimization for global optimization
Improved Reptile Search Optimization Algorithm Using Chaotic Map and Simulated Annealing for Feature Selection in Medical Field
IEEE ACCESS
IF3.6
Detecting Spam Email With Machine Learning Optimized With Bio-Inspired Metaheuristic Algorithms使用生物启发的元启发式算法优化的机器学习来检测垃圾邮件
IEEE ACCESS
IF3.6
A Neighborhood Regression Optimization Algorithm for Computationally Expensive Optimization Problems

