Related Experiment Video
Updated: Jan 12, 2026

Author Spotlight: Advancing Protein Engineering – Harnessing Evolution Through PRANCE and Lab Automation
Published on: January 12, 2024
An improved differential evolution algorithm based on reinforcement learning and its application
Guangwei Yang1,2, Peng Sun1, Jieyong Zhang3
1Information and Navigation College, Air Force Engineering University, Xi'an 710077, China.
None:
As a typical swarm intelligence optimization method, the Differential Evolution (DE) algorithm exhibits excellent performance in solving high-dimensional complex problems; however, its parameter sensitivity and premature convergence issues still restrict its practical application effectiveness. Therefore, this paper proposes an improved Differential Evolution algorithm based on reinforcement learning, namely RLDE. First, it adopts the Halton sequence to realize the uniform initialization of the population space, which effectively improves the ergodicity of the initial solution set. Second, it establishes a dynamic parameter adjustment mechanism based on the policy gradient network, and realizes the online adaptive optimization of the scaling factor and crossover probability through the reinforcement learning framework. Furthermore, it classifies the population according to individual fitness values and implements a differentiated mutation strategy. To verify the effectiveness of the proposed algorithm, 26 standard test functions are used for optimization testing, and comparisons are conducted with multiple heuristic optimization algorithms in 10, 30, and 50 dimensions respectively. Experimental results demonstrate that the proposed algorithm significantly enhances the global optimization performance. Furthermore, by modeling and solving the Unmanned Aerial Vehicle (UAV) task assignment problem, the engineering practical value of the algorithm in real-world scenarios is verified from various indicators.
Related Concept Videos
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Reinforcement Schedules
Once a behavior is learned,...
Generalization, Discrimination, and Extinction
Generalization occurs when a behavior reinforced in one context is performed in similar situations. For instance, a student who studies diligently for calculus and receives excellent grades might apply the same study habits to psychology and history, expecting similar results. Generalization shows how learning in one setting can influence behavior in...
Observational Learning
Differential Leveling
Law of Effect
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
