Related Experiment Video
Updated: Dec 13, 2025

Age-dependent Dynamics of Locomotion in Caenorhabditis elegans: A Lyapunov Exponent Analysis
Published on: September 23, 2025
Evolutionary reinforcement learning of dynamical large deviations
Stephen Whitelam1, Daniel Jacobson2, Isaac Tamblyn3
1Molecular Foundry, Lawrence Berkeley National Laboratory, 1 Cyclotron Road, Berkeley, California 94720, USA.
This study introduces evolutionary reinforcement learning to calculate dynamical large deviations. This method uses agents to model stochastic processes, enabling the computation of rate functions for complex physics problems.
Area of Science:
- Computational Physics
- Machine Learning
- Statistical Mechanics
Background:
- Dynamical large deviations are crucial for understanding rare events in stochastic systems.
- Calculating these deviations often involves computationally intensive methods.
- Existing frameworks may not fully capture the complexities of path-extensive quantities.
Purpose of the Study:
- To develop a novel method for bounding and calculating the likelihood of dynamical large deviations.
- To leverage evolutionary reinforcement learning for analyzing stochastic models.
- To bridge the gap between physics problems and machine learning frameworks.
Main Methods:
- An agent, representing a stochastic model, propagates continuous-time Monte Carlo trajectories.
- Rewards are assigned based on the values of path-extensive quantities.
- Evolutionary algorithms optimize agents to improve the calculation of large-deviation rate functions.
- For large state spaces, neural networks parameterize the model's rates.
Main Results:
- Demonstrated the feasibility of using evolutionary reinforcement learning to bound and calculate dynamical large deviations.
- Showcased the method's applicability to models with varying state space sizes.
- Successfully linked path-extensive physics problems to a machine learning framework.
Conclusions:
- Evolutionary reinforcement learning offers a powerful new approach for tackling complex problems in statistical mechanics and physics.
- This framework facilitates the computation of large-deviation rate functions, previously a significant challenge.
- The study highlights the potential of integrating advanced machine learning techniques into physical modeling.
Related Concept Videos
Entropy Change in Reversible Processes
The statement can be further generalized to prove that entropy is a state function. Take a cyclic process between any two points on a p-V diagram.
Observational Learning
Avoidance Learning and Learned Helplessness
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
Generalization, Discrimination, and Extinction
Generalization occurs when a behavior reinforced in one context is performed in similar situations. For instance, a student who studies diligently for calculus and receives excellent grades might apply the same study habits to psychology and history, expecting similar results. Generalization shows how learning in one setting can influence behavior in...
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Dynamic Equilibrium

