Related Experiment Video
Updated: Oct 17, 2025

Studying Cell Rolling Trajectories on Asymmetric Receptor Patterns
Published on: February 13, 2011
Reinforcement learning of rare diffusive dynamics
Avishek Das1, Dominic C Rose2, Juan P Garrahan2
1Department of Chemistry, University of California, Berkeley, California 94609, USA.
Abstract:
We present a method to probe rare molecular dynamics trajectories directly using reinforcement learning. We consider trajectories that are conditioned to transition between regions of configuration space in finite time, such as those relevant in the study of reactive events, and trajectories exhibiting rare fluctuations of time-integrated quantities in the long time limit, such as those relevant in the calculation of large deviation functions. In both cases, reinforcement learning techniques are used to optimize an added force that minimizes the Kullback-Leibler divergence between the conditioned trajectory ensemble and a driven one. Under the optimized added force, the system evolves the rare fluctuation as a typical one, affording a variational estimate of its likelihood in the original trajectory ensemble. Low variance gradients employing value functions are proposed to increase the convergence of the optimal force. The method we develop employing these gradients leads to efficient and accurate estimates of both the optimal force and the likelihood of the rare event for a variety of model systems.
Related Concept Videos
Passive Diffusion: Overview and Kinetics
When administered orally, drugs establish a substantial concentration gradient between the gastrointestinal (GI) lumen and the bloodstream, expediting...
Observational Learning
Instinctive Drift
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Laminar Flow: Problem Solving
Rolling Resistance: Problem Solving

