Related Experiment Video
Updated: May 8, 2026

Operant Learning of Drosophila at the Torque Meter
Published on: June 16, 2008
Reinforcement learning with thermal fluctuations at the nanoscale
Francesco Boccardo1,2, Olivier Pierre-Louis1
1<a href="https://ror.org/0323bey33">Institut Lumière Matière</a>, UMR5306, Université Lyon 1 - CNRS, Villeurbanne, France.
Abstract:
Reinforcement Learning offers a framework to learn to choose actions in order to control a system. However, at small scales Brownian fluctuations limit the control of nanomachine actuation or nanonavigation and of the molecular machinery of life. We analyze this regime using the general framework of Markov decision processes. We show that at the nanoscale, while optimal control actions should bring an improvement proportional to the small ratio of the applied force times a length scale over the temperature, the learned improvement is smaller and proportional to the square of this small ratio. Consequently, the efficiency of learning, which compares the learning improvement to the theoretical optimal improvement, drops to zero. Nevertheless, these limitations can be circumvented by using actions learned at a lower temperature. These results are illustrated with simulations of the control of the shape of small particle clusters.
Related Concept Videos
Temperature and Thermal Equilibrium
The concept of temperature has evolved from the common concepts of hot and cold. The scientific definition of temperature explains more than just our sense of hot and cold. Temperature is operationally defined as the quantity measured with a thermometer. Furthermore, temperature is...
Thermal expansion and Thermal stress: Problem Solving
To solve the problem, first, identify the known and unknown quantities. The initial length (L) of the bridge is 1275 m, the coefficient of linear expansion (α) for steel is 12 x 10-6/°C, and the change in temperature (ΔT) is 55...

