Multi-timescale reinforcement learning in the brain.

Paul Masset1,2, Pablo Tano3, HyungGoo R Kim1,2,4,5

  • 1Department of Molecular and Cellular Biology, Harvard University, USA.

Summary

This study reveals that reinforcement learning agents benefit from multiple timescales, not just one. Dopamine neurons in mice exhibit diverse temporal discounting, suggesting cell-specific properties crucial for adaptive behavior.