Related Experiment Video
Updated: Mar 7, 2026

A Flexible Platform for Monitoring Cerebellum-Dependent Sensory Associative Learning
Published on: January 19, 2022
Potential for Reinforcement Learning in the Cerebellum
Richard W Prager1, Richard Apps2
1University of Cambridge, Cambridge CB2 1PZ, UK rwp12@cam.ac.uk.
None:
This article explores how simple reinforcement learning algorithms might be implemented by the anatomy of the cerebellum. In doing this, we highlight which anatomical and physiological details are most important for assessing algorithmic fit, and we discuss which algorithm components are easiest to accommodate in a neural system. We describe hypothetical cerebellar implementations of four reinforcement learning algorithms and discuss the anatomical plausibility of the various components required. We show how one of the algorithms can learn to generate short sequences of actions without continuous information on the resulting changes to the environment. We finish with simulations that illustrate the way that the algorithms learn to solve the problem of balancing an inverted pendulum, commonly known as the cart-pole problem. We highlight two physiological features: reward signals and combining information across time, that indicate that some sort of reinforcement learning adaptation may be taking place. We also describe why the commonly used algorithmic feature, an eligibility trace, presents particular problems to implement in known neural anatomy.
More Related Videos
Related Concept Videos
Role of Cerebellum and Prefrontal Cortex in Memory
Observational Learning
Cerebellum: Anatomical Regions
Cerebellar Structure
Externally, the cerebellum features a highly convoluted surface with numerous folia (narrow ridges) separated by shallow sulci (grooves). The cerebellum is divided into two hemispheres by a thin median structure known as the vermis. The...
Major Somatic Sensory Pathways
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Diencephalon: Thalamus and Information Relay

