Related Experiment Video
Updated: Jun 28, 2025

Novel Object Exploration as a Potential Assay for Higher Order Repetitive Behaviors in Mice
Published on: August 20, 2016
Dopamine encoding of novelty facilitates efficient uncertainty-driven exploration
Yuhao Wang1, Armin Lak2, Sanjay G Manohar3
1MRC Brain Network Dynamics Unit, University of Oxford, Oxford, United Kingdom.
Abstract:
When facing an unfamiliar environment, animals need to explore to gain new knowledge about which actions provide reward, but also put the newly acquired knowledge to use as quickly as possible. Optimal reinforcement learning strategies should therefore assess the uncertainties of these action-reward associations and utilise them to inform decision making. We propose a novel model whereby direct and indirect striatal pathways act together to estimate both the mean and variance of reward distributions, and mesolimbic dopaminergic neurons provide transient novelty signals, facilitating effective uncertainty-driven exploration. We utilised electrophysiological recording data to verify our model of the basal ganglia, and we fitted exploration strategies derived from the neural model to data from behavioural experiments. We also compared the performance of directed exploration strategies inspired by our basal ganglia model with other exploration algorithms including classic variants of upper confidence bound (UCB) strategy in simulation. The exploration strategies inspired by the basal ganglia model can achieve overall superior performance in simulation, and we found qualitatively similar results in fitting model to behavioural data compared with the fitting of more idealised normative models with less implementation level detail. Overall, our results suggest that transient dopamine levels in the basal ganglia that encode novelty could contribute to an uncertainty representation which efficiently drives exploration in reinforcement learning.
Related Concept Videos
Uncertainty: Overview
The Availability Heuristic
Randomized Experiments
Simple randomization
Simple...
The Anchoring-and-Adjustment Heuristic
Generalization, Discrimination, and Extinction
Generalization occurs when a behavior reinforced in one context is performed in similar situations. For instance, a student who studies diligently for calculus and receives excellent grades might apply the same study habits to psychology and history, expecting similar results. Generalization shows how learning in one setting can influence behavior in...

