Related Experiment Video
Updated: Aug 6, 2025

New Variations for Strategy Set-shifting in the Rat
Published on: January 23, 2017
On the normative advantages of dopamine and striatal opponency for learning and choice
Alana Jaskir1, Michael J Frank1
1Department of Cognitive, Linguistic and Psychological Sciences, Carney Institute for Brain Science, Brown University, Providence, United States.
Abstract:
The basal ganglia (BG) contribute to reinforcement learning (RL) and decision-making, but unlike artificial RL agents, it relies on complex circuitry and dynamic dopamine modulation of opponent striatal pathways to do so. We develop the OpAL* model to assess the normative advantages of this circuitry. In OpAL*, learning induces opponent pathways to differentially emphasize the history of positive or negative outcomes for each action. Dynamic DA modulation then amplifies the pathway most tuned for the task environment. This efficient coding mechanism avoids a vexing explore-exploit tradeoff that plagues traditional RL models in sparse reward environments. OpAL* exhibits robust advantages over alternative models, particularly in environments with sparse reward and large action spaces. These advantages depend on opponent and nonlinear Hebbian plasticity mechanisms previously thought to be pathological. Finally, OpAL* captures risky choice patterns arising from DA and environmental manipulations across species, suggesting that they result from a normative biological mechanism.
More Related Videos
07:05Operant Protocols for Assessing the Cost-benefit Analysis During Reinforced Decision Making by Rodents
Published on: September 10, 2018
08:07Simultaneous Detection of c-Fos Activation from Mesolimbic and Mesocortical Dopamine Reward Sites Following Naive Sugar and Fat Ingestion in Rats
Published on: August 24, 2016
Related Concept Videos
Operant Conditioning
Reinforcement in operant conditioning can be positive or negative, both of which serve to increase the likelihood of a behavior. Positive...
Timing and Consequences on Behavior
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
Law of Effect
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
Associative Learning
Classical conditioning, also known...
Purposive Learning
Operant Conditioning Intervention
In operant conditioning, behaviors that are...