Related Experiment Video
Updated: Sep 11, 2025

06:57
Pavlovian Conditioned Approach Training in Rats
Published on: February 4, 2016
11.0K
Tonic dopamine and biases in value learning linked through a biologically inspired reinforcement learning model
Sandra Romero Pinto1,2,3, Naoshige Uchida4
1Department of Molecular and Cellular Biology, Center for Brain Science, Harvard University, Cambridge, MA, USA. sr4265@columbia.edu.
Nature Communications
|August 13, 2025
Summary
Variations in dopamine levels alter how the brain learns from rewards and punishments, leading to biased predictions. This mechanism, involving dopamine receptors, may explain psychiatric disorder symptoms.
Area of Science:
- Neuroscience
- Computational Psychiatry
- Reinforcement Learning
Background:
- Biased future predictions are a key feature of psychiatric disorders.
- Understanding the neural mechanisms of value learning is crucial for developing effective treatments.
Purpose of the Study:
- To investigate the mechanisms underlying biased value learning.
- To model how synaptic plasticity and basal ganglia circuits contribute to prediction biases.
Main Methods:
- Utilized reinforcement learning models.
- Incorporated recent findings on synaptic plasticity.
- Examined opponent circuit mechanisms in the basal ganglia.
Main Results:
- Variations in tonic dopamine shift the balance of learning from positive and negative reward prediction errors.
- Dopamine receptor (D1 and D2) dose-occupancy curves and affinities explain biased value predictions.
- The model successfully explains biased value learning in both mice and humans.
Conclusions:
- Tonic dopamine levels significantly modulate learning processes.
- The proposed mechanism provides a foundation for understanding basal ganglia function.
- This research offers insights into the neurobiological underpinnings of psychiatric disorders.
Related Concept Videos
Cognitive Learning
519
Cognitive learning is based on purposive behavior, incidental learning, and insight learning.
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
519
Purposive Learning
207
E. C. Tolman emphasized the purposiveness of behavior — the idea that much of our behavior is goal-directed. For instance, employees who aim for a promotion work diligently to meet their targets. Tolman argued that when classical conditioning and operant conditioning occur, the organism acquires certain expectations. In classical conditioning, a child might fear a dog because they expect it to bite. In operant conditioning, a person might consistently work overtime because they expect a...
207
Observational Learning
312
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
312
Instinctive Drift
324
Instinctive drift refers to the tendency of animals to revert to their innate behaviors despite repeated reinforcement. Breland and Breland demonstrated this concept in an experiment with a raccoon. The raccoon was trained to pick up two coins and place them in a container in exchange for food. Initially, the raccoon learned to associate the coins with food, making them a conditioned stimulus or a substitute for food. However, over time, the raccoon became less willing to put the coins into the...
324
Reinforcement
341
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
341
Law of Effect
1.6K
B.F. Skinner, a prominent figure in behavioral psychology, introduced operant conditioning by emphasizing the role of consequences in shaping behavior. This theory builds upon the law of effect proposed by Edward Thorndike, which posits that behaviors followed by satisfying outcomes are likely to be repeated. In contrast, those followed by unsatisfying outcomes are less likely to recur.
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
1.6K

