Dopaminergic action prediction errors serve as a value-free teaching signal

Francesca Greenstreet1, Hernando Martinez Vergara1,2, Yvonne Johansson1

  • 1Sainsbury Wellcome Centre for Neural Circuits and Behaviour, University College London, London, UK.

Nature
|May 14, 2025
PubMed
Summary

Mice learn through two dopamine signals: one for rewards and another for repeating actions. Action prediction errors reinforce repetitive learning in the striatum, working with reward signals for stable associations.

Related Concept Videos

Time-Domain Interpretation of PD Control01:07

Time-Domain Interpretation of PD Control

Proportional-Derivative (PD) control is a widely used control method in various engineering systems to enhance stability and performance. In a system with only proportional control, common issues include high maximum overshoot and oscillation, observed in both the error signal and its rate of change. This behavior can be divided into three distinct phases: initial overshoot, subsequent undershoot, and gradual stabilization.
Consider the example of control of motor torque. Initially, a positive...
74
Purposive Learning01:22

Purposive Learning

E. C. Tolman emphasized the purposiveness of behavior — the idea that much of our behavior is goal-directed. For instance, employees who aim for a promotion work diligently to meet their targets. Tolman argued that when classical conditioning and operant conditioning occur, the organism acquires certain expectations. In classical conditioning, a child might fear a dog because they expect it to bite. In operant conditioning, a person might consistently work overtime because they expect a...
93
Drugs Affecting Neurotransmitter Synthesis01:29

Drugs Affecting Neurotransmitter Synthesis

Drugs affecting neurotransmitter synthesis can impact the adrenergic neuron and the synthesis of neurotransmitters. For example, α-methyltyrosine and carbidopa target specific enzymes involved in catecholamine synthesis. α-methyltyrosine inhibits the enzyme tyrosine hydroxylase, which converts tyrosine into dopamine. By blocking this enzyme, α-methyltyrosine reduces dopamine production and other catecholamines. Carbidopa, on the other hand, inhibits the enzyme dopa decarboxylase,...
1.2K
Hindsight Biases01:12

Hindsight Biases

Hindsight bias leads you to believe that the event you just experienced was predictable, even though it really wasn’t. In other words, you knew all along that things would turn out the way they did. Can you relate this to the phrase "Hindsight is 20/20" now? 
3.4K
Regression Toward the Mean01:52

Regression Toward the Mean

Regression toward the mean (“RTM”) is a phenomenon in which extremely high or low values—for example, and individual’s blood pressure at a particular moment—appear closer to a group’s average upon remeasuring. Although this statistical peculiarity is the result of random error and chance, it has been problematic across various medical, scientific, financial and psychological applications. In particular, RTM, if not taken into account, can interfere when...
6.3K
Law of Effect01:06

Law of Effect

B.F. Skinner, a prominent figure in behavioral psychology, introduced operant conditioning by emphasizing the role of consequences in shaping behavior. This theory builds upon the law of effect proposed by Edward Thorndike, which posits that behaviors followed by satisfying outcomes are likely to be repeated. In contrast, those followed by unsatisfying outcomes are less likely to recur.
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
1.3K