Related Experiment Video
Updated: Jul 16, 2026

Pavlovian Conditioned Approach Training in Rats
Published on: February 4, 2016
PVLV: the primary value and learned value Pavlovian learning algorithm.
Randall C O'Reilly1, Michael J Frank, Thomas E Hazy
1Department of Psychology, University of Colorado, Boulder, CO 80309, USA. oreilly@psych.colorado.edu
The primary value learned value (PVLV) model offers a biologically plausible alternative to TD algorithms for understanding dopamine neuron activity in reward learning. PVLV better explains neural data and predicts distinct anatomical pathways for reward conditioning.
Area of Science:
- Neuroscience
- Computational Neuroscience
- Behavioral Neuroscience
Background:
- Dopamine (DA) neurons play a crucial role in reward prediction and learning.
- Existing models like temporal-differences (TD) algorithm have limitations in explaining biological mechanisms.
- A more biologically grounded model is needed to understand DA neuron firing properties.
Purpose of the Study:
- To introduce and validate the primary value learned value (PVLV) model.
- To present PVLV as a biologically plausible alternative to the TD algorithm for DA neuron function.
- To explain reward-predictive firing properties of DA neurons.
Main Methods:
- Developed the PVLV model, separating primary value (PV) and learned value (LV) systems.
- PV system linked to Rescorla-Wagner/delta-rule and ventral striatum/nucleus accumbens.
- LV system linked to central nucleus of the amygdala.
Main Results:
- The PVLV model successfully accounts for key aspects of DA neuron firing data.
- PVLV predicts distinct neural pathways for primary and learned rewards, supported by existing data.
- The model explains the anatomical dissociation of first- and second-order conditioning, unlike TD models.
Conclusions:
- The PVLV model provides a biologically plausible framework for understanding reward learning.
- PVLV offers a more robust explanation of DA neuron function compared to TD algorithms.
- The model highlights the distinct roles of the PV and LV systems in reward processing.
More Related Videos
Related Concept Videos
Classical Conditioning
Ivan Pavlov observed that dogs salivated...
Behaviorism
The core premise of behaviorism is its focus on observable behavior rather than internal thoughts or feelings. This approach argues that true scientific...
Principles of Classical Conditioning
During the...
Associative Learning
Classical conditioning, also known...
Real-World Application of Classical Conditioning
Higher-order, or second-order, conditioning occurs when a neutral stimulus becomes associated with an already established conditioned stimulus through repeated pairings. For instance, if a dog has been...
Operant Conditioning
Reinforcement in operant conditioning can be positive or negative, both of which serve to increase the likelihood of a behavior. Positive...

