Related Experiment Video
Updated: Apr 5, 2026

Pavlovian Conditioned Approach Training in Rats
Published on: February 4, 2016
Contextual modulation of value signals in reward and punishment learning
Stefano Palminteri1,2, Mehdi Khamassi3,4, Mateus Joffily4,5
1Institute of Cognitive Neuroscience (ICN), University College London (UCL), London WC1N 3AR, UK.
Abstract:
Compared with reward seeking, punishment avoidance learning is less clearly understood at both the computational and neurobiological levels. Here we demonstrate, using computational modelling and fMRI in humans, that learning option values in a relative--context-dependent--scale offers a simple computational solution for avoidance learning. The context (or state) value sets the reference point to which an outcome should be compared before updating the option value. Consequently, in contexts with an overall negative expected value, successful punishment avoidance acquires a positive value, thus reinforcing the response. As revealed by post-learning assessment of options values, contextual influences are enhanced when subjects are informed about the result of the forgone alternative (counterfactual information). This is mirrored at the neural level by a shift in negative outcome encoding from the anterior insula to the ventral striatum, suggesting that value contextualization also limits the need to mobilize an opponent punishment learning system.
Related Concept Videos
Operant Conditioning
Reinforcement in operant conditioning can be positive or negative, both of which serve to increase the likelihood of a behavior. Positive...
Timing and Consequences on Behavior
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
Punishment
Punishment can be positive or negative. Positive punishment involves adding an undesirable stimulus, such as scolding, to decrease a behavior. Negative punishment involves removing a desirable stimulus, such as taking away a favorite toy, to decrease behavior....
Primary and Secondary Reinforcers
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
Behavior Modification
A real-world application of operant conditioning principles is applied...
Generalization, Discrimination, and Extinction
Generalization occurs when a behavior reinforced in one context is performed in similar situations. For instance, a student who studies diligently for calculus and receives excellent grades might apply the same study habits to psychology and history, expecting similar results. Generalization shows how learning in one setting can influence behavior in...

