Hedging your bets by learning reward correlations in the human brain.

Klaus Wunderlich1, Mkael Symmonds, Peter Bossaerts

  • 1Wellcome Trust Center for Neuroimaging, University College London, London WC1N 3BG, UK. k.wunderlich@ucl.ac.uk

Neuron
|September 29, 2011
PubMed
Summary

Humans can learn complex reward correlations, using this information to optimize choices and minimize outcome variance. This ability is neurally represented in the brain, aiding adaptive sampling strategies.

Related Concept Videos

Motivational Bias01:25

Motivational Bias

Cognitive bias results from limitations in thinking and information processing, leading to systematic errors in judgment. Conversely, motivational bias stems from personal desires or emotions, causing distortions in perception to align with self-interest. Motivational bias influences how individuals perceive and attribute causes to events, often shaped by personal needs, goals, and self-esteem preservation. This bias can distort judgment, leading to inaccurate assessments of success, failure,...
Hindsight Biases01:12

Hindsight Biases

Hindsight bias leads you to believe that the event you just experienced was predictable, even though it really wasn’t. In other words, you knew all along that things would turn out the way they did. Can you relate this to the phrase "Hindsight is 20/20" now?
Cognitive Learning01:21

Cognitive Learning

Cognitive learning is based on purposive behavior, incidental learning, and insight learning.
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
Timing and Consequences on Behavior01:08

Timing and Consequences on Behavior

In operant conditioning, the timing of reinforcement is crucial. For animals like rats and cats, immediate reinforcement (within a few seconds) is much more effective than delayed reinforcement. For example, a food reward for a rat needs to follow within 30 seconds of pressing a bar to be effective. 
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant factor...
Cause and Effect01:53

Cause and Effect

While variables are sometimes correlated because one does cause the other, it could also be that some other factor, a confounding variable, is actually causing the systematic movement in our variables of interest. For instance, as sales in ice cream increase, so does the overall rate of crime. Is it possible that indulging in your favorite flavor of ice cream could send you on a crime spree? Or, after committing crime do you think you might decide to treat yourself to a cone?
Associative Learning01:27

Associative Learning

Associative learning is a fundamental concept in behavioral psychology, wherein a connection is established between two stimuli or events, leading to a learned response. This process is critical in understanding how behaviors are acquired and modified. Conditioning, the mechanism through which associations are formed, can be divided into two main types: classical conditioning and operant conditioning, each elucidating different aspects of associative learning.
Classical conditioning, also known...