Dopamine ramps are a consequence of reward prediction errors

Samuel J Gershman1

  • 1Department of Brain and Cognitive Sciences, MIT, Cambridge, MA 02139, U.S.A. sjgershm@mit.edu.

Neural Computation
|December 11, 2013
PubMed
Summary

Temporal difference learning models explain dopamine's role in reward prediction error. New findings show these models can also predict gradual dopamine ramping towards a goal, reconciling theory with experimental data.

Related Concept Videos

Drug Abuse and Addiction: Pharmacological Phenomena01:15

Drug Abuse and Addiction: Pharmacological Phenomena

Drug dependence, abuse, and addiction are complex phenomena that can precipitate various abnormal states. Physical dependence refers to a state of pharmacological adaptation to a drug. This adaptation often results in tolerance—a reduced response to the drug after repeated administrations. When the drug use is abruptly stopped, withdrawal symptoms occur due to the body's need to readjust from the pharmacologically induced imbalance. However, tolerance and withdrawal symptoms do not...
1.6K
Drugs Affecting Neurotransmitter Synthesis01:29

Drugs Affecting Neurotransmitter Synthesis

Drugs affecting neurotransmitter synthesis can impact the adrenergic neuron and the synthesis of neurotransmitters. For example, α-methyltyrosine and carbidopa target specific enzymes involved in catecholamine synthesis. α-methyltyrosine inhibits the enzyme tyrosine hydroxylase, which converts tyrosine into dopamine. By blocking this enzyme, α-methyltyrosine reduces dopamine production and other catecholamines. Carbidopa, on the other hand, inhibits the enzyme dopa decarboxylase,...
2.5K
Timing and Consequences on Behavior01:08

Timing and Consequences on Behavior

In operant conditioning, the timing of reinforcement is crucial. For animals like rats and cats, immediate reinforcement (within a few seconds) is much more effective than delayed reinforcement. For example, a food reward for a rat needs to follow within 30 seconds of pressing a bar to be effective. 
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
927
Adrenergic Agonists: Indirect-Acting Agents01:25

Adrenergic Agonists: Indirect-Acting Agents

Indirect-acting adrenergic agonists potentiate the effects of endogenous catecholamines through different mechanisms without directly binding to adrenoceptors.
One mechanism involves depleting stored catecholamines by displacing them from synaptic vesicles. These agents, known as "displacers," are transported into vesicles at the expense of noradrenaline. Examples include amphetamine and tyramine, which lack a catechol moiety, resulting in prolonged action, improved oral...
2.8K
Instinctive Drift01:05

Instinctive Drift

Instinctive drift refers to the tendency of animals to revert to their innate behaviors despite repeated reinforcement. Breland and Breland demonstrated this concept in an experiment with a raccoon. The raccoon was trained to pick up two coins and place them in a container in exchange for food. Initially, the raccoon learned to associate the coins with food, making them a conditioned stimulus or a substitute for food. However, over time, the raccoon became less willing to put the coins into the...
1.5K
Reinforcement Schedules01:24

Reinforcement Schedules

Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
740