Related Experiment Video
Updated: Feb 21, 2026

Recording Single Neurons' Action Potentials from Freely Moving Pigeons Across Three Stages of Learning
Published on: June 2, 2014
Dynamic Encoding of Reward Prediction Error Signals in the Pigeon Ventral Tegmental Area during Reinforcement
Zhigang Shang1,2, Jiashuo Zhang1,2, Mengmeng Li1,2
1School of Electrical and Information Engineering, Zhengzhou University, Zhengzhou 450001, China.
Abstract:
Reward prediction errors (RPEs) guide learning by comparing expected and obtained outcomes. In mammals, ventral tegmental area (VTA) activity is closely linked to RPE-like signaling, yet how avian VTA dynamics evolve during reinforcement learning remains less well characterized. Here we recorded VTA spiking in pigeons (two females and one male) performing a cue-guided operant task in which a green cue (cue+) predicted reward contingent on a key peck, whereas a red cue (cue-) was unrewarded. Using a 16-channel microwire array, we analyzed pooled channel-level multiunit activity (MUA) aligned to task events. Across sessions, cue+ trials showed a learning-related redistribution of event-locked modulation: outcome-locked activity was more prominent early in training, while cue-locked modulation became stronger as performance stabilized, consistent with a temporal-difference-like shift of prediction-related signals. Cue- trials were sparse after early learning and showed limited cue-locked modulation in the available dataset. Together, these results provide initial evidence that pigeon VTA pooled MUA exhibits learning-related dynamics consistent with RPE-like processing and support cross-species comparisons of dopaminergic learning signals.
Related Concept Videos
Observational Learning
Operant Conditioning
Reinforcement in operant conditioning can be positive or negative, both of which serve to increase the likelihood of a behavior. Positive...
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Reinforcement Schedules
Once a behavior is learned,...
Avoidance Learning and Learned Helplessness
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
Timing and Consequences on Behavior
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...

