Related Experiment Video
Updated: Feb 11, 2026

Determining Ultrasonic Vocalization Preferences in Mice using a Two-choice Playback Test
Published on: September 3, 2015
Place preference and vocal learning rely on distinct reinforcers in songbirds
Don Murdoch1, Ruidong Chen1, Jesse H Goldberg2
1Department of Neurobiology and Behavior, Cornell University, Ithaca, NY, 14853, USA.
Abstract:
In reinforcement learning (RL) agents are typically tasked with maximizing a single objective function such as reward. But it remains poorly understood how agents might pursue distinct objectives at once. In machines, multiobjective RL can be achieved by dividing a single agent into multiple sub-agents, each of which is shaped by agent-specific reinforcement, but it remains unknown if animals adopt this strategy. Here we use songbirds to test if navigation and singing, two behaviors with distinct objectives, can be differentially reinforced. We demonstrate that strobe flashes aversively condition place preference but not song syllables. Brief noise bursts aversively condition song syllables but positively reinforce place preference. Thus distinct behavior-generating systems, or agencies, within a single animal can be shaped by correspondingly distinct reinforcement signals. Our findings suggest that spatially segregated vocal circuits can solve a credit assignment problem associated with multiobjective learning.
Related Concept Videos
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Corrosion of Reinforcement
However, over time and under certain conditions like carbonation, chloride ingress, and cracking this protective state can be compromised. Steel has areas with...
Reinforcement Schedules
Once a behavior is learned,...
Reinforcements in Concrete
Fiber Reinforced Concrete
Reinforced Brick Masonry
To fortify brick walls...

