Related Experiment Video
Updated: Sep 15, 2025

12:09
Studying Food Reward and Motivation in Humans
Published on: March 19, 2014
23.7K
Dissociable Effects of Curiosity and Hedonic Valence on Reinforcement Learning
Biorxiv : the Preprint Server for Biology
|July 16, 2025
Summary
Monkeys explored uncertain options more when losses were possible, driven by optimistic beliefs about novelty. Valence influenced learning from losses and choice-making, revealing distinct curiosity and motivation effects.
Area of Science:
- Neuroscience
- Behavioral Economics
- Cognitive Science
Background:
- Curiosity and exploration are crucial for learning and decision-making under uncertainty.
- The impact of outcome valence (positive vs. negative) on exploration strategies across species is not fully understood.
- Previous human studies often rely on explicit instructions, limiting insight into experience-driven exploration.
Purpose of the Study:
- To investigate how hedonic valence influences novelty seeking, exploration, and reinforcement learning in rhesus macaques.
- To determine if valence-dependent exploration arises from curiosity or prior beliefs.
- To examine the roles of approach and avoidance motivation in learning and decision-making.
Main Methods:
- Rhesus macaques were studied using visual tokens as secondary reinforcers.
- Exploration of novel, uncertain options was compared between conditions predicting gains and losses.
- Reinforcement learning, choice balks, and directed vs. random exploration were analyzed.
Main Results:
- Monkeys showed increased novelty seeking when outcomes involved potential losses compared to gains.
- Heightened novelty seeking was attributed to optimistic prior beliefs about novelty, not a direct valence effect on curiosity.
- Monkeys learned faster from losses than gains, indicating loss aversion.
- Choice balks served as strategic responses to approach-avoidance conflicts and uncertainty, modulating exploration.
Conclusions:
- Hedonic valence does not categorically alter curiosity but influences exploration through prior beliefs and motivational states.
- Approach and avoidance motivations differentially impact reinforcement learning and decision-making strategies.
- The findings reveal dissociable effects of curiosity and valence on explore-exploit tradeoffs and adaptive behavior.
Related Concept Videos
Cognitive Learning
539
Cognitive learning is based on purposive behavior, incidental learning, and insight learning.
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
539
Instinctive Drift
329
Instinctive drift refers to the tendency of animals to revert to their innate behaviors despite repeated reinforcement. Breland and Breland demonstrated this concept in an experiment with a raccoon. The raccoon was trained to pick up two coins and place them in a container in exchange for food. Initially, the raccoon learned to associate the coins with food, making them a conditioned stimulus or a substitute for food. However, over time, the raccoon became less willing to put the coins into the...
329
Primary and Secondary Reinforcers
414
In psychology, reinforcement is a key concept in behavior modification. B.F. Skinner demonstrated this with his experiments involving rats in what is known as a Skinner box. The rats learned to press a lever to receive food, a primary reinforcer that fulfilled their innate need for nourishment.
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
414
Generalization, Discrimination, and Extinction
813
Generalization, discrimination, and extinction are key concepts in operant conditioning that influence how behaviors are learned and maintained.
Generalization occurs when a behavior reinforced in one context is performed in similar situations. For instance, a student who studies diligently for calculus and receives excellent grades might apply the same study habits to psychology and history, expecting similar results. Generalization shows how learning in one setting can influence behavior in...
Generalization occurs when a behavior reinforced in one context is performed in similar situations. For instance, a student who studies diligently for calculus and receives excellent grades might apply the same study habits to psychology and history, expecting similar results. Generalization shows how learning in one setting can influence behavior in...
813
Behaviorism
2.7K
The field of behaviorism was pioneered by figures such as Ivan Pavlov, John B. Watson, and B.F. Skinner fundamentally shifted the focus of psychology to the observable and controllable aspects of human and animal behavior. This shift marked a critical evolution in the discipline, emphasizing scientific rigor and experimental methodology.
The core premise of behaviorism is its focus on observable behavior rather than internal thoughts or feelings. This approach argues that true scientific...
The core premise of behaviorism is its focus on observable behavior rather than internal thoughts or feelings. This approach argues that true scientific...
2.7K
Purposive Learning
208
E. C. Tolman emphasized the purposiveness of behavior — the idea that much of our behavior is goal-directed. For instance, employees who aim for a promotion work diligently to meet their targets. Tolman argued that when classical conditioning and operant conditioning occur, the organism acquires certain expectations. In classical conditioning, a child might fear a dog because they expect it to bite. In operant conditioning, a person might consistently work overtime because they expect a...
208

