Related Experiment Video
Updated: Sep 13, 2025

Investigating Motor Skill Learning Processes with a Robotic Manipulandum
Published on: February 12, 2017
Decoding fairness: A reinforcement learning perspective.
Guozhong Zheng1, Jiqiang Zhang2, Xin Ou1
1Shaanxi Normal University, School of Physics and Information Technology, Xi'an 710061, People's Republic of China.
Human behavior in the ultimatum game (UG) shows a preference for fairness, challenging traditional economic models. This study demonstrates that internal motivations, not external factors, drive fair decision-making through reinforcement learning.
Area of Science:
- Behavioral Economics
- Computational Neuroscience
- Game Theory
Background:
- Behavioral experiments on the ultimatum game (UG) show humans prefer fairness, contradicting orthodox economic predictions.
- Existing explanations often rely on external (exogenous) factors within imitation learning frameworks.
Purpose of the Study:
- To investigate the emergence of fairness in the ultimatum game using a reinforcement learning (RL) paradigm.
- To determine if endogenous incentives alone can explain fairness without external factors.
Main Methods:
- Applied Q learning to the ultimatum game, assigning two Q tables per player for proposer and responder roles.
- Analyzed a two-player scenario, extending to latticed populations with various role assignments (random, fixed, rotating).
Main Results:
- Fairness emerged prominently in the UG when players valued both experience and future rewards.
- The probability of successful deals increased with higher offers, aligning with empirical observations.
- The system exhibited two phases, stabilizing into either fair or rational strategies, robust across different conditions.
Conclusions:
- Endogenous incentives within a reinforcement learning framework are sufficient to explain the emergence of fairness in the ultimatum game.
- Exogenous factors are not necessary to account for observed fair behavior in economic decision-making.
More Related Videos
08:24The Joint Effect of Social Comparison and Social Distance on Evaluation of Intertemporal Choice Outcomes in Event-related Potential Studies
Published on: August 25, 2023
07:05Operant Protocols for Assessing the Cost-benefit Analysis During Reinforced Decision Making by Rodents
Published on: September 10, 2018
Related Concept Videos
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Observational Learning
Law of Effect
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
Primary and Secondary Reinforcers
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
Generalization, Discrimination, and Extinction
Generalization occurs when a behavior reinforced in one context is performed in similar situations. For instance, a student who studies diligently for calculus and receives excellent grades might apply the same study habits to psychology and history, expecting similar results. Generalization shows how learning in one setting can influence behavior in...
Reinforcement Schedules
Once a behavior is learned,...