Related Experiment Video
Updated: Jan 8, 2026

Operant Protocols for Assessing the Cost-benefit Analysis During Reinforced Decision Making by Rodents
Published on: September 10, 2018
Human reinforcement learning processes and biases: computational characterization and possible applications to
1Laboratoire de Neurosciences Cognitives et Computationnelles, Institut National de la Santé et de la Recherche Médicale, Paris, France.
None:
The reinforcement learning framework provides a computational and behavioral foundation for understanding how agents learn to maximize rewards and minimize punishments through interaction with their environment. This framework has been widely applied across disciplines, including artificial intelligence, animal psychology, and economics. Over the last decade, a growing body of research has shown that human reinforcement learning often deviates from normative standards, exhibiting systematic biases. The first aim of this paper is to propose a conceptual framework and a taxonomy for evaluating computational biases within reinforcement learning. We specifically propose a distinction between praxic biases, characterized by a mismatch between internal representations and selected actions, and epistemic biases, characterized by a mismatch between past experiences and internal representations. Building on this foundation, we characterize and discuss two primary types of epistemic biases: relative valuation and biased update. We describe their behavioral signatures and discuss their potential adaptive roles. Finally, we eleborate on how these findings may shape future developments in both theoretical and applied domains. Notably, despite being widely used in clinical and educational settings, reinforcement-based interventions have been comparatively neglected in the domains of behavioral public policy and decision-making improvement, particularly when compared to more popular approaches such as nudges and boosts. In this review, we offer an explanation for this comparative neglect that we believe rooted in common historical and epistemological misconceptions, and advocate for a greater integration of reinforcement learning into the design of behavioral public policy.
Related Concept Videos
Behavior Modification
A real-world application of operant conditioning principles is applied...
Cognitive Learning
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
Behaviorism
The core premise of behaviorism is its focus on observable behavior rather than internal thoughts or feelings. This approach argues that true scientific...
Behavioral Genetics and Its Designs
The primary methodologies used in behavior genetics include family studies, twin studies, and adoption studies, each providing unique...
Law of Effect
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
Operant Conditioning
Reinforcement in operant conditioning can be positive or negative, both of which serve to increase the likelihood of a behavior. Positive...

