Jove
Visualize
联系我们
JoVE
x logofacebook logolinkedin logoyoutube logo
关于 JoVE
概览领导团队博客JoVE 帮助中心
作者
出版流程编辑委员会范围与政策同行评审常见问题投稿
图书馆员
用户评价订阅访问资源图书馆顾问委员会常见问题
研究
JoVE JournalMethods CollectionsJoVE Encyclopedia of Experiments存档
教育
JoVE CoreJoVE BusinessJoVE Science EducationJoVE Lab Manual教师资源中心教师网站
使用条款与条件
隐私政策
政策

相关概念视频

Incentive Theory: Pull Theory of Motivation01:18

Incentive Theory: Pull Theory of Motivation

485
Incentive theory, or the "pull theory" of motivation, suggests that external rewards primarily drive behavior. Individuals are motivated to engage in activities when they anticipate a desirable outcome. This is why people often work hard for promotions or study intensively to achieve high grades. These incentives can be tangible, physical rewards such as money or promotions, or intangible, non-physical rewards like praise and social recognition.
The theory differentiates between...
485
Primary and Secondary Reinforcers01:23

Primary and Secondary Reinforcers

304
In psychology, reinforcement is a key concept in behavior modification. B.F. Skinner demonstrated this with his experiments involving rats in what is known as a Skinner box. The rats learned to press a lever to receive food, a primary reinforcer that fulfilled their innate need for nourishment.
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
304
Timing and Consequences on Behavior01:08

Timing and Consequences on Behavior

122
In operant conditioning, the timing of reinforcement is crucial. For animals like rats and cats, immediate reinforcement (within a few seconds) is much more effective than delayed reinforcement. For example, a food reward for a rat needs to follow within 30 seconds of pressing a bar to be effective. 
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
122
Reinforcement01:23

Reinforcement

280
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
280
Reinforcement Schedules01:24

Reinforcement Schedules

205
Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
205
Instinctive Drift01:05

Instinctive Drift

254
Instinctive drift refers to the tendency of animals to revert to their innate behaviors despite repeated reinforcement. Breland and Breland demonstrated this concept in an experiment with a raccoon. The raccoon was trained to pick up two coins and place them in a container in exchange for food. Initially, the raccoon learned to associate the coins with food, making them a conditioned stimulus or a substitute for food. However, over time, the raccoon became less willing to put the coins into the...
254

您也可能阅读

相关文章

通过共同作者、期刊和引用图与本文相关的文章。

排序
Same author

Tuning tasks to how we learn.

Nature human behaviour·2026
Same author

Latent subdimensions of anxiety and depression differentially influence exertion of effort in pursuit of reward versus avoidance of threat.

Translational psychiatry·2026
Same author

A habit and working memory model as an alternative account of human reward-based learning.

Nature human behaviour·2025
Same author

Striatal dopamine can enhance both fast working memory, and slow reinforcement learning, while reducing implicit effort cost sensitivity.

Nature communications·2025
Same author

Episodic memory contributions to working memory-supported reinforcement learning.

Journal of experimental psychology. Learning, memory, and cognition·2025
Same author

Naturally disengaging control to reveal habits.

Research square·2025

相关实验视频

Updated: Jul 23, 2025

Pavlovian Conditioned Approach Training in Rats
06:57

Pavlovian Conditioned Approach Training in Rats

Published on: February 4, 2016

11.0K

内在奖励解释了强化学习中的情境敏感估值.

Gaia Molinaro1, Anne G E Collins1,2

  • 1Department of Psychology, University of California, Berkeley, Berkeley, California, United States of America.

PLoS biology
|July 17, 2023
PubMed
概括

人类的选择受到环境的影响. 我们的新模型显示了内部目标,而不仅仅是外部奖励,在强化学习 (RL) 中塑造主观价值,改进了情境敏感估值的预测.

科学领域:

  • 认知科学 认知科学
  • 神经科学是一个神经科学.
  • 计算心理学 计算心理学

背景情况:

  • 人类的决策是敏感的环境,在这个环境中,一个选项的感知价值的变化基于可用的替代品.
  • 这种情境敏感的估值在强化学习 (RL) 中被观察到,强化学习是通过试错学习的过程.
  • 范围适应一直是主要的解释,表明选项是基于经验丰富的价值范围重新缩放的.

研究的目的:

  • 提出和测试一个替代的机制,以RL的情境敏感估值.
  • 调查内部定义的目标在塑造主观价值的作用.
  • 引入和验证一种新的内在增强的RL模型.

主要方法:

  • 开发了一种内在增强的RL模型,将外部奖励与内部目标实现信号相结合.
  • 在七项研究中测试了该模型,包括现有数据集和一项新的预注册实验.
  • 在解释情境敏感估值时,将模型的性能与范围调整进行了比较.

主要成果:

  • 本质上增强的RL模型解释了情境敏感的估值和范围调整一样有效,或者比范围调整更好.
  • 研究结果表明,内部,目标依赖的奖励在人类RL中起着重要作用.
  • 该模型在依赖上下文的场景中预测人类行为的准确性有所提高.

更多相关视频

Studying Food Reward and Motivation in Humans
12:09

Studying Food Reward and Motivation in Humans

Published on: March 19, 2014

23.5K
A Conflict Model of Reward-seeking Behavior in Male Rats
06:11

A Conflict Model of Reward-seeking Behavior in Male Rats

Published on: February 20, 2019

7.5K

相关实验视频

Last Updated: Jul 23, 2025

Pavlovian Conditioned Approach Training in Rats
06:57

Pavlovian Conditioned Approach Training in Rats

Published on: February 4, 2016

11.0K
Studying Food Reward and Motivation in Humans
12:09

Studying Food Reward and Motivation in Humans

Published on: March 19, 2014

23.5K
A Conflict Model of Reward-seeking Behavior in Male Rats
06:11

A Conflict Model of Reward-seeking Behavior in Male Rats

Published on: February 20, 2019

7.5K

结论:

  • 内部生成的目标对于理解决策中的主观价值至关重要.
  • 整合内在奖励信号增强了人类行为的标准RL模型.
  • 这项研究重新定义了在强化学习和决策理论中对奖励处理的理解.