短暂的奖励任务:为什么子很难学习它?
Daniel Peng1, Zohaib Iqbal1, Thomas R Zentall1
1Department of Psychology, University of Kentucky.
概括
子在短暂的奖励任务中扎,直到奖励概率降低. 将奖励概率降低到50%使子能够学习最佳选择策略.
科学领域:
- 动物认知 动物认知
- 行为神经科学 行为神经科学
- 比较心理学比较心理学
背景情况:
- 短暂奖励任务涉及在奖励的刺激A和B之间做出选择.
- ,和灵长类动物学会了这项任务,但子和老鼠却没有.
- 以前的研究表明,子的困难可能源于刺激结果的相似性.
研究的目的:
- 为了调查为什么子无法学习短暂的奖励任务.
- 为了测试这种假设,即刺激结果相似性阻碍了学习.
- 探索奖励大小和概率在学习中的作用.
主要方法:
- 实验1:在选择B之后修改刺激 (C) 来测试结果相似性假设 (组AC,BC,BB).
- 实验2:将选择A和B的奖励概率降低到子的50%.
- 行为观察和选择模式的分析,以评估学习.
主要成果:
- 实验1中的子未能在所有修改条件 (AC,BC,BB) 中学习最佳策略.
- 实验2中的子成功地学会了在奖励概率为50%时做出最佳选择.
- 一个和两个奖励之间的感知价值的差异可能比0.5和一个奖励之间的差异影响力更小.
结论:
- 刺激结果的相似性似乎不是子在标准的短暂奖励任务中遇到困难的主要原因.
- 奖励概率显著影响子学习最佳策略的能力.
- 奖励的相对值,特别是分数奖励和某些奖励之间的差异,是学习的关键因素.
相关概念视频
Timing and Consequences on Behavior
157
In operant conditioning, the timing of reinforcement is crucial. For animals like rats and cats, immediate reinforcement (within a few seconds) is much more effective than delayed reinforcement. For example, a food reward for a rat needs to follow within 30 seconds of pressing a bar to be effective.
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
157
Instinctive Drift
330
Instinctive drift refers to the tendency of animals to revert to their innate behaviors despite repeated reinforcement. Breland and Breland demonstrated this concept in an experiment with a raccoon. The raccoon was trained to pick up two coins and place them in a container in exchange for food. Initially, the raccoon learned to associate the coins with food, making them a conditioned stimulus or a substitute for food. However, over time, the raccoon became less willing to put the coins into the...
330


