关于关于延迟奖励的提前信息的价值
Alejandro Macías1,2, Armando Machado3, Marco Vasconcelos3
1William James Center for Research, University of Aveiro, Aveiro, Portugal. josea.maciasaa@gmail.com.
Animal cognition
|March 1, 2024
概括
子表现出对选择的偏好,这些选择标志着延迟奖励,而不是那些不这样做的人. 这种对信息信号的偏好随着长时间延迟和短时间延迟之间的比率增加而加剧.
科学领域:
- 动物行为 动物行为
- 认知科学 认知科学
- 行为经济学是一种行为经济学.
背景情况:
- 动物往往更喜欢具有可预测线索的结果,而不是不确定的结果.
- 在延迟奖励下理解决策在行为研究中至关重要.
研究的目的:
- 为了调查子是否更喜欢信号延迟来奖励而不是无信号延迟.
- 确定延迟时间的比例如何影响这种偏好.
主要方法:
- 子在两个提供不同延迟奖励的替代方案之间做出选择.
- 一个替代方案提供了可靠的延迟信号 (信息化),而另一个则没有 (非信息化).
- 短暂延误与长时间延误的比例有系统地变化.
主要成果:
- 子始终更喜欢信息化选项而不是非信息化选项.
- 随着长时间延迟与短时间延迟的比率的增加,对信号延迟的偏好得到了加强.
- 一个修改后的 Δ-Σ 假设准确地模拟了观察到的行为.
结论:
- 子表现出对信号延迟的偏好,表明了预测信息的价值.
- 这种偏好是由延迟比率的大小调节的.
- 调查结果表明,预警延误在决策中提供了工具性优势.
相关概念视频
Timing and Consequences on Behavior
90
In operant conditioning, the timing of reinforcement is crucial. For animals like rats and cats, immediate reinforcement (within a few seconds) is much more effective than delayed reinforcement. For example, a food reward for a rat needs to follow within 30 seconds of pressing a bar to be effective.
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
90
Prediction Intervals
2.3K
The interval estimate of any variable is known as the prediction interval. It helps decide if a point estimate is dependable.
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y.
2.3K
Reinforcement Schedules
145
Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
Once a behavior is learned,...
145
Primary and Secondary Reinforcers
253
In psychology, reinforcement is a key concept in behavior modification. B.F. Skinner demonstrated this with his experiments involving rats in what is known as a Skinner box. The rats learned to press a lever to receive food, a primary reinforcer that fulfilled their innate need for nourishment.
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
253
Hindsight Biases
3.4K
Hindsight bias leads you to believe that the event you just experienced was predictable, even though it really wasn’t. In other words, you knew all along that things would turn out the way they did. Can you relate this to the phrase "Hindsight is 20/20" now?
3.4K
Real-World Application of Classical Conditioning
556
Classical conditioning not only includes the initial pairing of stimuli but also extends to more complex forms, such as higher-order conditioning. Higher-order conditioning involves creating associations beyond the primary conditioned stimulus, resulting in a chain of conditioned responses.
Higher-order, or second-order, conditioning occurs when a neutral stimulus becomes associated with an already established conditioned stimulus through repeated pairings. For instance, if a dog has been...
Higher-order, or second-order, conditioning occurs when a neutral stimulus becomes associated with an already established conditioned stimulus through repeated pairings. For instance, if a dog has been...
556


