相关实验视频
Updated: Sep 11, 2025

06:57
Pavlovian Conditioned Approach Training in Rats
Published on: February 4, 2016
11.0K
调性多巴胺和价值学习中的偏见通过一种生物启发的强化学习模型联系在一起
Sandra Romero Pinto1,2,3, Naoshige Uchida4
1Department of Molecular and Cellular Biology, Center for Brain Science, Harvard University, Cambridge, MA, USA. sr4265@columbia.edu.
Nature communications
|August 13, 2025
概括
多巴胺水平的变化会改变大脑如何从奖励和惩罚中学习,从而导致有偏见的预测. 这种涉及多巴胺受体的机制可能解释精神疾病症状.
科学领域:
- 神经科学是一个神经科学.
- 计算精神病学是一种计算精神病学.
- 强化学习是一种强化学习.
背景情况:
- 偏见的未来预测是精神疾病的一个关键特征.
- 了解价值学习的神经机制对于开发有效治疗方法至关重要.
研究的目的:
- 为了研究背后存在偏见价值学习的机制.
- 为了建模突触可塑性和基底质回路如何对预测偏差有所贡献.
主要方法:
- 使用了强化学习模型.
- 结合了近期关于突触可塑性的发现.
- 检查了基底质中的对手电路机制.
主要成果:
- 增强性多巴胺的变化将学习平衡从积极和消极的奖励预测错误转移.
- 多巴胺受体 (D1和D2) 剂量占用曲线和亲和关系解释了偏见的价值预测.
- 该模型成功地解释了小鼠和人类的偏见价值学习.
结论:
- 性多巴胺水平显著调节学习过程.
- 拟议的机制为了解基底腺功能提供了基础.
- 这项研究提供了关于精神疾病的神经生物学基础的见解.
相关概念视频
Cognitive Learning
519
Cognitive learning is based on purposive behavior, incidental learning, and insight learning.
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
519
Purposive Learning
207
E. C. Tolman emphasized the purposiveness of behavior — the idea that much of our behavior is goal-directed. For instance, employees who aim for a promotion work diligently to meet their targets. Tolman argued that when classical conditioning and operant conditioning occur, the organism acquires certain expectations. In classical conditioning, a child might fear a dog because they expect it to bite. In operant conditioning, a person might consistently work overtime because they expect a...
207
Observational Learning
312
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
312
Instinctive Drift
324
Instinctive drift refers to the tendency of animals to revert to their innate behaviors despite repeated reinforcement. Breland and Breland demonstrated this concept in an experiment with a raccoon. The raccoon was trained to pick up two coins and place them in a container in exchange for food. Initially, the raccoon learned to associate the coins with food, making them a conditioned stimulus or a substitute for food. However, over time, the raccoon became less willing to put the coins into the...
324
Reinforcement
341
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
341
Law of Effect
1.6K
B.F. Skinner, a prominent figure in behavioral psychology, introduced operant conditioning by emphasizing the role of consequences in shaping behavior. This theory builds upon the law of effect proposed by Edward Thorndike, which posits that behaviors followed by satisfying outcomes are likely to be repeated. In contrast, those followed by unsatisfying outcomes are less likely to recur.
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
1.6K

