ドーパミンを媒介する補強によって 自然な行動が学習されます
Jonathan Kasdin1, Alison Duffy2, Nathan Nadler1
1Department of Neuroscience, Zuckerman Mind Brain Behavior Institute, Columbia University, New York, NY, USA.
Nature
|March 13, 2025
まとめ
斑馬のドーパミンは 音声学習を誘導し 音声の精度が 強化学習に似ています これはドーパミンベースの補強学習が 複雑な自然な行動の 基礎にあることを示しています
科学分野:
- 神経科学
- 動物 の 行動
- 計算神経科学
背景:
- 言語のような 自然な運動能力は 試行錯誤による学習で得られます
- ドーパミンは 報酬予測の誤りをコードすることで この学習を導くと仮定されています
- 以前の研究では ドーパミンが成虫のゼブラフィンチで 性能予測の誤りをコードしていることが示されました
研究 の 目的:
- 若いゼブラフィンチの声の学習が ドーパミンベースの強化によって起こるかどうかを調査する.
- ドーパミンの活性を調べるため
主な方法:
- 幼いゼブラフィンチの歌の学習経路を追跡する
- エリアXでドーパミンの活動をモニターする
- ドーパミンの活動と 曲の変動の関係を分析する
主要な成果:
- ドーパミンは 大人の歌に近い音節で活性化され 遠く離れた音節で抑制されます
- ドーパミンの活動は 将来の曲の進化を予測し 行動における 駆動的な役割を示しました
- ドーパミンの活動は,現在の曲と最近の曲のコントラストと相関しており,予測エラーのエンコーディングと一致しています.
結論:
- 幼いゼブラフィンチは ドーパミンを媒介した補強学習で 声の習得を行います
- ドーパミンが予測エラーを 暗号化する役割は 複雑な自然行動学習において 極めて重要なようです
- 強化学習モデルは生物学的システムにおける 自然な行動の習得を説明できる.
関連する概念動画
Behaviorism
2.2K
The field of behaviorism was pioneered by figures such as Ivan Pavlov, John B. Watson, and B.F. Skinner fundamentally shifted the focus of psychology to the observable and controllable aspects of human and animal behavior. This shift marked a critical evolution in the discipline, emphasizing scientific rigor and experimental methodology.
The core premise of behaviorism is its focus on observable behavior rather than internal thoughts or feelings. This approach argues that true scientific...
The core premise of behaviorism is its focus on observable behavior rather than internal thoughts or feelings. This approach argues that true scientific...
2.2K
Operant Conditioning
1.5K
Operant conditioning, a key concept in behavioral psychology, involves using reinforcement and punishment to alter the likelihood of a behavior being repeated. B.F. introduced this type of conditioning. Skinner focused on voluntary behaviors and the consequences that follow them, influencing whether these behaviors will be strengthened or diminished.
Reinforcement in operant conditioning can be positive or negative, both of which serve to increase the likelihood of a behavior. Positive...
Reinforcement in operant conditioning can be positive or negative, both of which serve to increase the likelihood of a behavior. Positive...
1.5K
Law of Effect
1.3K
B.F. Skinner, a prominent figure in behavioral psychology, introduced operant conditioning by emphasizing the role of consequences in shaping behavior. This theory builds upon the law of effect proposed by Edward Thorndike, which posits that behaviors followed by satisfying outcomes are likely to be repeated. In contrast, those followed by unsatisfying outcomes are less likely to recur.
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
1.3K
Operant Conditioning Intervention
33
Operant conditioning serves as a foundational principle in therapeutic interventions aimed at modifying maladaptive behaviors. Central to this approach is the notion that behaviors, both adaptive and maladaptive, are learned through reinforcement. By analyzing the environmental factors that reinforce problematic behaviors, clinicians can design interventions to weaken these reinforcements and replace maladaptive behaviors with healthier alternatives.
In operant conditioning, behaviors that are...
In operant conditioning, behaviors that are...
33
Timing and Consequences on Behavior
71
In operant conditioning, the timing of reinforcement is crucial. For animals like rats and cats, immediate reinforcement (within a few seconds) is much more effective than delayed reinforcement. For example, a food reward for a rat needs to follow within 30 seconds of pressing a bar to be effective.
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
71
Instinctive Drift
173
Instinctive drift refers to the tendency of animals to revert to their innate behaviors despite repeated reinforcement. Breland and Breland demonstrated this concept in an experiment with a raccoon. The raccoon was trained to pick up two coins and place them in a container in exchange for food. Initially, the raccoon learned to associate the coins with food, making them a conditioned stimulus or a substitute for food. However, over time, the raccoon became less willing to put the coins into the...
173


