関連する実験動画
Updated: Sep 9, 2025

A Conflict Model of Reward-seeking Behavior in Male Rats
Published on: February 20, 2019
処罰を回避し,同時に報酬を追求する際に,作業記憶と補強学習がどのように相互作用するか
Peter F Hitchcock1, Joonhwa Kim2, Michael J Frank3
1Department of Psychology, Emory University.
人間では 補強学習 (RL) と 作業記憶 (WM) を別々に用いる. この研究では WMは即時の罰を回避するのに役立つが RLは罰の記憶に苦しんでおり 個人の違いがこの相互作用に影響していることが示されています
科学分野:
- 認知神経科学
- 行動心理学
- 計算精神科
背景:
- 人間の適応行動は強化学習 (RL) と作業記憶 (WM) システムに依存しています.
- 報酬と罰の同時学習における RLとWMの相互作用は十分に理解されていません.
- 精神症状がこれらの学習システムに与える影響については,さらなる調査が必要である.
研究 の 目的:
- RLとWMの間の相互作用を調査する 罰を回避する学習.
- 報酬と罰の同時学習で WM が RL にどのように影響するかを調べる.
- 鬱や不安や反省が 学習プロセスに与える影響を評価する
主な方法:
- 新しい報酬/罰 RL-WMタスクがオンラインサンプル (N=298) に与えられました.
- 学習パターンとシステムの相互作用を分析するために計算モデルが使用されました.
- 交差した試験段階では,RLベースの保持とRLに対するWMの影響を評価した.
主要な成果:
- 作業記憶 (WM) は即時処罰回避を容易にしたが,補強学習 (RL) システムではこれをうまく記憶できなかった.
- 学習の途中でRLを鈍化するWMの証拠が観察されましたが,この効果はさらなる学習後に減少しました.
- 個々の差異はWM-RLの相互作用を調節し,一部の個体はWMの保持促進を示した.
- 鬱や不安や反省にもかかわらず 課題の遂行能力は ほとんど変わらなかった
結論:
- WMシステムは短期的な処罰回避をサポートし,RLシステムは処罰情報の制限された保持を示しています.
- 個々の差異は,WMがRLに影響を与える程度に大きな影響を与えます.
- このタスクの行動能力は うつ病や不安症のような 精神症状を内在化させるのに 適しています
さらに関連する動画
14:24An Appetitive Spatial Working Memory Task for Mice in a Semi-Automated 8-Arm Radial Maze, Reducing Fearful Memory Association in the Maze
Published on: July 29, 2025
08:05A Prediction Error-driven Retrieval Procedure for Destabilizing and Rewriting Maladaptive Reward Memories in Hazardous Drinkers
Published on: January 5, 2018
関連する概念動画
Avoidance Learning and Learned Helplessness
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
Timing and Consequences on Behavior
Humans, however, can respond to delayed reinforcers. We often make decisions between immediate small rewards and delayed larger rewards. This ability to delay gratification is a significant...
Cognitive Learning
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
Associative Learning
Classical conditioning, also known...
Working Memory
Operant Conditioning
Reinforcement in operant conditioning can be positive or negative, both of which serve to increase the likelihood of a behavior. Positive...