探索基于奖励的学习策略对第二语言语音的有效性
Craig A Thorburn1, Ellen Lau2, Naomi H Feldman2,3
1Department of Psychology, University of Texas at Austin, Sarah M. & Charles E. Seay Bldg 108 E Dean Keeton St, Austin, TX, 78712, USA. craig.thorburn@austin.utexas.edu.
Psychonomic bulletin & review
|August 7, 2024
概括
成年人可以使用强化学习有效地学习新的语音,特别是当声音具有功能意义时. 一个深度强化网络与实验中的人类学习行为密切匹配.
科学领域:
- 认知科学 认知科学
- 神经科学是一个神经科学.
- 计算语言学 计算语言学
背景情况:
- 成年人经常在传统环境中与非母语语言类别学习作斗争.
- 一个视频游戏范例表明,当声音具有功能意义时,学习效率高.
- 行为和神经数据表明,强化学习机制参与了语音类别的获取.
研究的目的:
- 以计算形式形式化和测试强化学习是语音类别学习的基础的假设.
- 将深度强化学习网络与语音学习的监督模型进行比较.
- 研究特定神经电路在有效的语音学习中的作用.
主要方法:
- 实施了深度强化学习网络,以建模环境输入和行动之间的映射.
- 将强化网络的表现与监督学习模型进行了比较.
- 通过两项实验评估模型:学习合成的听觉噪声令牌和改善语音声音区分.
主要成果:
- 在这两次实验中,强化网络的行为与人类的表现非常相匹配.
- 强化模型和监督模型都显示了可比的性能.
- 模型输出的相似性表明,仅基于奖励的学习就没有固有的计算优势.
结论:
- 强化学习机制与成人语音类别学习有关.
- 特定的神经回路,特别是条纹体和上区域之间的联系,对于有效的学习至关重要.
- 虽然强化学习模拟了人类的行为,但对监督学习的计算效益可能是最小的,这凸显了神经机制的重要性.
更多相关视频
相关概念视频
Law of Effect
1.3K
B.F. Skinner, a prominent figure in behavioral psychology, introduced operant conditioning by emphasizing the role of consequences in shaping behavior. This theory builds upon the law of effect proposed by Edward Thorndike, which posits that behaviors followed by satisfying outcomes are likely to be repeated. In contrast, those followed by unsatisfying outcomes are less likely to recur.
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
Edward Thorndike's foundational work involved studying learning in animals, particularly using puzzle...
1.3K
Reinforcement
193
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
193
Language Development
329
Children master language quickly and with relative ease, supported by both biological predisposition and reinforcement. B. F. Skinner (1957) proposed that language is learned through reinforcement, while Noam Chomsky (1965) argued that language acquisition mechanisms are biologically determined.
The critical period for language acquisition suggests that the ability to acquire language is at its peak early in life. As people age, this proficiency decreases. Language development begins very...
The critical period for language acquisition suggests that the ability to acquire language is at its peak early in life. As people age, this proficiency decreases. Language development begins very...
329
Primary and Secondary Reinforcers
222
In psychology, reinforcement is a key concept in behavior modification. B.F. Skinner demonstrated this with his experiments involving rats in what is known as a Skinner box. The rats learned to press a lever to receive food, a primary reinforcer that fulfilled their innate need for nourishment.
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
Effective reinforcers for humans vary depending on the individual and the context. Primary reinforcers, such as food, water, sleep, shelter, and pleasure, have inherent value and satisfy basic biological...
222
Reinforcement Schedules
139
Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
Once a behavior is learned,...
139
Elaborative Rehearsals
83
Elaborative rehearsal is a crucial cognitive strategy that strengthens information encoding in long-term memory by making meaningful connections between new data and pre-existing knowledge. This approach contrasts with maintenance rehearsal, which involves simple repetition without delving into the significance of the information. While maintenance rehearsal might temporarily keep information active in short-term memory, it is less effective for long-term retention.
The effectiveness of...
The effectiveness of...
83


