相关实验视频
Updated: Jun 15, 2026

06:31
Force and Position Control in Humans - The Role of Augmented Feedback
Published on: June 19, 2016
7.8K
强化学习可能会揭开人类有限的运动学习效率的神秘性,这是由于视觉-主观认知不匹配
Kyungrak Choi1, Yoonsuck Choe2, Hangue Park1,3,4
1Department of Electrical and Computer Engineering, Texas A&M University, College Station, TX 77843, USA.
International journal of neural systems
|April 24, 2024
概括
视觉和自身感知之间的感觉不匹配阻碍了运动学习. 模拟显示更大的感觉偏移和类似的奖励幅度降低了准确性,而探索率影响了速度和准确性.
科学领域:
- 发动机控制器 发动机控制器
- 计算神经科学是一种计算神经科学.
- 人与计算机的互动.
背景情况:
- 众所周知,视觉和自身感知之间的感觉不匹配限制了运动学习.
- 这种感官不匹配对运动学习结果的确切影响和机制尚不清楚.
研究的目的:
- 通过计算来研究视觉与自身感知感官不匹配如何影响运动学习.
- 量化传感不匹配大小与运动控制精度/速度之间的关系.
主要方法:
- 在模拟环境中利用了强化学习算法.
- 在曲任务中使用了肘部关节的简化生物机械模型.
- 在不同程度的视觉-自身感知角度偏移和奖励幅度下模拟运动学习.
主要成果:
- 视觉和自身感知之间的感知角偏移增加与运动控制精度下降直接相关.
- 两种感官模式之间的峰值奖励幅度更大的相似性也导致了运动控制精度的降低.
- 勘探速度极大地影响了性能:不足的速度限制了任务完成的速度,而过度的速度损害了运动控制的准确性.
结论:
- 视觉-自身感知感官不匹配显著降低了运动学习性能,特别是准确性.
- 感官偏移的程度和奖励一致性是运动控制有效性的关键决定因素.
- 适度的探索速度对于在运动学习中平衡速度和准确性至关重要,强调了它的重要性.
相关概念视频
Hindsight Biases
Hindsight bias leads you to believe that the event you just experienced was predictable, even though it really wasn’t. In other words, you knew all along that things would turn out the way they did. Can you relate this to the phrase "Hindsight is 20/20" now?
Social Facilitation
Not all intergroup interactions lead to negative outcomes. Sometimes, being in a group situation can improve performance. Social facilitation occurs when an individual performs better when an audience is watching than when the individual performs the behavior alone. This typically occurs when people are performing a task for which they are skilled.
Higher Mental Functions of Brain: Learning and Memory
Memory is one of the most vital higher mental functions of the brain. Memory is closely related to learning because it enables us to retain information and experiences from our past to use them in our present life. It also helps us to remember facts, events, and skills, such as riding a bike or swimming. There are two types of memory — declarative memory, which involves memorizing facts or events, and procedural memory, which enables us to remember how to do something like writing or playing an...
Parallel Processing
The brain processes sensory information rapidly due to parallel processing, which involves sending data across multiple neural pathways at the same time. This method allows the brain to manage various sensory qualities, such as shapes, colors, movements, and locations, all concurrently. For instance, when observing a forest landscape, the brain simultaneously processes the movement of leaves, the shapes of trees, the depth between them, and the various shades of green. This enables a quick and...
Cognitive Learning
Cognitive learning is based on purposive behavior, incidental learning, and insight learning.
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
E. C. Tolman's theory of purposive behavior emphasizes that much behavior is goal-directed. He argued that to understand behavior, we must look at the entire sequence of actions leading to a goal. For instance, high school students study hard, not just due to past reinforcement but also to achieve the goal of getting into a good college.
Tolman introduced the idea that behavior is influenced by...
Purposive Learning
E. C. Tolman emphasized the purposiveness of behavior — the idea that much of our behavior is goal-directed. For instance, employees who aim for a promotion work diligently to meet their targets. Tolman argued that when classical conditioning and operant conditioning occur, the organism acquires certain expectations. In classical conditioning, a child might fear a dog because they expect it to bite. In operant conditioning, a person might consistently work overtime because they expect a bonus...

