Associative Learning
Purposive Learning
The Anchoring-and-Adjustment Heuristic
Observational Learning
Cognitive Learning
Hindsight Biases
您也可能阅读
通过共同作者、期刊和引用图与本文相关的文章。
在强化学习 (RL) 中的优势学习 (AL) 提供了稳定性,但趋同速度较慢. 基于Occam's Razor的AL (ORAL) 可自适应地调整动作差距,提高融合速度和复杂任务的性能.
科学领域:
背景情况:
研究的目的:
主要方法:
主要成果:
结论: