Masking and Demasking Agents
Collisions in Multiple Dimensions: Problem Solving
Reinforcement
Three-Dimensional Force System:Problem Solving
Avoidance Learning and Learned Helplessness
Observational Learning
您也可能阅读
通过共同作者、期刊和引用图与本文相关的文章。
Updated: Jul 23, 2025

The HoneyComb Paradigm for Research on Collective Human Behavior
Published on: January 19, 2019
He Cai1, Yaoguo Luo1, Huanli Gao1
1School of Automation Science and Engineering, South China University of Technology, Guangzhou 510641, China.
本研究介绍了一种多相半静态训练方法,用于群体对抗中的多代理深度强化学习 (MDRL). 这种方法提高了培训效率,使较弱的代理商能够更有效地从更强的代理商那里学习.
科学领域:
背景情况:
研究的目的:
主要方法:
主要成果:
结论: