相关实验视频
Updated: Jul 17, 2025

06:48
The HoneyComb Paradigm for Research on Collective Human Behavior
Published on: January 19, 2019
9.4K
在人类群体中建立合作架构,并进行深度强化学习
Kevin R McKee1, Andrea Tacchetti2, Michiel A Bakker2
1Google DeepMind, London, UK. kevinrmckee@google.com.
Nature human behaviour
|September 7, 2023
概括
深度学习成功地鼓励了游戏中的小组合作. 经过培训的社会规划师通过适应性网络策略将合作率提高了77.7%.
科学领域:
- 计算社会科学 计算社会科学
- 人工智能在社会动态中的作用
背景情况:
- 鼓励群体合作是社会动态中持续存在的挑战.
- 以前的策略通常涉及孤立不合作的个人.
研究的目的:
- 将深度学习应用到动态结构网络,以加强群体合作.
- 开发和测试一个人工智能驱动的"社会计划器",以优化社交互动.
主要方法:
- 利用深度强化学习和模拟来训练一个社会规划师.
- 社会规划者建议参与者之间建立网络联系 (建立/破坏联系).
- 在一个集团合作游戏中测试了人工智能策略,其中涉及到真正的货币风险.
主要成果:
- 使用社会规划器的群体实现了77.7%的合作率,明显高于静态网络 (42.8%).
- 社会规划者采用了和解的方法,将叛逃者纳入合作社区.
- 这与传统的分离叛逃者的方法形成了鲜明对比.
结论:
- 人工智能驱动的网络结构可以有效地促进群体中的亲社会行为.
- 适应性,调和性策略比隔离更有效地促进合作.
- 这种方法为管理群体动态和合作提供了一种新的方法.
相关概念视频
Observational Learning
207
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
207
Reinforcement
273
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
273
Reinforcement Schedules
202
Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
Once a behavior is learned,...
202
Associative Learning
434
Associative learning is a fundamental concept in behavioral psychology, wherein a connection is established between two stimuli or events, leading to a learned response. This process is critical in understanding how behaviors are acquired and modified. Conditioning, the mechanism through which associations are formed, can be divided into two main types: classical conditioning and operant conditioning, each elucidating different aspects of associative learning.
Classical conditioning, also known...
Classical conditioning, also known...
434
Avoidance Learning and Learned Helplessness
1.8K
Avoidance learning and learned helplessness are critical concepts in understanding behavioral responses to negative stimuli.
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
Avoidance learning occurs when an organism learns that a specific behavior can prevent an unpleasant outcome. For example, a student who receives a bad grade may start studying harder to avoid future poor grades. This behavior persists even when the negative outcome is no longer present. Avoidance learning is powerful because it maintains behavior in the absence of the...
1.8K
Purposive Learning
139
E. C. Tolman emphasized the purposiveness of behavior — the idea that much of our behavior is goal-directed. For instance, employees who aim for a promotion work diligently to meet their targets. Tolman argued that when classical conditioning and operant conditioning occur, the organism acquires certain expectations. In classical conditioning, a child might fear a dog because they expect it to bite. In operant conditioning, a person might consistently work overtime because they expect a...
139

