Replicator-mutator dynamics of the rock-paper-scissors game: Learning through mistakes.

Suman Chakraborty1, Ishita Agarwal1, Sagar Chakraborty1

  • 1Department of Physics, Indian Institute of Technology Kanpur, Uttar Pradesh 208016, India.

Physical Review. E
|April 18, 2024
PubMed
Summary

Mistakes in learning models can surprisingly improve player strategies and even lead to rational Nash-equilibrium outcomes in games like rock-paper-scissors. This research introduces a new Hamiltonian structure for the replicator-mutator equation.

Related Concept Videos

Nonconscious Mimicry01:13

Nonconscious Mimicry

Nonconscious mimicry occurs when individuals alter their mannerisms to match the behaviors and expressions of those nearby, without intention.
4.6K
Observational Learning01:12

Observational Learning

Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
168
Instinctive Drift01:05

Instinctive Drift

Instinctive drift refers to the tendency of animals to revert to their innate behaviors despite repeated reinforcement. Breland and Breland demonstrated this concept in an experiment with a raccoon. The raccoon was trained to pick up two coins and place them in a container in exchange for food. Initially, the raccoon learned to associate the coins with food, making them a conditioned stimulus or a substitute for food. However, over time, the raccoon became less willing to put the coins into the...
209
Reinforcement Schedules01:24

Reinforcement Schedules

Positive reinforcement is a powerful method for teaching new behaviors to both animals and humans. B.F. Skinner demonstrated this with his experiments using rats in a Skinner box. When a rat pressed a lever, it received a food pellet. This immediate reward encouraged the rat to repeat the behavior. This method, where a reward follows every instance of the behavior, is known as continuous reinforcement. It is highly effective for establishing new behaviors quickly.
Once a behavior is learned,...
144
Cooperative Allosteric Transitions01:58

Cooperative Allosteric Transitions

2.3K