Related Experiment Video
Updated: Sep 19, 2025

The HoneyComb Paradigm for Research on Collective Human Behavior
Published on: January 19, 2019
Unilateral incentive alignment in two-agent stochastic games
Alex McAvoy1,2, Udari Madhushani Sehwag3,4, Christian Hilbe5
1School of Data Science and Society, University of North Carolina at Chapel Hill, Chapel Hill, NC 27599.
Abstract:
Multiagent learning is challenging when agents face mixed-motivation interactions, where conflicts of interest arise as agents independently try to optimize their respective outcomes. Recent advancements in evolutionary game theory have identified a class of "zero-determinant" strategies, which confer an agent with significant unilateral control over outcomes in repeated games. Building on these insights, we present a comprehensive generalization of zero-determinant strategies to stochastic games, encompassing dynamic environments. We propose an algorithm that allows an agent to discover strategies enforcing predetermined linear (or approximately linear) payoff relationships. Of particular interest is the relationship in which both payoffs are equal, which serves as a proxy for fairness in symmetric games. We demonstrate that an agent can discover strategies enforcing such relationships through experience alone, without coordinating with an opponent. In finding and using such a strategy, an agent ("enforcer") can incentivize optimal and equitable outcomes, circumventing potential exploitation. In particular, from the opponent's viewpoint, the enforcer transforms a mixed-motivation problem into a cooperative problem, paving the way for more collaboration and fairness in multiagent systems.
Related Concept Videos
Incentive Theory: Pull Theory of Motivation
The theory differentiates between...
Cooperative Allosteric Transitions
Secondary Motives: Affiliation Motivation and Aggression Motivation
Compensation Mechanisms
Respiratory Compensation
This mechanism addresses metabolic-induced pH imbalances by adjusting breathing rates. Respiratory compensation begins within minutes of detecting a pH...
Alternative Sets of Equilibrium Equations
One example of such a situation can be observed in a...
Robbers Cave

