Related Experiment Video
Updated: Aug 19, 2025

New Variations for Strategy Set-shifting in the Rat
Published on: January 23, 2017
Mastering the game of Stratego with model-free multiagent reinforcement learning
Julien Perolat1, Bart De Vylder1, Daniel Hennes1
1DeepMind Technologies Ltd., London, UK.
Abstract:
We introduce DeepNash, an autonomous agent that plays the imperfect information game Stratego at a human expert level. Stratego is one of the few iconic board games that artificial intelligence (AI) has not yet mastered. It is a game characterized by a twin challenge: It requires long-term strategic thinking as in chess, but it also requires dealing with imperfect information as in poker. The technique underpinning DeepNash uses a game-theoretic, model-free deep reinforcement learning method, without search, that learns to master Stratego through self-play from scratch. DeepNash beat existing state-of-the-art AI methods in Stratego and achieved a year-to-date (2022) and all-time top-three ranking on the Gravon games platform, competing with human expert players.
Related Concept Videos
Mechanistic Models: Compartment Models in Algorithms for Numerical Problem Solving
In individual population analyses, different algorithms are employed, such as Cauchy's method, which uses a...
Reinforcement
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Reinforcement Schedules
Once a behavior is learned,...
Observational Learning
Multi-input and Multi-variable systems
In the absence...
Associative Learning
Classical conditioning, also known...

