Related Experiment Video
Updated: May 10, 2026

Studying the Neural Basis of Adaptive Locomotor Behavior in Insects
Published on: April 13, 2011
Emergence of natural and robust bipedal walking by learning from biologically plausible objectives
Pierre Schumacher1,2, Thomas Geijtenbeek3, Vittorio Caggiano4
1Max-Planck Institute for Intelligent Systems, Tübingen, Germany.
Abstract:
Humans show unparalleled ability when maneuvering diverse terrains. While reinforcement learning (RL) has shown great promise for musculoskeletal simulation in the development of robust controllers, complex behaviors are only achievable under extensive use of motion data. We demonstrate that the combination of a recent RL algorithm with a biologically plausible reward is capable of learning controllers for 4 different musculoskeletal models and achieves locomotion with up to 90 muscles without demonstrations. Our controllers generalize to diverse and unseen terrains, while only a single adaptive objective function is needed for training. We validate our findings on four models in two different simulators. The RL agents perform robustly with complex 3D models, where reflex-controllers are difficult to apply, and produce close-to-natural motion. This is a first step for the motor control, biomechanics, and rehabilitation communities to generate complex human movements with RL, without using motion data or simple unrepresentative models.
Related Concept Videos
What is Natural Selection?
Limits to Natural Selection
The Colonization of Land
The Evidence for Evolution
Introduction to Joints
Muscles of the Leg that Move the Foot and Toes
Anterior Compartment
The anterior compartment includes muscles that contribute to the dorsiflexion of the foot. This compartment houses the tibialis anterior, extensor hallucis longus, and extensor digitorum longus muscles.

