Related Experiment Video
Updated: Oct 16, 2025

11:54
Real-Time Proxy-Control of Re-Parameterized Peripheral Signals using a Close-Loop Interface
Published on: May 8, 2021
4.7K
Decentralized control and local information for robust and adaptive decentralized Deep Reinforcement Learning
Malte Schilling1, Andrew Melnik2, Frank W Ohl3
1Machine Learning Group, Bielefeld University, 33501 Bielefeld, Germany.
Summary
Decentralized control architectures enhance learning speed and robustness in Deep Reinforcement Learning (DRL) for motor control tasks. This approach shows promise for more adaptive and generalized robotic behaviors.
Area of Science:
- Robotics
- Artificial Intelligence
- Computational Neuroscience
Background:
- Biological motor control utilizes decentralization for rapid, localized responses.
- Current Deep Reinforcement Learning (DRL) for motor control predominantly uses centralized architectures.
- Centralized DRL controllers must process extensive sensory information, potentially limiting efficiency.
Purpose of the Study:
- To investigate the benefits of decentralized control architectures in DRL for embodied sensori-motor control.
- To compare centralized versus decentralized DRL approaches for adaptive locomotion in a four-legged agent.
- To evaluate learning speed, robustness, and generalization capabilities across varying degrees of decentralization.
Main Methods:
- Developed and analyzed eight distinct control architectures for a four-legged agent.
- Systematically varied the degree of decentralization from fully centralized to fully decentralized.
- Assessed performance based on learning speed, robustness to hyperparameter settings, and generalization to untrained terrains.
Main Results:
- Distributed architectures significantly enhanced learning speed compared to centralized ones.
- Decentralized control demonstrated increased robustness, requiring less hyperparameter tuning and avoiding local minima.
- An intermediate decentralized architecture, integrating local information from neighboring legs, showed superior generalization to uneven terrains.
Conclusions:
- Distributing control into decentralized units leveraging local information offers significant advantages for DRL-based motor control.
- Decentralization improves learning efficiency, robustness, and generalization capabilities in adaptive locomotion tasks.
- This approach presents a promising direction for developing more resilient and adaptable DRL systems.
Related Concept Videos
Observational Learning
386
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
386
Reinforcement
457
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
457
Neural Control of Respiration
3.3K
The neural regulation of respiration is a meticulously coordinated process primarily controlled by the respiratory centers located within the brainstem. These centers, composed of specialized neurons, transmit nerve impulses that control the contraction and relaxation of our respiratory muscles.
Respiratory Centers in the Brainstem
Two primary areas comprise the respiratory center: the medullary respiratory center in the medulla oblongata and the pontine respiratory group in the pons. The...
Respiratory Centers in the Brainstem
Two primary areas comprise the respiratory center: the medullary respiratory center in the medulla oblongata and the pontine respiratory group in the pons. The...
3.3K
Neural Regulation
40.5K
Digestion begins with a cephalic phase that prepares the digestive system to receive food. When our brain processes visual or olfactory information about food, it triggers impulses in the cranial nerves innervating the salivary glands and stomach to prepare for food.
40.5K
Open and closed-loop control systems
1.1K
Control systems are foundational elements in automation and engineering. They are broadly categorized into open-loop and closed-loop systems. These classifications hinge on the presence or absence of feedback mechanisms, significantly influencing the system's performance, complexity, and application.
An open-loop control system operates without feedback from the output. It consists of two primary elements: the controller and the controlled process. The controller receives an input signal...
An open-loop control system operates without feedback from the output. It consists of two primary elements: the controller and the controlled process. The controller receives an input signal...
1.1K
Hierarchy of Motor Control
4.1K
The hierarchy of motor control refers to the different levels of organization and processing involved in controlling movement in the body. These levels range from higher cortical areas involved in planning and decision-making to lower spinal cord reflexes that respond automatically to external stimuli.
4.1K

