Related Experiment Video
Updated: Dec 28, 2025

11:18
Quantifying Learning in Young Infants: Tracking Leg Actions During a Discovery-learning Task
Published on: June 1, 2015
11.0K
Learning Mobile Manipulation through Deep Reinforcement Learning.
Cong Wang1,2,3,4, Qifeng Zhang1,2, Qiyan Tian1,2
1State Key Laboratory of Robotics, Shenyang Institute of Automation, Chinese Academy of Sciences, Shenyang 110016, China.
Sensors (Basel, Switzerland)
|February 14, 2020
Summary
This study introduces a novel deep reinforcement learning system for mobile manipulation. The system enables robots to autonomously grasp objects in unstructured environments using only onboard sensors.
Area of Science:
- Robotics
- Artificial Intelligence
- Machine Learning
Background:
- Mobile manipulation presents significant challenges due to the complex coordination required between a mobile base and a manipulator.
- Existing deep reinforcement learning (DRL) methods are often not suitable for mobile manipulation tasks, especially in unstructured environments.
Purpose of the Study:
- To investigate the application of deep reinforcement learning (DRL) for whole-body mobile manipulation in unstructured environments.
- To develop a novel mobile manipulation system utilizing only onboard sensors.
Main Methods:
- Proposed a new mobile manipulation system integrating state-of-the-art DRL algorithms with visual perception.
- Developed an efficient framework that decouples visual perception from DRL control for improved generalization.
- Enabled the system to learn from simulation and transfer to real-world testing.
Main Results:
- The proposed system demonstrated autonomous grasping of diverse objects in various simulated and real-world scenarios.
- Extensive simulations and experiments validated the system's effectiveness in complex environments.
- The decoupled framework facilitated successful generalization from simulation to real-world applications.
Conclusions:
- The developed deep reinforcement learning-based mobile manipulation system is effective for autonomous object grasping.
- The system's ability to generalize from simulation to the real world using onboard sensors is a key advancement.
- This research opens new possibilities for robust mobile manipulation in unstructured environments.
Related Concept Videos
Observational Learning
747
Albert Bandura's observational learning, also known as imitation or modeling, occurs when a person observes and imitates another's behavior. It is a quicker process than operant conditioning. A well-known example is the Bobo doll study, where children who saw an adult acting aggressively towards the doll were more likely to act aggressively when left alone, compared to those who observed a nonaggressive adult. Many psychologists view observational learning as a form of latent learning...
747
Reinforcement
739
Positive and negative reinforcement are key concepts in operant conditioning, a learning process where the consequences of a behavior affect the likelihood of that behavior being repeated.
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
Positive reinforcement occurs when a behavior is followed by the presentation of a rewarding stimulus, increasing the frequency of that behavior. For example:
739
Associative Learning
1.1K
Associative learning is a fundamental concept in behavioral psychology, wherein a connection is established between two stimuli or events, leading to a learned response. This process is critical in understanding how behaviors are acquired and modified. Conditioning, the mechanism through which associations are formed, can be divided into two main types: classical conditioning and operant conditioning, each elucidating different aspects of associative learning.
Classical conditioning, also known...
Classical conditioning, also known...
1.1K
Purposive Learning
382
E. C. Tolman emphasized the purposiveness of behavior — the idea that much of our behavior is goal-directed. For instance, employees who aim for a promotion work diligently to meet their targets. Tolman argued that when classical conditioning and operant conditioning occur, the organism acquires certain expectations. In classical conditioning, a child might fear a dog because they expect it to bite. In operant conditioning, a person might consistently work overtime because they expect a...
382

