Related Experiment Video
Updated: Jul 22, 2026

Acquisition of a High-precision Skilled Forelimb Reaching Task in Rats
Published on: June 22, 2015
A High-Efficient Reinforcement Learning Approach for Dexterous Manipulation
Jianhua Zhang1, Xuanyi Zhou2, Jinyu Zhou2
1College of Mechanical Engineering, Beijing University of Science and Technology, Beijing 100083, China.
Abstract:
Robotic hands have the potential to perform complex tasks in unstructured environments owing to their bionic design, inspired by the most agile biological hand. However, the modeling, planning and control of dexterous hands remain unresolved, open challenges, resulting in the simple movements and relatively clumsy motions of current robotic end effectors. This paper proposed a dynamic model based on generative adversarial architecture to learn the state mode of the dexterous hand, reducing the model's prediction error in long spans. An adaptive trajectory planning kernel was also developed to generate High-Value Area Trajectory (HVAT) data according to the control task and dynamic model, with adaptive trajectory adjustment achieved by changing the Levenberg-Marquardt (LM) coefficient and the linear searching coefficient. Furthermore, an improved Soft Actor-Critic (SAC) algorithm is designed by combining maximum entropy value iteration and HVAT value iteration. An experimental platform and simulation program were built to verify the proposed method with two manipulating tasks. The experimental results indicate that the proposed dexterous hand reinforcement learning algorithm has better training efficiency and requires fewer training samples to achieve quite satisfactory learning and control performance.
Related Concept Videos
Long-term Potentiation
Long-term Potentiation
Hebbian LTP
LTP can occur when presynaptic neurons...
Root-Locus Method
This system can be represented by a block diagram,...
Purposive Learning
Mnemonic Devices
Acronyms
Acronyms are created by using the initial letters of a series of words to form a new word or phrase. This approach condenses complex information into a single, memorable entity. For example,...
Methods of Medium Optimization

