Related Experiment Video
Updated: Jan 20, 2026

Development of an Audio-based Virtual Gaming Environment to Assist with Navigation Skills in the Blind
Published on: March 27, 2013
Navigation in Unknown Dynamic Environments Based on Deep Reinforcement Learning
Junjie Zeng1, Rusheng Ju2, Long Qin3
1College of Systems Engineering, National University of Defense Technology, Changsha 410073, China. zengjunjie13@nudt.edu.cn.
Abstract:
In this paper, we propose a novel Deep Reinforcement Learning (DRL) algorithm which can navigate non-holonomic robots with continuous control in an unknown dynamic environment with moving obstacles. We call the approach MK-A3C (Memory and Knowledge-based Asynchronous Advantage Actor-Critic) for short. As its first component, MK-A3C builds a GRU-based memory neural network to enhance the robot's capability for temporal reasoning. Robots without it tend to suffer from a lack of rationality in face of incomplete and noisy estimations for complex environments. Additionally, robots with certain memory ability endowed by MK-A3C can avoid local minima traps by estimating the environmental model. Secondly, MK-A3C combines the domain knowledge-based reward function and the transfer learning-based training task architecture, which can solve the non-convergence policies problems caused by sparse reward. These improvements of MK-A3C can efficiently navigate robots in unknown dynamic environments, and satisfy kinetic constraints while handling moving objects. Simulation experiments show that compared with existing methods, MK-A3C can realize successful robotic navigation in unknown and challenging environments by outputting continuous acceleration commands.
Related Concept Videos
09:01Development of an Audio-based Virtual Gaming Environment to Assist with Navigation Skills in the Blind
10:25Deep Learning-Based Segmentation of Cryo-Electron Tomograms
05:42Dynamic Navigation for Dental Implant Placement
08:12Two-photon Calcium Imaging in Mice Navigating a Virtual Reality Environment
04:17DNA Virus Detection System Based on RPA-CRISPR/Cas12a-SPM and Deep Learning
07:03Dynamic Navigation in Endodontics: Guided Access Cavity Preparation by Means of a Miniaturized Navigation System

