Related Experiment Video
Updated: May 14, 2026

The Modular Design and Production of an Intelligent Robot Based on a Closed-Loop Control Strategy
Published on: October 14, 2017
Autonomous Path Planning for USV Swarm Based on Dual-Module Learning MATD3
None:
The complex and dynamic maritime environment brings significant uncertainties to autonomous path planning tasks, posing substantial challenges for multiagent reinforcement learning (MARL). To address these challenges, this article analyzes the kinematic and dynamic models of unmanned surface vessels (USVs) as well as the environmental disturbances (e.g., wind, waves, and currents), and then proposes a dual-module learning multiagent twin delayed deep deterministic (DML-MATD3) policy gradient framework for USV swarm path planning based on the realistic physical conditions. The framework establishes a motion-generation module and a collision-avoidance module to enable simpler yet more effective reward designs for learning. Specifically, tailored reward functions are independently designed for each module. A potential field method (PFM) is introduced to provide dense and informative guidance for both motion-generation and collision-avoidance modules. Moreover, an adaptive energy consumption reward (AECR) is integrated into the motion-generation module to improve energy-efficient navigation under environmental disturbances. To further enhance exploration efficiency and reward responsiveness during early training, an Ornstein-Uhlenbeck noise-based action enhancement strategy (OU-AES) is employed. Extensive experiments against seven baseline algorithms demonstrate that the proposed DML-MATD3 consistently achieves faster convergence, improved training stability, shorter path lengths, reduced task execution times, and superior overall performance in complex maritime environments.
Related Concept Videos
One-Degree-of-Freedom System
A one-degree-of-freedom system is defined by an independent variable that determines its state and behavior. One example of a one-degree-of-freedom system is a simple harmonic oscillator, such as a...
Three-Dimensional Force System:Problem Solving
To solve a three-dimensional force system, first resolve each force into its respective scalar components. Do this using...
Relative Motion Analysis using Rotating Axes-Problem Solving
Here, in order to determine the magnitude of velocity and acceleration for point...
PID Controller
Indirect Motor Pathways
The vestibulospinal tract originates in the vestibular nuclei of the brainstem. The vestibular system detects changes in...
Multi-input and Multi-variable systems
In the absence of...