CORTEXA
← Browse
crossrefSensors2026-07-17Cited by 0

Deep Reinforcement Learning-Based Path-Following Control for Underactuated Autonomous Underwater Vehicles

Xin Pan, Lin Huang, Liangjin Li, Song Wang

Autonomous Underwater Vehicles (AUVs) face significant challenges in path-following control due to strong environmental disturbances and model uncertainties. To address these issues, this paper proposes a model-free deep reinforcement learning framework, named ILLT (Improved LOS-LSTM-TD3), which integrates an integral line-of-sight (LOS) guidance law with the twin delayed deep deterministic policy gradient (TD3) algorithm. The framework treats the LOS look-ahead distance as a learnable optimization variable and incorporates an LSTM network to capture temporal motion dependencies. A progressive unfreezing transfer learning strategy, combined with attention-based feature–current fusion, is designed to enhance domain adaptation under varying ocean currents. Simulation results demonstrate that ILLT reduces the average cross-track error by 48.5% compared to the baseline ILT algorithm and by 66.4% compared to traditional PID control, while achieving significantly faster convergence in target domains. Physical experiments in tank and lake environments further validate the algorithm’s feasibility and robustness, with tracking errors approaching simulation results under moderate current conditions. These findings confirm the effectiveness of the proposed framework for underactuated AUV path-following tasks.

View free PDFSource page

Related papers

crossrefSensors2025-08-12Cited by 5

A Deep Learning-Based Machine Vision System for Online Monitoring and Quality Evaluation During Multi-Layer Multi-Pass Welding

Van Doi Truong, Yunfeng Wang, Chanhee Won, Jonghun Yoon

Multi-layer multi-pass welding plays an important role in manufacturing industries such as nuclear power plants, pressure vessel manufacturing, and ship building. However, distortion or welding defects are still challenges; therefore, welding monitoring and quality control are es…

View free PDFSource page
crossrefSensors2025-08-21Cited by 4

Dynamic Allocation of C-V2X Communication Resources Based on Graph Attention Network and Deep Reinforcement Learning

Zhijuan Li, Guohong Li, Zhuofei Wu, Wei Zhang, Alessandro Bazzi

Vehicle-to-vehicle (V2V) and vehicle-to-network (V2N) communications are two key components of intelligent transport systems (ITSs) that can share spectrum resources through in-band overlay. V2V communication primarily supports traffic safety, whereas V2N primarily focuses on inf…

View free PDFSource page
crossrefSensors2026-06-27

Multi-Component Joint Maintenance Decision for Electro-Hydraulic Servo Fatigue Testing Machine Based on Multi-Head Deep Reinforcement Learning

Peng Liu, Guotai Huang, Jialu Xi, Jiaqi Wu

To address the challenge of maintenance decision-making for critical components in electro-hydraulic servo material fatigue testing machine, characterized by weak state observability and difficulty in degradation prediction, a multi-component joint maintenance decision-making met…

View free PDFSource page
crossrefSensors2026-04-01

Deep Reinforcement Learning for Autonomous Underwater Navigation: A Comparative Study with DWA and Digital Twin Validation

Zamirddine Mari, Mohamad Motasem Nawaf, Pierre Drap

Autonomous navigation in underwater environments is challenged by the absence of GPS, degraded visibility, and submerged obstacles. This article investigates these issues using the BlueROV2, an open platform for scientific experimentation. We propose a deep reinforcement learning…

View free PDFSource page
crossrefSensors2025-03-05Cited by 16

Machine Learning-Based Computer Vision for Depth Camera-Based Physiotherapy Movement Assessment: A Systematic Review

Yafeng Zhou, Fadilla ’Atyka Nor Rashid, Marizuana Mat Daud, Mohammad Kamrul Hasan, Wangmei Chen

Machine learning-based computer vision techniques using depth cameras have shown potential in physiotherapy movement assessment. However, a comprehensive understanding of their implementation, effectiveness, and limitations remains needed. Following PRISMA guidelines, we systemat…

View free PDFSource page