Skip to main navigation Skip to search Skip to main content

ASF-HDRL: hierarchical reinforcement learning with attention-guided high-level switching and stage-conditioned feature modulation for multi-AUV cooperative tracking

  • Northwestern Polytechnical University Xian

Research output: Contribution to journalArticlepeer-review

Abstract

The cooperative tracking of Multi-Autonomous Underwater vehicles (AUV) has shown great potential in fields such as ocean environmental monitoring, marine resource exploration, and underwater security. However, limited underwater acoustic communication range, sparse deployment of sensor nodes, and environmental uncertainties often lead to incomplete target trajectory information and partially observable states, resulting in decreased tracking accuracy and unstable performance in complex scenarios. To solve the above problems, this paper proposes a Hierarchical Deep Reinforcement Learning framework, termed ASF-HDRL (Attention-guided high-level Switching with stage-conditioned Feature modulation), which decouples high- and low-level policies and incorporates a target estimation algorithm to enhance adaptability and robustness in complex tasks. Simulation results demonstrate that ASF-HDRL achieves superior cooperative tracking performance in challenging simulation environments, outperforming several mainstream baseline methods in terms of convergence speed and tracking accuracy.

Original languageEnglish
Pages (from-to)466-482
Number of pages17
JournalCCF Transactions on Pervasive Computing and Interaction
Volume8
Issue number3
DOIs
StatePublished - Sep 2026

UN SDGs

This output contributes to the following UN Sustainable Development Goals (SDGs)

  1. SDG 14 - Life Below Water
    SDG 14 Life Below Water

Keywords

  • Autonomous underwater vehicles
  • Multi-agent reinforcement learning
  • Stage transitions
  • Underwater target tracking

Fingerprint

Dive into the research topics of 'ASF-HDRL: hierarchical reinforcement learning with attention-guided high-level switching and stage-conditioned feature modulation for multi-AUV cooperative tracking'. Together they form a unique fingerprint.

Cite this