跳到主要导航 跳到搜索 跳到主要内容

Goal-Conditioned Reinforcement Learning-Based Control for Multihorizon Underwater Manipulation Under Sparse Rewards

  • Northwestern Polytechnical University Xian

科研成果: 期刊稿件文章同行评审

摘要

Learning diverse goal-conditioned tasks under sparse rewards remains a major challenge in deep reinforcement learning (DRL), particularly for underwater robotic manipulation where dynamic disturbances hinder exploration and exacerbate hindsight bias. To address these challenges, this article proposes an integrated framework that combines a temporal-context attention-based adaptive intrinsic curiosity (TCA-AIC) mechanism with an importance-weighted similarity-based goal hindsight experience replay (IS-GHER) strategy. The TCA-AIC module enhances exploration efficiency and robustness to underwater disturbances by leveraging multistep temporal context and channel-wise attention in state prediction. Meanwhile, IS-GHER mitigates policy bias and stabilizes Q-value estimation by adaptively weighting hindsight relabeling samples based on goal similarity and training dynamics. The proposed framework is validated on several binaryreward underwater manipulation tasks in Gazebo simulation and further tested on a physical robot platform. Results show that it significantly improves learning efficiency and convergence speed compared with state-of-the-art baselines, particularly under sparse-reward and dynamically perturbed conditions.

源语言英语
期刊IEEE Transactions on Industrial Electronics
DOI
出版状态已接受/待刊 - 2026

学术指纹

探究 'Goal-Conditioned Reinforcement Learning-Based Control for Multihorizon Underwater Manipulation Under Sparse Rewards' 的科研主题。它们共同构成独一无二的学术指纹。

引用此