摘要
This study presents a reinforcement learning framework for solving constrained spacecraft target-attacker-defender (TAD) problem in 3-dimensional orbital environment. We formulate the TAD problem using state transition of relative orbital dynamics under impulsive maneuver and multiple practical constraints are considered. To solve the complex TAD game problem, a Hamilton-regularized multi-agent twin-delayed deep deterministic policy gradient algorithm is proposed which achieves an efficiently training by guiding the spacecrafts exploring the optimal strategy effectively. An adaptive curriculum learning methodology is designed which is driven automatically such that the convergence of strategies is improved progressively. The reward decomposition combined with nonlinear process-oriented and terminal condition rewards can solve the sparse reward issue and therefore make the strategy avoid converging to local optima. A sun-elevation angle adaptation method is introduced to amend the strategy that makes the constraint be satisfied in TAD game. Numerical simulations and comparison results demonstrate the effectiveness of the proposed method.
| 源语言 | 英语 |
|---|---|
| 期刊 | IEEE Transactions on Aerospace and Electronic Systems |
| DOI | |
| 出版状态 | 已接受/待刊 - 2026 |
指纹
探究 'Impulsive Maneuver Strategy Design for Target-Attacker -Defender Game of Spacecrafts via A Curriculum Learning Based Hamilton -Regularized MATD3 Algorithm' 的科研主题。它们共同构成独一无二的指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver