跳到主要导航 跳到搜索 跳到主要内容

Impulsive Maneuver Strategy Design for Target-Attacker -Defender Game of Spacecrafts via A Curriculum Learning Based Hamilton -Regularized MATD3 Algorithm

  • Northwestern Polytechnical University Xian

科研成果: 期刊稿件文章同行评审

摘要

This study presents a reinforcement learning framework for solving constrained spacecraft target-attacker-defender (TAD) problem in 3-dimensional orbital environment. We formulate the TAD problem using state transition of relative orbital dynamics under impulsive maneuver and multiple practical constraints are considered. To solve the complex TAD game problem, a Hamilton-regularized multi-agent twin-delayed deep deterministic policy gradient algorithm is proposed which achieves an efficiently training by guiding the spacecrafts exploring the optimal strategy effectively. An adaptive curriculum learning methodology is designed which is driven automatically such that the convergence of strategies is improved progressively. The reward decomposition combined with nonlinear process-oriented and terminal condition rewards can solve the sparse reward issue and therefore make the strategy avoid converging to local optima. A sun-elevation angle adaptation method is introduced to amend the strategy that makes the constraint be satisfied in TAD game. Numerical simulations and comparison results demonstrate the effectiveness of the proposed method.

指纹

探究 'Impulsive Maneuver Strategy Design for Target-Attacker -Defender Game of Spacecrafts via A Curriculum Learning Based Hamilton -Regularized MATD3 Algorithm' 的科研主题。它们共同构成独一无二的指纹。

引用此