摘要
This paper investigates an improved Proximity Policy Optimization(PPO)-based intelligent adaptive control algorithm for spacecraft swarm to achieve rapid rounding-up and high-precision tracking of targets in elliptical orbits. Firstly, an improved reinforcement learning algorithm is proposed. And a stability-constrained hybrid learning-control framework is developed, which can learn controller parameters by the improved algorithm. By analytically constraining the output space of the actor network through theoretical derivation, it can be guaranteed that all learned policies satisfy Lyapunov stability conditions a priori. Secondly, considering various constraints, a spacecraft swarm round-up algorithm is designed based on the learning-control framework. A hierarchical reward function is designed to integrate parameters with significantly different operational characteristics into a unified training framework. The composite reward function incorporates multiple constraints to guide the agent in learning optimal rounding-up strategies while achieving high-precision tracking control. Finally, numerical simulations with environment disturbances and sensor measurement noises demonstrate the effectiveness and robustness of the proposed controller.
| 源语言 | 英语 |
|---|---|
| 页(从-至) | 543-555 |
| 页数 | 13 |
| 期刊 | Acta Astronautica |
| 卷 | 249 |
| DOI | |
| 出版状态 | 已出版 - 12月 2026 |
学术指纹
探究 'Improved learning-based intelligent adaptive control for spacecraft swarm rounding-up in elliptical orbits' 的科研主题。它们共同构成独一无二的学术指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver