摘要
A deep reinforcement learning-based trajectory optimization method is investigated to address the challenges posed by complex propulsion-mode transitions and strongly coupled process constraints in the ascent trajectory optimization of RBCC-powered aerospace vehicles. A longitudinal dynamical and aerodynamic model covering the ejector, ramjet, scramjet, and rocket modes is constructed, and a continuous action space centered on angle-of-attack rate is formulated within a Markov decision process framework. A reward structure is designed to balance feasibility constraints, flight-load constraints, and terminal mission requirements. A guidance-channel compensation mechanism based on a time-sequenced action library is further introduced to steer value estimation during policy updates, thereby enhancing the convergence efficiency and training stability of the deep deterministic policy gradient algorithm under multimodal propulsion conditions. Simulation analyses indicate that the guidance mechanism can improve convergence behavior, reduce peak process loads, and yield smoother control inputs to a certain extent. The results demonstrate the feasibility and potential of incorporating a guidance channel into reinforcement learning-based trajectory optimization, providing an extensible pathway toward intelligent trajectory planning for complex combined-cycle aerospace vehicles.
| 投稿的翻译标题 | An improved DDPG-based trajectory optimization for ascent phase of vehicle with combined-cycle engine |
|---|---|
| 源语言 | 繁体中文 |
| 页(从-至) | 83-92 |
| 页数 | 10 |
| 期刊 | Aerospace Technology |
| 期 | 1 |
| DOI | |
| 出版状态 | 已出版 - 2月 2026 |
关键词
- DDPG method
- aerospace vehicle
- ascent stage
- combined cycle engine
- trajectory opti⁃ mization
指纹
探究 '基于改进 DDPG 的组合动力飞行器上升段轨迹优化' 的科研主题。它们共同构成独一无二的指纹。引用此
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver