跳到主要导航 跳到搜索 跳到主要内容

Domain Knowledge-Assisted Deep Reinforcement Learning Power Allocation for MIMO Radar Detection

  • Yuedong Wang
  • , Yan Liang
  • , Huixia Zhang
  • , Yijing Gu
  • Northwestern Polytechnical University Xian

科研成果: 期刊稿件文章同行评审

18 引用 (Scopus)

摘要

The power allocation of multiple-input multiple-output (MIMO) radars is a key point in target tracking and detection. The optimality of a multitarget and multiconstraint optimization problem strictly depends on a priori model, which is difficult to obtain in time-varying, complex, and noncooperative environments. Recently, deep reinforcement learning (DRL) has been applied for target tracking tasks, which provides a trial-and-error interactive learning mechanism to improve the policy. Unlike tracking tasks with complete target state transition models, it remains an open issue for DRL-based MIMO radar detection that requires efficiently adapting the control policy to the environment of randomly appearing targets and extensive power transmission actions, which leads to sparse final task rewards and hence slow policy learning for the agent. Through introducing both the analytic model (radar equation) and empirical rules (expert preferences) for domain knowledge, this article proposes a domain-knowledge-assisted DRL (DKADRL) framework in which a domain-knowledge-based timely reward generator is utilized to generate timely rewards that assist the agent's policy learning. To adjust the role of the timely rewards and the final task rewards, a reward fusion module is designed, which gradually increases the role of the final task rewards as the training process progresses, thus allows agent's policy to converge to the final optimization goal. The algorithm is validated under two target motion scenarios, showing the higher target detection probability and the faster training speed, compared to equal power allocation and proximal policy optimization (PPO)-based power allocation.

源语言英语
页(从-至)23117-23128
页数12
期刊IEEE Sensors Journal
22
23
DOI
出版状态已出版 - 1 12月 2022

学术指纹

探究 'Domain Knowledge-Assisted Deep Reinforcement Learning Power Allocation for MIMO Radar Detection' 的科研主题。它们共同构成独一无二的学术指纹。

引用此