跳到主要导航 跳到搜索 跳到主要内容

Weighted Mean Field Q-Learning for Large Scale Multiagent Systems

  • Northwestern Polytechnical University Xian
  • University of Adelaide

科研成果: 期刊稿件文章同行评审

6 引用 (Scopus)

摘要

Mean field reinforcement learning (MFRL) addresses the problem of dimensional explosion for large-scale multiagent systems. However, MFRL averages the actions of neighbors equally while discarding the diversity and distinct features between individuals, which may lead to poor performance in many application scenarios. In this article, a new MFRL algorithm termed temporal weighted mean filed Q-learning (TWMFQ) is proposed. TWMFQ introduces a temporal compensated multihead attention structure to construct the weighted mean-field framework, which can sort out the complex relationships within the swarm into the interactions between specific agent and the weighted virtual mean agent. This approach allows the mean Q-function to represent the swarm behavior more informatively and comprehensively. In addition, an advanced sampling mechanism called mixed experience replay is established, which enriches the diversity of samples and prevents the algorithm from falling into local optimal solution. The comparison experiments on MAgent and multi-USV platform justify the superior performance of TWMFQ across different population sizes.

源语言英语
页(从-至)7368-7378
页数11
期刊IEEE Transactions on Industrial Informatics
21
9
DOI
出版状态已出版 - 2025

指纹

探究 'Weighted Mean Field Q-Learning for Large Scale Multiagent Systems' 的科研主题。它们共同构成独一无二的指纹。

引用此