跳到主要导航 跳到搜索 跳到主要内容

Graph-based strategy evaluation for large-scale multiagent reinforcement learning

  • Yiyun Sun
  • , Meiqin Liu
  • , Senlin Zhang
  • , Ronghao Zheng
  • , Shanling Dong
  • Zhejiang University
  • Jinhua Institute of Zhejiang University

科研成果: 期刊稿件文章同行评审

摘要

In large-scale multiagent systems, the practical application of multiagent reinforcement learning (MARL) is hindered by the absence of robust reliability assurances. This gap has heightened the focus on strategy evaluation within the MARL framework, a domain that grapples with scalability issues in the joint strategy space. To address this concern, this paper introduces a novel two-stage graph-based strategy evaluation algorithm that significantly reduces the required sample capacity in the joint strategy space without compromising the evaluation quality. The proposed algorithm performs a hierarchical evaluation to compress sample capacity and employs a strategy-seeking model to seek a sink equilibrium (SE) joint strategy using the best responses. Moreover, a stopping condition is developed to achieve an approximately globally optimal SE strategy, accounting for the local optimal properties of the best-response-based algorithm. Case studies demonstrate that our algorithm achieves an approximately optimal SE joint strategy with superior sample efficiency compared with other approaches. The integration of MARL methods with the strategy evaluation algorithm proves to be an effective approach for establishing trustworthy MARL systems.

源语言英语
期刊论文编号182206
期刊Science China Information Sciences
68
8
DOI
出版状态已出版 - 8月 2025

学术指纹

探究 'Graph-based strategy evaluation for large-scale multiagent reinforcement learning' 的科研主题。它们共同构成独一无二的学术指纹。

引用此