跳到主要导航 跳到搜索 跳到主要内容

Dual-scale optimization of integrated energy systems: a novel MAPPO-based approach for hybrid game equilibrium

  • Xi'an Jiaotong University
  • Eindhoven University of Technology

科研成果: 期刊稿件文章同行评审

6 引用 (Scopus)

摘要

With the accelerating global energy transition toward carbon neutrality, integrated energy systems face complex optimization challenges arising from renewable energy intermittency, diversified user demands, and the inherent time-scale misalignment between supplier pricing strategies and user responsive behaviors, creating multi-level game environments that existing methods struggle to address effectively, leading to insufficient equipment utilization and inadequate energy resource exploitation. To this end, this paper first establishes a comprehensive physical model incorporating electricity-gas-heat multi-energy coupling with diverse conversion equipment (electric boilers, heat pumps, gas boilers, CHP systems, and P2G facilities) and renewable energy sources. Second, Multi-Time-scale Graph-enhanced Multi-Agent Proximal Policy Optimization framework (MT-MAPPO) is proposed to address the hybrid game equilibrium problem in integrated energy system optimization. The framework constructs a comprehensive “vertical Stackelberg + horizontal competition” dual-game structure that combines supplier-user hierarchical relationships with peer-to-peer competitive dynamics among users. Third, Centralized Training with Decentralized Execution (CTDE) and policy clipping mechanisms are employed, while graph neural networks are integrated to model user transaction network topologies for enhanced state representation learning. Finally, a cross-time-scale collaborative optimization mechanism is developed through forward simulation and rolling optimization, enabling effective information transfer and value feedback between different temporal decision hierarchies. Experimental validation demonstrates that MT-MAPPO exhibits excellent performance in multi-agent energy system optimization, achieving good convergence speed and training stability. The average reward of MADQN and MADDPG is 5.4% and 16.2% lower than MAPPO.

源语言英语
期刊论文编号127468
期刊Applied Energy
409
DOI
出版状态已出版 - 15 4月 2026

联合国可持续发展目标

此成果有助于实现下列可持续发展目标:

  1. 可持续发展目标 7 - 经济适用的清洁能源
    可持续发展目标 7 经济适用的清洁能源

学术指纹

探究 'Dual-scale optimization of integrated energy systems: a novel MAPPO-based approach for hybrid game equilibrium' 的科研主题。它们共同构成独一无二的学术指纹。

引用此