消费者
强化学习
计算机科学
可扩展性
能源市场
马尔可夫决策过程
点对点
多智能体系统
分布式计算
双重拍卖
透明度(行为)
马尔可夫过程
人工智能
计算机安全
可再生能源
微观经济学
工程类
共同价值拍卖
经济
电气工程
统计
数据库
数学
作者
Dawei Qiu,Jianhong Wang,Zihang Dong,Yi Wang,Goran Štrbac
标识
DOI:10.1109/tpwrs.2022.3217922
摘要
With increasing numbers of prosumers employed with multi-energy systems (MES) towards higher energy utilization efficiency, an advanced energy management scheme is becoming increasingly important. The incorporation of MES into the existential energy market holds promise for future power systems. The continuous double auction (CDA) market, in a decentralized manner, makes it ideal for enabling peer-to-peer (P2P) energy trading due to its high transparency and efficiency. However, the CDA market is difficult to model when considering the highly stochastic and dynamic behaviors of market participants. For this reason, we formulate this task as a Decentralized Partially Observed Markov Decision Process and propose a novel multi-agent reinforcement learning method that allows each prosumer agent to stabilize the training performance with mean-field approximation and also to maintain the scalability and privacy with market public information. Case studies constructed on a real-world scenario of 100 prosumers show that our method captures the economic benefits of the P2P energy trading paradigm without violating the prosumers' privacy, and outperforms the state-of-the-art methods in terms of policy performance, scalability, and computational performance.
科研通智能强力驱动
Strongly Powered by AbleSci AI