arrow
返回

Deep Reinforcement Learning Microgrid Optimization Strategy Considering Priority Flexible Demand Side

delete2022-03-14
delete21
delete
OA
AI
J
Jinsong Sang
孙
孙宏斌 (Hongbin Sun) *
寇磊 封面图
寇磊 (Lei Kou)
DOI:10.3390/s22062256delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
As an efficient way to integrate multiple distributed energy resources (DERs) and the user side, a microgrid is mainly faced with the problems of small-scale volatility, uncertainty, intermittency and demand-side uncertainty of DERs. The traditional microgrid has a single form and cannot meet the flexible energy dispatch between the complex demand side and the microgrid. In response to this problem, the overall environment of wind power, thermostatically controlled loads (TCLs), energy storage systems (ESSs), price-responsive loads and the main grid is proposed. Secondly, the centralized control of the microgrid operation is convenient for the control of the reactive power and voltage of the distributed power supply and the adjustment of the grid frequency. However, there is a problem in that the flexible loads aggregate and generate peaks during the electricity price valley. The existing research takes into account the power constraints of the microgrid and fails to ensure a sufficient supply of electric energy for a single flexible load. This paper considers the response priority of each unit component of TCLs and ESSs on the basis of the overall environment operation of the microgrid so as to ensure the power supply of the flexible load of the microgrid and save the power input cost to the greatest extent. Finally, the simulation optimization of the environment can be expressed as a Markov decision process (MDP) process. It combines two stages of offline and online operations in the training process. The addition of multiple threads with the lack of historical data learning leads to low learning efficiency. The asynchronous advantage actor-critic (Memory A3C, M-A3C) with the experience replay pool memory library is added to solve the data correlation and nonstatic distribution problems during training. The multithreaded working feature of M-A3C can efficiently learn the resource priority allocation on the demand side of the microgrid and improve the flexible scheduling of the demand side of the microgrid, which greatly reduces the input cost. Comparison of the researched cost optimization results with the results obtained with the proximal policy optimization (PPO) algorithm reveals that the proposed algorithm has better performance in terms of convergence and optimization economics.
Keyword:
microgrid
energy storage
flexible load
reinforcement learning
deep learning
energy optimization
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Sensors 封面图
Sensors
IF:
3.5
论文数:
7.2W
被引数:
20.9W

机构

C
Changchun Institute Technology
学者数:
691
论文数: 440
被引数: 12
引用论文

引用论文

Fast Decomposed Energy Flow in Large-Scale Integrated Electricity-Gas-Heat Energy Systems
err2018-10-01
err101
errOAAI
errMassrur, Hamid Reza; Niknam, Taher; Aghaei, Jamshid; Shafie-Khah, Miadreza; Catalao, Joao P. S.
err分享
err收藏
A novel microgrid support management system based on stochastic mixed-integer linear programming
errENERGY
IF9.4
err2021-05-01
err55
PREAI
errGomes, I. L. R.; Melicio, R.; Mendes, V. M. F.
err分享
err收藏
err分享
err收藏
err分享
err收藏
Optimal Operation of Industrial Energy Hubs in Smart Grids
err2015-03-01
err112
PREAI
errPaudyal, Sumit; Canizares, Claudio A.; Bhattacharya, Kankar
err分享
err收藏
err分享
err收藏
学者 查看更多内容