arrow
Return

Multi-agent Battery Storage Management using MPC-based Reinforcement Learning

delete2021-08-09
delete4
delete
OA
AI
A
Arash Bahari Kordabad *
W
Wenqi Cai
S
Sébastien Gros
DOI:10.1109/CCTA48906.2021.9659202delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
In this paper, we present the use of Model Predictive Control (MPC) based on Reinforcement Learning (RL) to find the optimal policy for a multi-agent battery storage system. A time-varying prediction of the power price and production-demand uncertainty are considered. We focus on optimizing an economic objective cost while avoiding very low or very high state of charge, which can damage the battery. We consider the bounded power provided by the main grid and the constraints on the power input and state of each agent. A parametrized MPC-scheme is used as a function approximator for the deterministic policy gradient method and RL optimizes the closed-loop performance by updating the parameters. Simulation results demonstrate that the proposed method is able to tackle the constraints and deliver the optimal policy.

Journal

I
IEEE Conference on Control Technology and Applications
IF:
0
Papers:
34
Citations:
0

Organization

No organization information available