返回
Formula-E race strategy development using distributed policy gradient reinforcement learning
DOI:10.1016/j.knosys.2021.106781.png)
摘要
En 中文
Energy and thermal management is a crucial element in Formula-E race strategy development. In this study, the race-level strategy development is formulated into a Markov decision process (MDP) problem featuring a hybrid-type action space. Deep Deterministic Policy Gradient (DDPG) reinforcement learning is implemented under distributed architecture Ape-X and integrated with the prioritized experience replay and reward shaping techniques to optimize a hybrid-type set of actions of both continuous and discrete components. Soft boundary violation penalties in reward shaping, significantly improves the performance of DDPG and makes it capable of generating faster race finishing solutions. The new proposed method has shown superior performance in comparison to the Monte Carlo Tree Search (MCTS) with policy gradient reinforcement learning, which solves this problem in a fully discrete action space as presented in the literature. The advantages are faster race finishing time and better handling of ambient temperature rise. (C) 2021 Elsevier B.V. All rights reserved.
Keyword:
Energy management
Formula-E race strategy
Deep deterministic policy gradient
Reinforcement leaning
AI总结
对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。
期刊
K
IF:
7.6
论文数:
1.3W
被引数:
4.5W
机构
引用论文
Adherence of sickle erythrocytes to vascular endothelial cells: requirement for both cell membrane changes and plasma factors
Blood
IF0
A mathematical representation of an energy management strategy for hybrid energy storage system in electric vehicle and real time optimization using a genetic algorithm电动汽车混合储能系统能量管理策略的数学表示和遗传算法的实时优化
APPLIED ENERGY
IF11
Trip-oriented stochastic optimal energy management strategy for plug-in hybrid electric bus
ENERGY
IF9.4

