arrow
返回

Max-plus approximation for reinforcement learning

delete2021-07-01
delete4
PRE
AI
V
Vinícius Mariano Gonçalves *
DOI:10.1016/j.automatica.2021.109623delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Max-Plus Algebra has been applied in several contexts, especially in the control of discrete events systems. In this article, we discuss another application closely related to control: the use of Max-Plus algebra concepts in the context of reinforcement learning. Max-Plus Algebra and reinforcement learning are strongly linked due to the latter's dependence on the Bellman Equation which, in some cases, is a linear Max-Plus equation. This fact motivates the application of Max-Plus algebra to approximate the value function, central to the Bellman Equation and thus also to reinforcement learning. This article proposes conditions so that this approach can be done in a simple way and following the philosophy of reinforcement learning: explore the environment, receive the rewards and use this information to improve the knowledge of the value function. The proposed conditions are related to two matrices and impose on them a relationship that is analogous to the concept of weak inverses in traditional algebra. (C) 2021 Elsevier Ltd. All rights reserved.
Keyword:
Max-Plus Algebra
Reinforcement learning
Dynamic programming
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Automatica 封面图
Automatica
IF:
5.9
论文数:
1.2W
被引数:
5.2W

机构

U
Universidade Federal de Minas Gerais
学者数:
2.5W
论文数: 1.5W
被引数: 1.4W
引用论文

引用论文

Optimizing Chemical Reactions with Deep Reinforcement Learning
err2017-12-15
err327
errOAAI
errZhou, Zhenpeng; Li, Xiaocheng; Zare, Richard N.
err分享
err收藏
On max-plus linear dynamical system theory: The regulation problem关于max-plus线性动力系统理论: 调节问题
err2017-01-01
err19
errOAAI
errGoncalves, Vinicius Mariano; Maia, Carlos Andrey; Hardouin, Laurent
err分享
err收藏
没有更多内容