arrow
返回

Reinforcement learning for control: Performance, stability, and deep approximators

delete2018-01-01
delete277
delete
OA
AI
L
Lucian Buşoniu *
T
Tim de Bruin
D
Domagoj Tolić
J
Jens Kober
I
Ivana Palunko
DOI:10.1016/j.arcontrol.2018.09.005delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Reinforcement learning (RL) offers powerful algorithms to search for optimal controllers of systems with nonlinear, possibly stochastic dynamics that are unknown or highly uncertain. This review mainly covers artificial-intelligence approaches to RL, from the viewpoint of the control engineer. We explain how approximate representations of the solution make RL feasible for problems with continuous states and control actions. Stability is a central concern in control, and we argue that while the control-theoretic RL subfield called adaptive dynamic programming is dedicated to it, stability of RL largely remains an open question. We also cover in detail the case where deep neural networks are used for approximation, leading to the field of deep RL, which has shown great success in recent years. With the control practitioner in mind, we outline opportunities and pitfalls of deep RL; and we close the survey with an outlook that - among other things - points out some avenues for bridging the gap between control and artificial-intelligence RL techniques. (C) 2018 Elsevier Ltd. All rights reserved.
Keyword:
Reinforcement learning
Optimal control
Deep learning
Stability
Function approximation
Adaptive dynamic programming
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Annual Reviews in Control 封面图
Annual Reviews in Control
IF:
10.7
论文数:
829
被引数:
5.9K

机构

D
Delft University of Technology
学者数:
2.6W
论文数: 2.5W
被引数: 3.8W
T
Technical University of Cluj Napoca
学者数:
2.1K
论文数: 1.6K
被引数: 1.2K
U
University of Dubrovnik
学者数:
172
论文数: 149
被引数: 0
学者 查看更多机构
引用论文

引用论文

Carbon dioxide in enhanced oil recovery
err1993-09-01
err0
errOAAI
errMartin Blunt; F.John Fayers; Franklin M. Orr
err分享
err收藏
Mini‐Mental State Examination
err2002-04-30
err0
PREAI
errJoseph R. Cockrell; Marshal F. Folstein
err分享
err收藏
Natural Actor-Critic
err2008-03-01
err643
PREAI
errPeters, Jan; Schaal, Stefan
err分享
err收藏
Structural features of glasses in the GaF3-SnF2 system
err2006-06-01
err0
PREAI
errL. N. Ignat’eva; N. V. Surovtsev; E. B. Merkulov; V. M. Buznik
err分享
err收藏
Opening Platforms: How, When and Why?
err2008-01-01
err0
PREAI
errThomas R. Eisenmann; Geoffrey Parker; Marshall W. Van Alstyne
err分享
err收藏
Quantum Monte Carlo study of the cooperative binding of NO2 to fragment models of carbon nanotubes
err2008-12-01
err0
errOAAI
errJohn W. Lawson; Charles W. Bauschlicher; Julien Toulouse; Claudia Filippi; C.J. Umrigar
err分享
err收藏
学者 查看更多内容