返回
Stochastic Linear Quadratic Optimal Control Problem: A Reinforcement Learning Method
DOI:10.1109/TAC.2022.3181248.png)
摘要
En 中文
This article adopts a reinforcement learning (RL) method to solve infinite horizon continuous-time stochastic linear quadratic problems, where the drift and diffusion terms in the dynamics may depend on both the state and control. Based on the Bellman's dynamic programming principle, we presented an online RL algorithm to attain optimal control with partial system information. This algorithm computes the optimal control, rather than estimates the system coefficients, and solves the related Riccati equation. It only requires local trajectory information, which significantly simplifies the calculation process. We shed light on our theoretical findings using two numerical examples.
Keyword:
Optimal control
Stochastic processes
Heuristic algorithms
Trajectory
Mathematics
Mathematical models
Riccati equations
Linear quadratic (LQ) problem
reinforcement learning (RL)
stochastic optimal control
期刊
IF:
7
论文数:
1.3W
被引数:
6.7W
机构
引用论文
Adaptive optimal control for continuous-time linear systems based on policy iteration基于策略迭代的连续时间线性系统自适应最优控制
AUTOMATICA
IF5.9
Structure of the C-Terminal Domain of Human La Protein Reveals a Novel RNA Recognition Motif Coupled to a Helical Nuclear Retention Element
Structure
IF0
Reinforcement learning for a class of continuous-time input constrained optimal control problems一类连续时间输入约束最优控制问题的强化学习
AUTOMATICA
IF5.9

