返回
Reinforcement Learning Control With Knowledge Shaping
DOI:10.1109/TNNLS.2023.3243631.png)
摘要
En 中文
We aim at creating a transfer reinforcement learning framework that allows the design of learning controllers to leverage prior knowledge extracted from previously learned tasks and previous data to improve the learning performance of new tasks. Toward this goal, we formalize knowledge transfer by expressing knowledge in the value function in our problem construct, which is referred to as reinforcement learning with knowledge shaping (RL-KS). Unlike most transfer learning studies that are empirical in nature, our results include not only simulation verifications but also an analysis of algorithm convergence and solution optimality. Also different from the well-established potential-based reward shaping methods which are built on proofs of policy invariance, our RL-KS approach allows us to advance toward a new theoretical result on positive knowledge transfer. Furthermore, our contributions include two principled ways that cover a range of realization schemes to represent prior knowledge in RL-KS. We provide extensive and systematic evaluations of the proposed RL-KS method. The evaluation environments not only include classical RL benchmark problems but also include a challenging task of real-time control of a robotic lower limb with a human user in the loop.
Keyword:
Task analysis
Knowledge transfer
Reinforcement learning
Transfer learning
Silicon
Knowledge representation
Convergence
Reinforcement learning (RL)
reward shaping
transfer learning
value function
期刊
IF:
8.9
论文数:
7.5K
被引数:
7.2W
机构
引用论文
Study of bi-directional buck-boost converter topologies for application in electrical vehicle motor drives应用于电动汽车电机驱动的双向buck-boost变换器拓扑研究
Finite-Approximation-Error-Based Discrete-Time Iterative Adaptive Dynamic Programming基于有限近似误差的离散时间迭代自适应动态规划
Optimal control of unknown nonaffine nonlinear discrete-time systems based on adaptive dynamic programming基于自适应动态规划的未知非仿射非线性离散系统最优控制
AUTOMATICA
IF5.9

