返回
SRL-TR2: A Safe Reinforcement Learning Based TRajectory TRacker Framework
DOI:10.1109/TITS.2023.3250720.png)
摘要
En 中文
This paper aims to solve the trajectory tracking con-trol problem for an autonomous vehicle based on reinforcement learning methods. Existing reinforcement learning approaches have found limited successful applications on safety-critical tasks in the real world mainly due to two challenges: 1) sim-to-real transfer; 2) closed-loop stability and safety concern. In this paper, we propose an actor-critic-style framework SRL-TR2, in which the RL-based TRajectory TRackers are trained under the safety constraints and then deployed to a full-size vehicle as the lateral controller. To improve the generalization ability, we adopt a light-weight adapter State and Action Space Alignment (SASA) to establish mapping relations between the simulation and reality. To address the safety concern, we leverage an expert strategy to take over the control when the safety constraints are not satisfied. Hence, we conduct safe explorations during the training process and improve the stability of the policy. The experiments show that our agents can achieve one-shot transfer across simulation scenarios and unseen realistic scenarios, finishing the field tests with average running time less than 10 ms/step and average lateral error less than 0.1 m under the speed ranging from 12 km/h to 18 km/h. A video of the field tests is available at https://youtu.be/pjWcN_fV24g.
Keyword:
Trajectory tracking
safe reinforcement learning
sim-to-real transfer
期刊
IF:
8.4
论文数:
9.7K
被引数:
6.3W
机构
引用论文
Study of bi-directional buck-boost converter topologies for application in electrical vehicle motor drives应用于电动汽车电机驱动的双向buck-boost变换器拓扑研究
Multi-Agent Deep Reinforcement Learning for Large-Scale Traffic Signal Control面向大规模交通信号控制的多智能体深度强化学习
A Graph Embedding Framework for Maximum Mean Discrepancy-Based Domain Adaptation Algorithms基于最大均值差异的域自适应算法的图嵌入框架

