arrow
返回

SRL-TR2: A Safe Reinforcement Learning Based TRajectory TRacker Framework

delete2023-06-01
delete3
PRE
AI
C
Chengyu Wang
L
Luhan Wang *
Z
Zhaoming Lu
X
Xinghe Chu
Z
Zhengrui Shi
J
Jiayin Deng
T
Tianyang Su
G
Guochu Shou
X
Xiangming Wen
DOI:10.1109/TITS.2023.3250720delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This paper aims to solve the trajectory tracking con-trol problem for an autonomous vehicle based on reinforcement learning methods. Existing reinforcement learning approaches have found limited successful applications on safety-critical tasks in the real world mainly due to two challenges: 1) sim-to-real transfer; 2) closed-loop stability and safety concern. In this paper, we propose an actor-critic-style framework SRL-TR2, in which the RL-based TRajectory TRackers are trained under the safety constraints and then deployed to a full-size vehicle as the lateral controller. To improve the generalization ability, we adopt a light-weight adapter State and Action Space Alignment (SASA) to establish mapping relations between the simulation and reality. To address the safety concern, we leverage an expert strategy to take over the control when the safety constraints are not satisfied. Hence, we conduct safe explorations during the training process and improve the stability of the policy. The experiments show that our agents can achieve one-shot transfer across simulation scenarios and unseen realistic scenarios, finishing the field tests with average running time less than 10 ms/step and average lateral error less than 0.1 m under the speed ranging from 12 km/h to 18 km/h. A video of the field tests is available at https://youtu.be/pjWcN_fV24g.
Keyword:
Trajectory tracking
safe reinforcement learning
sim-to-real transfer

期刊

IEEE Transactions on Intelligent Transportation Systems 封面图
IEEE Transactions on Intelligent Transportation Systems
IF:
8.4
论文数:
9.7K
被引数:
6.3W

机构

B
beijing university of posts & telecommunications
学者数:
1.4W
论文数: 1.2W
被引数: 9
引用论文

引用论文

err
IF0
err
err0
PREAI
err
err分享
err收藏
Developments in numerical treatments for large data sets of XPS images
err2016-02-25
err0
PREAI
errSolène Béchu; Mireille Richard‐Plouet; Vincent Fernandez; John Walton; Neal Fairley
err分享
err收藏
Nested PID steering control for lane keeping in autonomous vehicles
err2011-12-01
err352
PREAI
errMarino, Riccardo; Scalzi, Stefano; Netto, Mariana
err分享
err收藏
A Reinforcement Learning-Based Adaptive Path Tracking Approach for Autonomous Driving
err2020-10-01
err84
PREAI
errShan, Yunxiao; Zheng, Boli; Chen, Longsheng; Chen, Long; Chen, De
err分享
err收藏
Learning dexterous in-hand manipulation学习灵巧的手操作
err2019-11-18
err818
errOAAI
errAndrychowicz, Marcin; Baker, Bowen; Chociej, Maciek; Jozefowicz, Rafal; McGrew, Bob; Pachocki, Jakub; Petron, Arthur; Plappert, Matthias; Powell, Glenn; Ray, Alex; Schneider, Jonas; Sidor, Szymon; Tobin, Josh; Welinder, Peter; Weng, Lilian; Zaremba, Wojciech
err分享
err收藏
学者 查看更多内容