arrow
Return

Trajectory Generation for Multiprocess Robotic Tasks Based on Nested Dual-Memory Deep Deterministic Policy Gradient

delete2022-12-01
delete8
PRE
AI
F
Fengkang Ying
刘华山 cover
刘华山 (Huashan Liu) *
R
Rongxin Jiang
X
Xin Yin
DOI:10.1109/TMECH.2022.3160605delete
deleteOriginal
deleteOriginal request for help
deleteShare
deleteSave
Abstract

Abstract

En 中文
Though there are extensive works on deep reinforcement learning (DRL) for robotics, sequential trajectory generation for multiprocess robotic tasks based on DRL is yet to be explored. In this article, the multiprocess task is formulated as a Markov decision process, and a nested dual-memory deep deterministic policy gradient algorithm with dynamic criteria is proposed, to generalize the traditional trajectory planning with predefined target point into a trajectory exploration problem aiming at a target area without solving inverse kinematics. First, a dual-memory architecture with local-to-global strategy is introduced to enhance the performance. Second, a novel nested architecture is proposed to generate sequential trajectory segments successively and asynchronously for the multiprocess task. Third, a compound reward system is designed and a weight coefficient matrix is adopted to balance the position control and the orientation control based on Tait-Bryan angles. In addition, a virtual twin system is established to promote the training efficiency, where the trajectory generated in simulation can be directly applied to the real physical platform. Finally, experimental results on both simulated and real-world applications have verified the performance of the proposed approach.
Keywords:
Compound reward system
dual-memory deep deterministic policy gradient
nested architecture
trajectory generation
virtual twin system

Journal

I
IEEE-ASME Transactions on Mechatronics
IF:
7.3
Papers:
5.4K
Citations:
2.4W

Organization

D
Donghua University
Scholars:
2.0W
Papers: 1.4W
Citations: 2.9W