arrow
返回

Distributed Neural Networks Training for Robotic Manipulation With Consensus Algorithm

delete2024-02-01
delete11
PRE
AI
W
Wenxing Liu
H
Hanlin Niu *
I
Inmo Jang
G
Guido Herrmann
J
Joaquín Carrasco
DOI:10.1109/TNNLS.2022.3191021delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
In this article, we propose an algorithm that combines actor-critic-based off-policy method with consensus-based distributed training to deal with multiagent deep reinforcement learning problems. Specifically, convergence analysis of a consensus algorithm for a type of nonlinear system with a Lyapunov method is developed, and we use this result to analyze the convergence properties of the actor training parameters and the critic training parameters in our algorithm. Through the convergence analysis, it can be verified that all agents will converge to the same optimal model as the training time goes to infinity. To validate the implementation of our algorithm, a multiagent training framework is proposed to train each Universal Robot 5 (UR5) robot arm to reach the random target position. Finally, experiments are provided to demonstrate the effectiveness and feasibility of the proposed algorithm.
Keyword:
Training
Reinforcement learning
Convergence
Task analysis
Robot kinematics
Manipulators
Privacy
Consensus
deep reinforcement learning
Lyapunov methods
manipulator
multiagent systems

期刊

IEEE Transactions on Neural Networks and Learning Systems 封面图
IEEE Transactions on Neural Networks and Learning Systems
IF:
8.9
论文数:
7.6K
被引数:
7.2W

机构

U
University of Manchester
学者数:
5.7W
论文数: 5.3W
被引数: 7.4W
引用论文

引用论文

Listeria monocytogenes Adaptation and Growth at Low Temperatures
err2018-01-01
err0
PREAI
errJoshua C. Saldivar; Morgan L. Davis; Michael G. Johnson; Steven C. Ricke
err分享
err收藏
err分享
err收藏
err分享
err收藏
err分享
err收藏
Six-DOF Spacecraft Optimal Trajectory Planning and Real-Time Attitude Control: A Deep Neural Network-Based Approach
err2020-11-01
err113
errOAAI
errChai, Runqi; Tsourdos, Antonios; Savvaris, Al; Chai, Senchun; Xia, Yuanqing; Chen, C. L. Philip
err分享
err收藏
Priming
err2002-01-01
err0
PREAI
errAnthony D. Wagner; Wilma Koutstaal
err分享
err收藏
学者 查看更多内容