arrow
返回

Distributed Non-Communicating Multi-Robot Collision Avoidance via Map-Based Deep Reinforcement Learning

delete2020-08-27
delete24
delete
OA
AI
G
Guangda Chen
S
Shunyi Yao
J
Jun Ma
L
Lifan Pan
陈
陈垣 (Yuan Chen)
P
Pei Xu
J
Jianmin Ji *
陈孝平 封面图
陈孝平 (Xiaoping Chen)
DOI:10.3390/s20174836delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
It is challenging to avoid obstacles safely and efficiently for multiple robots of different shapes in distributed and communication-free scenarios, where robots do not communicate with each other and only sense other robots' positions and obstacles around them. Most existing multi-robot collision avoidance systems either require communication between robots or require expensive movement data of other robots, like velocities, accelerations and paths. In this paper, we propose a map-based deep reinforcement learning approach for multi-robot collision avoidance in a distributed and communication-free environment. We use the egocentric local grid map of a robot to represent the environmental information around it including its shape and observable appearances of other robots and obstacles, which can be easily generated by using multiple sensors or sensor fusion. Then we apply the distributed proximal policy optimization (DPPO) algorithm to train a convolutional neural network that directly maps three frames of egocentric local grid maps and the robot's relative local goal positions into low-level robot control commands. Compared to other methods, the map-based approach is more robust to noisy sensor data, does not require robots' movement data and considers sizes and shapes of related robots, which make it to be more efficient and easier to be deployed to real robots. We first train the neural network in a specified simulator of multiple mobile robots using DPPO, where a multi-stage curriculum learning strategy for multiple scenarios is used to improve the performance. Then we deploy the trained model to real robots to perform collision avoidance in their navigation without tedious parameter tuning. We evaluate the approach with multiple scenarios both in the simulator and on four differential-drive mobile robots in the real world. Both qualitative and quantitative experiments show that our approach is efficient and outperforms existing DRL-based approaches in many indicators. We also conduct ablation studies showing the positive effects of using egocentric grid maps and multi-stage curriculum learning.
Keyword:
multi-robot navigation
distributed collision avoidance
deep reinforcement learning
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Sensors 封面图
Sensors
IF:
3.5
论文数:
7.2W
被引数:
20.9W

机构

U
university of science & technology of china, cas
学者数:
3.2W
论文数: 2.7W
被引数: 74
C
chinese academy of sciences
学者数:
56.7W
论文数: 45.0W
被引数: 704
引用论文

引用论文

S1-Leitlinie Post-COVID/Long-COVIDS1-Leitlinie后COVID/Long-COVID
err2021-09-02
err0
errOAAI
errAndreas Rembert Koczulla; Tobias Ankermann; Uta Behrends; Peter Berlit; Sebastian Böing; Folke Brinkmann; Christian Franke; Rainer Glöckl; Christian Gogoll; Thomas Hummel; Juliane Kronsbein; Thomas Maibaum; Eva M. J. Peters; Michael Pfeifer; Thomas Platz; Matthias Pletz; Georg Pongratz; Frank Powitz; Klaus F. Rabe; Carmen Scheibenbogen; Andreas Stallmach; Michael Stegbauer; Hans Otto Wagner; Christiane Waller; Hubert Wirtz; Andreas Zeiher; Ralf Harun Zwick
err分享
err收藏
A mechanism for scheduling multi robot intelligent warehouse system face with dynamic demand
err2018-12-14
err60
PREAI
errLi, Zhi; Barenji, Ali Vatankhah; Jiang, Jiazhi; Zhong, Ray Y.; Xu, Gangyan
err分享
err收藏
Mo(Co)6Induced Cleavage of Oximes
err2006-08-22
err0
PREAI
errFlorence Geneste; Nadia Racelma; Alec Moradpour
err分享
err收藏
Impact of Race on Prostate-Specific Antigen Outcome After Radical Prostatectomy for Clinically Localized Adenocarcinoma of the Prostate
err2002-06-15
err0
PREAI
errChaundre K. Cross; Delray Shultz; S. Bruce Malkowicz; William C. Huang; Richard Whittington; John E. Tomaszewski; Andrew A. Renshaw; Jerome P. Richie; Anthony V. D’Amico
err分享
err收藏
The Hybrid Reciprocal Velocity Obstacle混合往复速度障碍物
err2011-08-01
err306
errOAAI
errSnape, Jamie; van den Berg, Jur; Guy, Stephen J.; Manocha, Dinesh
err分享
err收藏
学者 查看更多内容