arrow
返回

Distributed Multirobot Path Planning Based on MRDWA-MADDPG

delete2023-10-15
delete8
PRE
AI
Q
Qichao Wu
R
Rui Lin *
任
任子武 (Ziwu Ren)
DOI:10.1109/JSEN.2023.3310519delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Multirobot path planning in complex environments is a challenging research area. This article proposes a path planning method for multirobot systems based on distributed multiagent deep reinforcement learning. We propose a multirobot dynamic window approach (MRDWA) in which a central controller facilitates sensor information sharing among robots, enabling locally optimal path planning considering the behavior of other robots. We incorporate the output velocity information into the observation function to form an efficient and low-dimensional state representation. Additionally, we employ the multiagent deep deterministic policy gradient (MADDPG) reinforcement learning algorithm to directly map part of the observation information to motion commands for multiple robots, enabling effective obstacle avoidance strategies. An improved action module is developed by using velocity and angular velocity increments and an action selector to refine the output. Furthermore, we introduce a multirobot reward module utilizing heuristic functions to guide the robots to quickly and efficiently identify feasible paths. We also propose a multirobot dynamic constraint reward function to optimize the multirobot trajectories. The MRDWA-MADDPG algorithm is validated through simulations and real-world experiments, demonstrating its effectiveness in diverse complex multirobot path planning scenarios. Our method outperforms conventional algorithms in terms of success rate, arrival time, and overall decision making in complex scenarios. Moreover, our method has a faster computation speed and a shorter training time, produces smoother trajectories, and is easier to deploy on real robots than other learning-based methods.
Keyword:
Robots
Multi-robot systems
Robot kinematics
Collision avoidance
Robot sensing systems
Training
Heuristic algorithms
Deep reinforcement learning
multirobot dynamic window approach (MRDWA)
multirobot path planning
sensor application
trajectory optimization

期刊

IEEE Sensors Journal 封面图
IEEE Sensors Journal
IF:
4.5
论文数:
2.2W
被引数:
7.3W

机构

S
soochow university - china
学者数:
5.2W
论文数: 3.6W
被引数: 82
引用论文

引用论文

Seafood Enzymes
err2012-04-26
err0
PREAI
errM. K. Nielsen; H. H. Nielsen
err分享
err收藏
Reversion of cardiovascular remodelling in renovascular hypertensive 2K‐1C rats by renin–angiotensin system inhibitors
err2020-08-14
err0
PREAI
errJosé Wilson do Nascimento Corrêa; Cibele Maria Prado; Maria Elena Riul; Alice Valença Araújo; Marcos Antonio Rossi; Lusiane Maria Bendhack
err分享
err收藏
err分享
err收藏
Conflict-based search for optimal multi-agent pathfinding基于冲突的多agent最优寻路搜索
err2015-02-01
err635
PREAI
errSharon, Guni; Stern, Roni; Felner, Ariel; Sturtevant, Nathan R.
err分享
err收藏
err分享
err收藏
Hydration Behaviour of Ecocement in Presence of Metakaolin
err2004-01-01
err0
errOAAI
errNabajyoti SAIKIA; Akira USAMI; Shigeru KATO; Toshinori KOJIMA
err分享
err收藏
学者 查看更多内容