arrow
返回

A DDPG-based solution for optimal consensus of continuous-time linear multi-agent systems

delete2023-07-18
delete3
PRE
AI
Y
Ye Li
刘忠信 封面图
刘忠信 (Zhongxin Liu) *
M
Malika Sader
陈
陈增强 (Zengqiang Chen)
DOI:10.1007/s11431-022-2216-9delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Modeling a system in engineering applications is a time-consuming and labor-intensive task, as system parameters may change with temperature, component aging, etc. In this paper, a novel data-driven model-free optimal controller based on deep deterministic policy gradient (DDPG) is proposed to address the problem of continuous-time leader-following multi-agent consensus. To deal with the problem of the dimensional explosion of state space and action space, two different types of neural nets are utilized to fit them instead of the time-consuming state iteration process. With minimal energy consumption, the proposed controller achieves consensus only based on the consensus error and does not require any initial admissible policies. Besides, the controller is self-learning, which means it can achieve optimal control by learning in real time as the system parameters change. Finally, the proofs of convergence and stability, as well as some simulation experiments, are provided to verify the algorithm's effectiveness.
Keyword:
leader-following consensus
optimal control
reinforcement learning
deep deterministic policy gradient (DDPG)

期刊

Science China-Technological Sciences 封面图
Science China-Technological Sciences
IF:
4.9
论文数:
5.0K
被引数:
9.9K

机构

N
nankai university
学者数:
4.8W
论文数: 3.3W
被引数: 74
引用论文

引用论文

Maximizing Convergence Speed for Second Order Consensus in Leaderless Multi-Agent Systems
err2022-02-01
err17
PREAI
errDifilippo, Gianvito; Fanti, Maria Pia; Mangini, Agostino Marcello
err分享
err收藏
err分享
err收藏
A data-efficient goal-directed deep reinforcement learning method for robot visuomotor skill
err2021-10-01
err7
PREAI
errJiang, Rong; Wang, Zhipeng; He, Bin; Zhou, Yanmin; Li, Gang; Zhu, Zhongpan
err分享
err收藏
学者 查看更多内容