arrow
返回

Applications in Traffic Signal Control: A Distributed Policy Gradient Decomposition Algorithm

delete2024-02-01
delete0
PRE
AI
P
Pengcheng Dai
W
Wenwu Yu *
H
He Wang
J
Jiahui Jiang
DOI:10.1109/TII.2023.3296887delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
This article explores the application of the multiagent reinforcement learning (MARL) algorithm in addressing the large-scale traffic signal control (TSC) problem. To address the TSC problem in complex urban traffic networks, most existing algorithms focus on optimizing local traffic flow at each intersection through decentralized training based on either the local observations or messages from its neighboring intersections, which, however, lacks the concept of cooperative learning. To conquer such limitations, a novel distributed critic with decentralized actor (DCDA) framework is proposed, which allows the communication messages and temporal difference (TD) losses to be exchanged among neighboring intersections. Specially, by considering the traffic network as a communication network of agents (more precisely, intersections and lanes are considered as agents and edges, respectively), a distributed global average TD loss estimation algorithm is designed in the distributed critic step to estimate the global average TD loss estimation and enhance collaboration among agents. Moreover, in the decentralized actor step, the policy gradient decomposition method is adopted for each agents to learn its local policy solely based on its local action-value function. By adhering to the DCDA framework, a novel distributed policy gradient decomposition (DPGD) algorithm is further proposed to address the TSC problem. Empirical experiments demonstrate that the efficiency, robustness, and stability of the DPGD algorithm outperform the state-of-the-art MARL algorithms in both the environments of cooperative adaptive cruise control and adaptive traffic signal control.
Keyword:
Vehicle dynamics
Estimation
Roads
Informatics
Electronic mail
Training
Scalability
Distributed critic with decentralized actor (DCDA)
distributed policy gradient decomposition (DPGD)
multiagent reinforcement learning (MARL)
traffic signal control (TSC)

期刊

IEEE Transactions on Industrial Informatics 封面图
IEEE Transactions on Industrial Informatics
IF:
9.9
论文数:
8.6K
被引数:
6.0W

机构

S
southeast university - china
学者数:
5.3W
论文数: 4.9W
被引数: 57
引用论文

引用论文

Akt1 in murine chondrocytes controls cartilage calcification during endochondral ossification under physiologic and pathologic conditions
err2010-02-25
err0
PREAI
errAtsushi Fukai; Naohiro Kawamura; Taku Saito; Yasushi Oshima; Toshiyuki Ikeda; Fumitaka Kugimiya; Akiro Higashikawa; Fumiko Yano; Naoshi Ogata; Kozo Nakamura; Ung‐Il Chung; Hiroshi Kawaguchi
err分享
err收藏
Transcriptome Analysis of NPFR Neurons Reveals a Connection Between Proteome Diversity and Social Behavior
err2021-03-31
err0
errOAAI
errJulia Ryvkin; Assa Bentzur; Anat Shmueli; Miriam Tannenbaum; Omri Shallom; Shiran Dokarker; Jennifer I. C. Benichou; Mali Levi; Galit Shohat-Ophir
err分享
err收藏
err分享
err收藏
学者 查看更多内容