Multi-Agent Deep Reinforcement Learning for Regional Traffic Signal Control Based on Dynamic Weight Decomposition

Conventional traffic signal control methodologies are deficient in adapting to rapid traffic flow variations and capturing the complex dynamic interactions between intersections within regional road networks. In order to address this specific issue, the present study proposed the Qatten (Q-value Attention Network)-TSC algorithm. The algorithm was constructed on the basis of the dynamic weighted value decomposition principle and was built upon the multi-agent QMIX (Q-value Mixed Network) framework. The model employed a multi-head attention mechanism to effectively fuse individual agent Q-values with global states and individual features to compute global Q-values. Furthermore, the model incorporated multidimensional state information to comprehensively characterize complex traffic networks. Extensive experiments were conducted on small- and large-scale SUMO simulation platforms based on the real road network of Yangzhou. The experimental results demonstrated that in comparison to VDN and QMIX, Qatten-TSC attained average reward increments of 26.4% and 3.46%, correspondingly, in small-scale road networks, and 34.12% and 12.81%, correspondingly, in large-scale road networks. Furthermore, in large-scale scenarios, the average time loss was reduced by 19.28% and 7.15%, respectively, while the average speed increased by 3.00% and 0.87%, respectively. In addition, the baseline algorithm (Qatten) is unstable and poor-performing. The dynamic weighting mechanism is robust and effective, even as the road network complexity increases.

Authors

Institutions

Publication Details

Journal
Sensors
Published
2026-09-11
DOI
https://doi.org/10.3390/s26185766
Primary Topic
Traffic control and management
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Multi-Agent Deep Reinforcement Learning for Regional Traffic Signal Control Based on Dynamic Weight Decomposition

Zhenghua Zhang, Peng Shi
Sensors
Traffic control and management
article

Multi-Agent Deep Reinforcement Learning for Regional Traffic Signal Control Based on Dynamic Weight Decomposition

Zhenghua Zhang, Peng Shi
article en

Abstract

Conventional traffic signal control methodologies are deficient in adapting to rapid traffic flow variations and capturing the complex dynamic interactions between intersections within regional road networks. In order to address this specific issue, the present study proposed the Qatten (Q-value Attention Network)-TSC algorithm. The algorithm was constructed on the basis of the dynamic weighted value decomposition principle and was built upon the multi-agent QMIX (Q-value Mixed Network) framework. The model employed a multi-head attention mechanism to effectively fuse individual agent Q-values with global states and individual features to compute global Q-values. Furthermore, the model incorporated multidimensional state information to comprehensively characterize complex traffic networks. Extensive experiments were conducted on small- and large-scale SUMO simulation platforms based on the real road network of Yangzhou. The experimental results demonstrated that in comparison to VDN and QMIX, Qatten-TSC attained average reward increments of 26.4% and 3.46%, correspondingly, in small-scale road networks, and 34.12% and 12.81%, correspondingly, in large-scale road networks. Furthermore, in large-scale scenarios, the average time loss was reduced by 19.28% and 7.15%, respectively, while the average speed increased by 3.00% and 0.87%, respectively. In addition, the baseline algorithm (Qatten) is unstable and poor-performing. The dynamic weighting mechanism is robust and effective, even as the road network complexity increases.

SensorsVol. 26(18)
Yango University (CN), Yangzhou Vocational University (CN), Yangzhou University (CN)
No poverty
Openalex Percentile: Top 15%
Traffic control and management
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.

Multi-Agent Deep Reinforcement Learning for Regional Traffic Signal Control Based on Dynamic Weight Decomposition — Zhenghua Zhang, Peng Shi · Sensors (2026) | TGRS Research Map | TGRS