vix.ing · top · new · best · stats

Optimizing Traffic Lights with Multi-agent Deep Reinforcement Learning and V2X communication

2020/02/23 by Azhar Hussain, Tong Wang, Hussain, Azhar +3
Computer Science · Engineering · #Artificial Intelligence (cs.AI) #Autonomous Vehicle Technology and Safety #FOS: Computer and information sciences #Machine Learning (cs.LG) #Multiagent Systems (cs.MA) #Traffic control and management #Vehicular Ad Hoc Networks (VANETs) #cs.AI #cs.LG #cs.MA

paper · pdf · doi:10.48550/arxiv.2002.09853

7 Figure, Table 1

arxiv created 2020/02/23 · openalex publication_date 2020/02/23 · arxiv updated 2020/02/25 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28

Abstract

We consider a system to optimize duration of traffic signals using multi-agent deep reinforcement learning and Vehicle-to-Everything (V2X) communication. This system aims at analyzing independent and shared rewards for multi-agents to control duration of traffic lights. A learning agent traffic light gets information along its lanes within a circular V2X coverage. The duration cycles of traffic light are modeled as Markov decision Processes. We investigate four variations of reward functions. The first two are unshared-rewards: based on waiting number, and waiting time of vehicles between two cycles of traffic light. The third and fourth functions are: shared-rewards based on waiting cars, and waiting time for all agents. Each agent has a memory for optimization through target network and prioritized experience replay. We evaluate multi-agents through the Simulation of Urban MObility (SUMO) simulator. The results prove effectiveness of the proposed system to optimize traffic signals and reduce average waiting cars to 41.5 % as compared to the traditional periodic traffic control system.

Related