Skip to content
Preprint

SelectLight: Learning to Select Signal Plans Generated by Distributed Model Predictive Control for Urban Traffic Networks

Aug 2026 · 0 citations · 68 references
Engineering Computer Science

TL;DR

SelectLight is proposed, which implements post-optimization selection by allowing a multi-agent reinforcement learning (MARL) policy to choose directly from plans generated online by DMPC, and achieves the best delay-related performance and that its advantage widens with demand.

Abstract

Coordinated traffic signal control across urban networks must adapt to changing demand while satisfying operational constraints. Multi-objective distributed model predictive control (DMPC) can construct feasible signal plans online, but prescribed rules for selecting among trade-off solutions cannot learn from realized closed-loop outcomes. We propose SelectLight, which implements post-optimization selection by allowing a multi-agent reinforcement learning (MARL) policy to choose directly from plans generated online by DMPC. At each control update, state-pruned multi-objective dynamic programming (SP-MODP) evaluates plans with a Newellian point--spatial queue model and returns a bounded set of mutually nondominated candidate signal plans for total queueing delay, peak queue accumulation, and total number of stops. A topology-aware attention policy trained with independent proximal policy optimization (IPPO) selects one unmodified plan from each variable-size set. This confines learning to candidate selection, preserves the prescribed signal timing constraints, and leaves the selected plan and its predicted objective trade-offs available for inspection. Experiments on two 28-intersection SUMO networks show that SelectLight achieves the best delay-related performance and that its advantage widens with demand. At twice the baseline demand, it reduces queueing delay and waiting time by 5.57% and 6.44%, respectively, relative to the strongest baseline. SelectLight also incurs the lowest transfer loss under every tested demand shift. With a 120 s prediction horizon, the per-intersection 99th-percentile SP-MODP solution time is 5.408 ms, well below the 5 s control interval.

View source

Similar papers

Open access Sep 2026

Hybrid-RL-RB: A Constraint-Aware Reinforcement Learning and Rule-Based Algorithm for Multi-Intersection Traffic Signal Control

Traffic signal control plays a critical role in mitigating congestion and improving urban mobility, particularly in multi-intersection networks where fixed-time strategies cannot adapt to fluctuating demand. Although reinforcement learning has shown strong potential for adaptive signal optimization, purely learning-bas...

Mohammed El Kaim Billah, Mohammed-Alamine El Houssaini, Abedelfettah Mabrouk et al. · 0 citations
Open access Aug 2026

Deep reinforcement learning-based traffic signal control in multi-intersection environments: a comparative study of DQN variants

The findings demonstrate the potential of DRL-based traffic signal control in controlled simulation conditions and highlight that algorithm performance is strongly influenced by traffic policy design and environmental complexity.

D. Prastiyanto, A. A. Manaf, Muhammad Ahnaf Maulana et al. · 0 citations

Machine Learning based Distributed Traffic Signal Control

This study proposes a distributed traffic signal control framework built upon a Machine Learning (ML) paradigm utilizing Reinforcement Learning (RL), and demonstrates the effectiveness of the proposed approach, with vehicle queue lengths and average waiting times reduced by 35% on roads leading to the junctions, compar...

Alireza Rezaee, Amirhossein Safdari · 1 citation
Open access Sep 2026

Design management for smart urban traffic: an IoV-Based multi-agent AI framework for adaptive signal control

An adaptive multi-agent traffic-management framework that responds to real-time urban traffic conditions using Internet of Vehicles traffic-state information and artificial intelligence provides an adaptive AI-based approach for coordinated traffic-signal control and supports traffic-efficiency-related sustainability o...

Yuan-Yuan Fei, Yan Zhang · 0 citations
Open access Aug 2026

Deep reinforcement learning-based adaptive traffic signal control in urban networks

Results demonstrate the usefulness of deep reinforcement learning in the creation of intelligent and adaptive traffic signal control services in the city and prove the usefulness of computational feasibility and robustness even when the size of the network grows.

Manisha Aeri, K. Purohit, Lata Nautiyal et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.