arXiv AI By Dickens Kwesiga, Nishu Choudhary, Angshuman Guin, Michael Hunter

Explainable Reinforcement Learning for Adaptive Traffic Signal Control

Read the original on arXiv AI →

arXiv:2607. 03703v1 Announce Type: new Abstract: Reinforcement Learning (RL) has emerged as a powerful paradigm for adaptive traffic signal control.

Machine-generated by The Flow from the publisher's headline and feed description — not written or checked by a human. The full article lives at arXiv AI.

arXiv AI
Jun 29

OverFlowLight: Real-Time Gridlock Prevention and Traffic Signal Optimization for Urban Intersections

arXiv:2606. 27381v1 Announce Type: cross Abstract: Queue overflow, a severe consequence of urban traffic congestion, occurs when vehicle queues exceed intersection capacity, obstructing upstream traffic and triggering cascading gridlocks.

By Mingyuan Li, Boyang Huang, Tianqi Jiang, Chenpu Li, Chunyu Liu, Yang Li, Ruimin Li, Qiang Wu
arXiv Machine Learning
Aug 20

SIGMA: Symmetry-aware, Intelligent, Geometric, Multi-objective Adaptive Control for Robust, Dependable Traffic Management

SIGMA is a reinforcement‑learning framework for traffic signal control that incorporates a large language model to adaptively tune multiple objectives based on natural‑language emergency commands. It uses rotational data augmentation to learn orientation‑invariant policies and an offline‑to‑online training pipeline to ensure stable deployment. Experiments in SUMO on four Kolkata intersections show that SIGMA reduces waiting times, queue lengths, and improves throughput compared to fixed‑time, actuated, and DQN baselines, with ablation studies confirming robustness to component failures and geometric rotations.

By Pratham Payra, Jagadish B, Tanmay Sen, Tanujit Chakraborty
arXiv Machine Learning
Aug 13

Language-Structured Relational Q-Learning for Threat-Aware Control in Safety-Critical Driving

arXiv:2608. 11498v1 Announce Type: cross Abstract: Natural-language-based scenario generation offers an intuitive means of describing rare and complex driving interactions, yet it is still uncertain whether training with language-structured data leads to truly adaptive control policies.

By Aditya Humnabadkar, Huaizhong Zhang, Ardhendu Behera