Logo Lanfrica
  • Accueil
  • Atlas
  • Analyses
  • Documentation
  • Sign in

© 2026 Lanfrica. Tous droits réservés. Tous les droits d'auteur des ressources affichées sur le site Web Lanfrica appartiennent aux détenteurs de droits d'auteur d'origine, sauf indication contraire explicite.

Adaptive Traffic Signal Control Using Multi-Agent Reinforcement Learning: A Comparison of Control Strategies

Domaine:

mobility

Type de record:

paper
Créateur:
MahBadAbdAbd
Éditeur:
MDP
Hôte:
Urban traffic congestion remains a persistent challenge for conventional fixed-time signal control, particularly under fluctuating and asymmetric demand. Although multi-agent reinforcement learning (MARL) has shown promise for adaptive traffic signal control, previous studies have often focused on isolated intersections, simplified synthetic networks, or deep-learning-based controllers without systematically comparing tabular and deep-value-based multi-agent approaches under equivalent operating conditions. This study addresses this gap by comparing three traffic signal control strategies: fixed-time control, Multi-Agent Tabular Q-Learning, and multi-agent Deep Q-Network control (MADQN). The evaluation was conducted in a microscopic traffic simulation environment using two complementary testbeds: a synthetic two-intersection corridor, which enables controlled analysis of multi-agent coordination, and a real-world digital twin of the 25 January Corridor in Assiut, Egypt, which tests controller robustness under asymmetric geometry and realistic turning movements. The controllers are assessed under low-, medium-, and high-demand scenarios using queue length, cumulative delay, and Time-To-Collision as operational and safety-related indicators. The results show that MARL-based controllers generally outperform fixed-time control, but their relative performance depends on demand intensity and network complexity. MADQN provides stronger generalization in low-demand and queue-dissipation conditions, whereas Tabular Q-Learning remains highly competitive and can achieve superior delay reduction in several medium- and high-demand cases. These findings indicate that deeper MARL architectures are not universally superior; rather, adaptive signal control deployment should match the controller architecture to the operational objective, traffic demand regime, and practical complexity of the target corridor.

Visit

doi.org

Licenses

https://creativecommons.org/licenses/by/4.0/

Similaires

Traffic Signal Control Model on Isolated Intersection Using Reinforcement Learning: A Case Study on Algiers City, AlgeriaSumoGym: a Framework for Performing Traffic Control using Reinforcement LearningFormation Strategy Optimization Using Multi-Agent Reinforcement Learning (MARL)Off-The-Grid Multi-Agent Reinforcement LearningOnabanjomicheal/Adaptive-Traffic-Signal-DQNUniversally Expressive Communication in Multi-Agent Reinforcement Learning

Traffic Signal Control Model on Isolated Intersection Using Reinforcement Learning: A Case Study on Algiers City, Algeria

Traffic jams and congestion in our cities are a major problem because of the huge increase in the nu

SumoGym: a Framework for Performing Traffic Control using Reinforcement Learning

SumoGym: a Framework for Performing Traffic Control using Reinforcement Learning

Poster presented at the Deep Learning Indaba 2022 by Jacobus Martin

Formation Strategy Optimization Using Multi-Agent Reinforcement Learning (MARL)

Formation Strategy Optimization Using Multi-Agent Reinforcement Learning (MARL)

Poster presented at the Deep Learning Indaba 2023 by Abdel Mfougouon Njupoun

Off-The-Grid Multi-Agent Reinforcement Learning

Off-The-Grid Multi-Agent Reinforcement Learning

Poster presented at the Deep Learning Indaba 2022 by Claude Formanek

Onabanjomicheal/Adaptive-Traffic-Signal-DQN

🚦 Adaptive traffic signal control using Deep Q-Networks (DQN) in SUMO — optimized for Lagos, Nigeria

Universally Expressive Communication in Multi-Agent Reinforcement Learning

Universally Expressive Communication in Multi-Agent Reinforcement Learning

Poster presented at the Deep Learning Indaba 2022 by Matthew Morris