PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 26, 2026Systems0 citationsOpen Access

Multi-Agent Deep Deterministic Policy Gradient-Based Coordinated Control for Urban Expressway Entrance–Arterial Interfaces

View Full Paper
SWShunchao WangZWZhigang WuWYWenwu Yu

Key Points

  • To develop a cooperative control framework for urban expressway-arterial interfaces using multi-agent reinforcement learning.
  • Developed a multi-agent reinforcement learning framework based on MADDPG.
  • Created an asynchronous control cycle to accommodate different timing needs of traffic controls.
  • Implemented conflict-aware reward design for density regulation and speed harmonization.
  • Delays congestion onset and reduces shockwave propagation.
  • At the mainline merge, average travel time decreased to 13.56 seconds.
  • Ramp occupancy lowered to 6.4%, while signalized approach delay decreased to 85.71 seconds.

Abstract

Coordinated control of ramp metering, variable speed limits, and intersection signals is critical for mitigating congestion and enhancing efficiency at urban expressway–arterial interfaces. Existing strategies often operate in isolation, leading to fragmented responses and limited adaptability under heterogeneous traffic demands. This study develops a multi-agent reinforcement learning framework based on MADDPG to achieve cooperative decision-making across heterogeneous controllers. An asynchronous control cycle mechanism is designed to accommodate different temporal requirements of ramp meters, speed limits, and signal controllers, ensuring practical feasibility in real-time operations. A conflict-aware reward design further embeds density regulation, speed harmonization, and spillback prevention to stabilize flow dynamics. Simulation experiments on a calibrated urban network demonstrate that the proposed framework delays congestion onset, reduces shockwave propagation, and improves throughput compared with classical benchmarks. In particular, at the mainline merge, average travel time is reduced to 13.56 s (62.4% of VSL-only); at the ramp, occupancy is lowered to 6.4% (40.6% of ALINEA); and at the signalized approach, average delay decreases to 85.71 s (62.7% of actuated control). These results highlight the scalability and deployment potential of the proposed cooperative control approach for system-level traffic management in mixed traffic environments.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Wang et al. (2026) studied this question.

synapsesocial.com/papers/699fe3f995ddcd3a253e8206https://doi.org/10.3390/systems14030231
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Combining multi-agent deep deterministic policy gradient and rerouting technique to improve traffic network performance under mixed traffic conditions2024 · 1 citations
  2. 2Deep reinforcement learning-based adaptive traffic signal control in urban networks2026
  3. 3Single‐Agent Reinforcement Learning Model for Adaptive Traffic Signal Control in Urban Corridors2026
  4. 4Transition-Sensitive Congestion Dynamics in Heterogeneous Urban Traffic Networks Under Coordinated Reinforcement Learning2026
  5. 5Signal Control of Urban Expressway Ramp Based on Reinforcement Learning2024 · 1 citations