REVIEW 4 cited by
Deep Deterministic Policy Gradient for Urban Traffic Light Control
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Traffic light timing optimization is still an active line of research despite the wealth of scientific literature on the topic, and the problem remains unsolved for any non-toy scenario. One of the key issues with traffic light optimization is the large scale of the input information that is available for the controlling agent, namely all the traffic data that is continually sampled by the traffic detectors that cover the urban network. This issue has in the past forced researchers to focus on agents that work on localized parts of the traffic network, typically on individual intersections, and to coordinate every individual agent in a multi-agent setup. In order to overcome the large scale of the available state information, we propose to rely on the ability of deep Learning approaches to handle large input spaces, in the form of Deep Deterministic Policy Gradient (DDPG) algorithm. We performed several experiments with a range of models, from the very simple one (one intersection) to the more complex one (a big city section).
Forward citations
Cited by 4 Pith papers
-
Explainable Reinforcement Learning for Adaptive Traffic Signal Control
Entity embeddings of lanes and phases plus hierarchical attention and action masking yield an explainable PPO traffic-signal controller that matches or beats baselines on delay while producing attention maps aligned w...
-
Integrating Transit Signal Priority into Multi-Agent Reinforcement Learning based Traffic Signal Control
Coordinated multi-agent reinforcement learning transit signal priority reduces simulated bus travel time by 27% across two intersections, outperforming independent agents (22%) while keeping side street delay increases small.
-
Large-Scale Traffic Signal Control Using a Novel Multi-Agent Reinforcement Learning
Co-DQL, a combination of double Q-learning, UCB exploration, mean field modeling, and local reward/state sharing, reduces simulated traffic delays relative to several MARL baselines.
-
An Open-Source Framework for Adaptive Traffic Signal Control
An open-source SUMO framework for adaptive traffic signal control is introduced, and experiments on a two-intersection network show Max-pressure outperforms deep reinforcement learning controllers.
Discussion (0). Continue with ORCID to comment.