Pith. sign in

REVIEW 1 cited by

Deep Deterministic Policy Gradient for Urban Traffic Light Control

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1703.09035 v2 pith:3BEU44PW submitted 2017-03-27 cs.NE

Deep Deterministic Policy Gradient for Urban Traffic Light Control

classification cs.NE
keywords trafficdeeplargelightagentavailabledeterministicgradient
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Traffic light timing optimization is still an active line of research despite the wealth of scientific literature on the topic, and the problem remains unsolved for any non-toy scenario. One of the key issues with traffic light optimization is the large scale of the input information that is available for the controlling agent, namely all the traffic data that is continually sampled by the traffic detectors that cover the urban network. This issue has in the past forced researchers to focus on agents that work on localized parts of the traffic network, typically on individual intersections, and to coordinate every individual agent in a multi-agent setup. In order to overcome the large scale of the available state information, we propose to rely on the ability of deep Learning approaches to handle large input spaces, in the form of Deep Deterministic Policy Gradient (DDPG) algorithm. We performed several experiments with a range of models, from the very simple one (one intersection) to the more complex one (a big city section).

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Explainable Reinforcement Learning for Adaptive Traffic Signal Control

    cs.AI 2026-07 conditional novelty 6.0

    Entity embeddings of lanes and phases plus hierarchical attention and action masking yield an explainable PPO traffic-signal controller that matches or beats baselines on delay while producing attention maps aligned w...