Pith. sign in

REVIEW 1 cited by

Using a Deep Reinforcement Learning Agent for Traffic Signal Control

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1611.01142 v1 pith:UBH2YA4U submitted 2016-11-03 cs.LG cs.SYeess.SY

Using a Deep Reinforcement Learning Agent for Traffic Signal Control

classification cs.LG cs.SYeess.SY
keywords trafficagentcontrolsignalaveragedeepstatesystems
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Ensuring transportation systems are efficient is a priority for modern society. Technological advances have made it possible for transportation systems to collect large volumes of varied data on an unprecedented scale. We propose a traffic signal control system which takes advantage of this new, high quality data, with minimal abstraction compared to other proposed systems. We apply modern deep reinforcement learning methods to build a truly adaptive traffic signal control agent in the traffic microsimulator SUMO. We propose a new state space, the discrete traffic state encoding, which is information dense. The discrete traffic state encoding is used as input to a deep convolutional neural network, trained using Q-learning with experience replay. Our agent was compared against a one hidden layer neural network traffic signal control agent and reduces average cumulative delay by 82%, average queue length by 66% and average travel time by 20%.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. A Distributionally Robust Multi-agent Reinforcement Learning Framework for Intelligent Intersection Control

    eess.SY 2026-07 conditional novelty 4.0

    An adaptive contextual-bandit worst-case estimator co-trained with MARL traffic controllers cuts worst-case and average queues by large margins on grid and Monaco networks and generalizes zero-shot to unseen demand.