Pith. sign in

REVIEW 1 cited by

Traffic Light Control Using Deep Policy-Gradient and Value-Function Based Reinforcement Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1704.08883 v2 pith:TIKUHHW4 submitted 2017-04-28 cs.LG

classification cs.LG
keywords controltrafficagentdeeplearningpolicy-gradientreinforcementvalue-function
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Recent advances in combining deep neural network architectures with reinforcement learning techniques have shown promising potential results in solving complex control problems with high dimensional state and action spaces. Inspired by these successes, in this paper, we build two kinds of reinforcement learning algorithms: deep policy-gradient and value-function based agents which can predict the best possible traffic signal for a traffic intersection. At each time step, these adaptive traffic light control agents receive a snapshot of the current state of a graphical traffic simulator and produce control signals. The policy-gradient based agent maps its observation directly to the control signal, however the value-function based agent first estimates values for all legal control signals. The agent then selects the optimal control action with the highest value. Our methods show promising results in a traffic network simulated in the SUMO traffic simulator, without suffering from instability issues during the training process.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. An Open-Source Framework for Adaptive Traffic Signal Control

    eess.SY 2019-09 conditional novelty 4.0 of 10

    An open-source SUMO framework for adaptive traffic signal control is introduced, and experiments on a two-intersection network show Max-pressure outperforms deep reinforcement learning controllers.

Pith tools