Pith. sign in

Paper Citation Record · LEDGER

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections

As of 22 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2506.16336.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.16336 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:14.797853Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact7
  • verified fuzzy11
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 117b6a38-8987-42aa-815d-6ed68ac816d6 · outbound

This paper cites Autonomous Driving at Unsignalized Intersections: A Review of Decision-Making Challenges and Reinforcement Learning-Based Solutions.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Autonomous Driving at Unsignalized Intersections: A Review of Decision-Making Challenges and Reinforcement Learning-Based Solutions

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:16.855488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:09.979709Z digest=sha256:5528e4d25261a67a852f1f601abd4652f24fcf0e0259967bd7236617cc3d45f4

Observation 0ee6d5c9-79b8-4efb-9c2a-ac0ea1a8ea53 · outbound

This paper cites Enhanced intelligent driver model to access the impact of driving strategies on traffic capacity,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Enhanced intelligent driver model to access the impact of driving strategies on traffic capacity,

Reference 2

Resolution
verified exact
raw_fallback, observed 2026-08-06T23:49:16.567981Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:10.045389Z digest=sha256:7f03fbfa6d2988134121a618d34a1d0a94bf63d4a742c314735adf6b0b473aff

Observation ab3dbb70-193c-45c9-bf85-6051c9ce7b9f · outbound

This paper cites Extended safety descriptor measurements for relative safety assessment in mixed road traffic,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Extended safety descriptor measurements for relative safety assessment in mixed road traffic,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:19.466504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:10.208393Z digest=sha256:0a7ce59ca060456e6576fed8b0bf303cfd13c65a67d28d94a71c7ffcd075d423

Observation 26bb3be2-7e4b-497b-84b4-8774b67b1c5e · outbound

This paper cites Analysis of optimal velocity model with explicit delay,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Analysis of optimal velocity model with explicit delay,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:10.359278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:10.359278Z digest=sha256:1ad5f7caf095e2102a0871481b3bcd187a81648030b9611d4df5e1505887ee7d

Observation f4ed6449-325b-4c5a-9c88-29689429fb15 · outbound

This paper cites End to end learning for self-driving cars,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections End to end learning for self-driving cars,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:19.321928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:10.458971Z digest=sha256:0456a546bf8826828771125eb64ed872f77c0cb3ae15cbb5c32b958ede0f2e92

Observation 5f3ccb8d-1c36-4a81-a3fb-29b51c65009f · outbound

This paper cites ChauffeurNet: Learning to Drive by Imitating the Best and Synthesizing the Worst.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections ChauffeurNet: Learning to Drive by Imitating the Best and Synthesizing the Worst

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:10.703645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:10.703645Z digest=sha256:a065cca9c7126ff6de949cfde4f5c77163579997cb73486738584c207b2a2b90

Observation df52118b-0f50-4af1-b9f2-511523201f97 · outbound

This paper cites Navigating Occluded Intersections with Autonomous Vehicles using Deep Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Navigating Occluded Intersections with Autonomous Vehicles using Deep Reinforcement Learning

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:16.254374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:10.825859Z digest=sha256:828646756fd99279d2b6aa751c9452a3d3876d6360cf5d90fa6e90fdd0c72711

Observation 41825fd5-c11b-4f55-b155-c0c9ee38ffed · outbound

This paper cites Learning Negotiating Behavior Between Cars in Intersections using Deep Q-Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning Negotiating Behavior Between Cars in Intersections using Deep Q-Learning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:16.032798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:10.964587Z digest=sha256:4d61d9d77a5b9696b7b55627765599194323b6ce55e8250943351213be255939

Observation 24ae827a-277c-43c0-9b83-ff267a4b87f0 · outbound

This paper cites Social Attention for Autonomous Decision-Making in Dense Traffic.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Social Attention for Autonomous Decision-Making in Dense Traffic

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:11.091194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:11.091194Z digest=sha256:f6324126de1bf2f3c4e49d3b93286f35b115704717948b2c0d1f2f70a46151dc

Observation 634a9400-a604-4c21-ab8e-e6b1a4d0976c · outbound

This paper cites A multi-task reinforcement learning approach for navigating unsignalized intersections,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections A multi-task reinforcement learning approach for navigating unsignalized intersections,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:19.189396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:11.215996Z digest=sha256:deb4f1bb0468d8dd1cb8ecff61d0984adf6219cab98c7367a92783a0162dbf69

Observation cf90fdee-691c-4a1c-9e0f-19d1e152bec2 · outbound

This paper cites Pomdp and hierarchical options mdp with continuous actions for autonomous driving at intersections,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Pomdp and hierarchical options mdp with continuous actions for autonomous driving at intersections,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:18.944152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:11.385053Z digest=sha256:6568408a5bcc775ca6229d1502d48a3c3a4fba8baa360703f52fcfffb3a1eb96

Observation 0ba67e87-a373-4f1b-b026-cfd37ff24656 · outbound

This paper cites Behavior Planning at Urban Intersections through Hierarchical Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Behavior Planning at Urban Intersections through Hierarchical Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:15.806518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:11.506469Z digest=sha256:fe5344048cd3826078a1c02aa18793d337371f66a4b42681b1b7b03e9b4b5052

Observation bbac736c-efe4-496e-b0c8-7f9c88380e02 · outbound

This paper cites Action and trajectory planning for urban autonomous driving with hierarchical reinforcement learning,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Action and trajectory planning for urban autonomous driving with hierarchical reinforcement learning,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:11.624122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:11.624122Z digest=sha256:f13cbfe2c1ee5301d56e63c4ec06023e864ebf91dc817c3016d951618ae2a41a

Observation 6bac7e8b-6c20-4214-9cb5-17bd944f04bb · outbound

This paper cites Safe reinforcement learning on autonomous vehicles,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Safe reinforcement learning on autonomous vehicles,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:18.721818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:11.854782Z digest=sha256:674965188dc40da140350d319e3a9622c6f19fd0c4cc49d32509e220be13089c

Observation eb576bdf-531c-4c0f-b9fd-d411cf778a2b · outbound

This paper cites Learning to Navigate Intersections with Unsupervised Driver Trait Inference.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning to Navigate Intersections with Unsupervised Driver Trait Inference

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:15.505005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:11.986404Z digest=sha256:1cc9de71413385726f7845424b4aef65538e87fbbfd810a194ec40bfe9e06c5b

Observation 6bc65589-4427-468b-a705-8815a3094f37 · outbound

This paper cites Human-like decision making at unsignalized intersections using social value ori- entation,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Human-like decision making at unsignalized intersections using social value ori- entation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:18.415770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:12.093783Z digest=sha256:f197f0a79f73517be4080744c259cddd5f2526d2b888df2830e1a2c50bba88af

Observation 91f99a7a-e874-4936-9487-37a3581c1db0 · outbound

This paper cites Interactive autonomous navigation with internal state inference and interactivity estimation,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Interactive autonomous navigation with internal state inference and interactivity estimation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:18.127625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:12.243166Z digest=sha256:30ea1517f470b47796fbd048221ef84e696176e689f7e8072f117aa9c25a03c3

Observation d8417676-b812-4f2e-ab62-313a174f5141 · outbound

This paper cites Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Between mdps and semi-mdps: A framework for temporal abstraction in reinforcement learning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:17.849826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:12.391343Z digest=sha256:68cf3c0f585e26b5a12847275d8fbfc0861758219811fbbd63fb5fc4ce75075e

Observation bcaaddc6-da14-4cf4-9159-1a266599e13a · outbound

This paper cites Feudal reinforcement learning,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Feudal reinforcement learning,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:12.500305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:12.500305Z digest=sha256:32c6df9f74bf5a99cc2ca9f982cf17fdaa302ad835c5a8cf74c3ac18f110a578

Observation b1b5c3f8-d11b-48c9-84c0-67c1bff50f6f · outbound

This paper cites Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Hierarchical Deep Reinforcement Learning: Integrating Temporal Abstraction and Intrinsic Motivation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:12.644637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:12.644637Z digest=sha256:8d14e2f45778de2ae832e3da2ea5ee2c79b0024afc659091aca70ce477e6069a

Observation 1a17800a-5966-4d80-8d93-b589c280009d · outbound

This paper cites Data-Efficient Hierarchical Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Data-Efficient Hierarchical Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:12.786912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:12.786912Z digest=sha256:a378700653b236a544992363090fb4fd1d1220614be530e95f40c06b63b9e24e

Observation 77b7febc-4a27-4da2-bba3-9c65c30fe74c · outbound

This paper cites Learning Multi-Level Hierarchies with Hindsight.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning Multi-Level Hierarchies with Hindsight

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:12.947806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:12.947806Z digest=sha256:d976f44d7470ad82b175acd4dff1d311b7d61b0d0f7585b9539974e58fcd7051

Observation a06babb2-978f-4762-b0cd-9c6a32bb8c83 · outbound

This paper cites Multipath: Multiple probabilistic anchor trajectory hypotheses for behavior prediction,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Multipath: Multiple probabilistic anchor trajectory hypotheses for behavior prediction,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:17.642287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:13.093012Z digest=sha256:cd884b9bd443cca95afab00208eaa7ecb7a401cc6da19123e69347e580179ee3

Observation 6461e5fb-e054-4672-83cf-106621710998 · outbound

This paper cites Learning Lane Graph Representations for Motion Forecasting.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning Lane Graph Representations for Motion Forecasting

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-06T23:49:15.123765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:13.400237Z digest=sha256:fe81859f123f90d33b80eb7ad1426322d1deaf3a17e30d559245b3bbb296a5ed

Observation 42429821-ae29-4de1-b089-6cf453b50753 · outbound

This paper cites TNT: Target-driveN Trajectory Prediction.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections TNT: Target-driveN Trajectory Prediction

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.523673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.523673Z digest=sha256:7f6c3fa077fff190392a34c79f92b831df5b3ebaf6c4ac91f5df7c324ded1623

Observation c66819fc-126f-4661-a848-982f4ad85d6f · outbound

This paper cites an unresolved cited work.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.652077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.652077Z digest=sha256:bbb217226e5afc53f359548b5b5e043d2071f9673013042073dfb36cb5e3c015

Observation 45fc5354-722d-43df-b207-f9136d67e99a · outbound

This paper cites Learning Interaction-aware Motion Prediction Model for Decision-making in Autonomous Driving.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning Interaction-aware Motion Prediction Model for Decision-making in Autonomous Driving

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.760119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.760119Z digest=sha256:a70c1a7760a55fce9077dd6e4f3911160c91a4e73269711c475fb002d0131711

Observation a999beb6-da79-46da-aaed-6ccd82a66ce4 · outbound

This paper cites Scene Transformer: A unified architecture for predicting multiple agent trajectories.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Scene Transformer: A unified architecture for predicting multiple agent trajectories

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.842331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.842331Z digest=sha256:38b34d5e80836c7775452ca150f792c3e8521388930bc9bef951f63031796c51

Observation b785a379-6c96-4f74-bb76-b26ffe255d2d · outbound

This paper cites VectorNet: Encoding HD Maps and Agent Dynamics from Vectorized Representation.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections VectorNet: Encoding HD Maps and Agent Dynamics from Vectorized Representation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.943431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.943431Z digest=sha256:a927d7f6d7928a6bab429b37b78288aeb9410264922ac64e143da1285bf5bba7

Observation e6714b33-faaf-4814-917b-6be7996d76dd · outbound

This paper cites Attention Is All You Need.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Attention Is All You Need

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.075473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.075473Z digest=sha256:ee410d7cb3221dcee95e93778c58d7c7e025a76c33cfbd2ff606ba4745b15dd7

Observation dbe86f36-d6bf-4406-b8f7-077166cb340c · outbound

This paper cites Separating axis theorem for oriented bounding boxes,.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Separating axis theorem for oriented bounding boxes,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:17.396641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:14.202926Z digest=sha256:729437a40f7e6efae80998492d1a7e5a5255cbdc3c33c3f19e67104fa54b5782

Observation 1a012cb7-39e4-4159-b9e1-ff95033a6cce · outbound

This paper cites Proximal Policy Optimization Algorithms.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Proximal Policy Optimization Algorithms

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.419292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.419292Z digest=sha256:a5f2045a8b9c2a38a2f2669417353422eb119acb1bfc970a8392536115859f69

Observation 21ce2981-ed5c-403d-a2ca-899d11809ae8 · outbound

This paper cites SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.539948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.539948Z digest=sha256:162ca7b029585008ace423783fcba926e250a1cf9430cda9b8722fc88e55d7bb

Observation 14840db5-cc5f-4874-bc3a-3300ae42800d · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Playing Atari with Deep Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.675888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.675888Z digest=sha256:cbee04e73e5eb1b263273a1469df481f688544056bf53d89293d1b32419d8584

Observation 7b6c6c6b-ad19-4b2e-8a59-d6bf083e2595 · outbound

This paper cites Soft Actor-Critic for Discrete Action Settings.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Soft Actor-Critic for Discrete Action Settings

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:14.797853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:14.797853Z digest=sha256:a27d0032a6502151fd05ed928213aba9623ce02278975b555a73a252dd004e12

Observation 8a9c4364-e1e3-4d73-9c04-347d07237822 · outbound

This paper cites Available: https://api.semanticscholar.org/CorpusID: 125587700.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Available: https://api.semanticscholar.org/CorpusID: 125587700

Reference 2009

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T23:49:17.124427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T23:49:14.338878Z digest=sha256:d9efb370cd235d6d94161b373a6333f0036cf21f51a82f77767bd3460f13b2ab

Observation 4b2a6ee0-ded9-4b99-b15c-865d88d30efe · outbound

This paper cites End to End Learning for Self-Driving Cars.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections End to End Learning for Self-Driving Cars

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:10.555549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:10.555549Z digest=sha256:e52686cc2c20f5b244af11af76752585ee47f2ed4ccf7c91c2e5b6bd1261990d

Observation 09da6045-488f-4dcb-ab9f-b449363a2107 · outbound

This paper cites MultiPath: Multiple Probabilistic Anchor Trajectory Hypotheses for Behavior Prediction.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections MultiPath: Multiple Probabilistic Anchor Trajectory Hypotheses for Behavior Prediction

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:13.253143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:13.253143Z digest=sha256:150ddf488552f50b48e2600072bf19e7094333a1f2dfbf5585c4243e82527f47

Observation e0ee5f6a-6b20-4728-b96b-b6eb5f78f3c2 · outbound

This paper cites Action and Trajectory Planning for Urban Autonomous Driving with Hierarchical Reinforcement Learning.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Action and Trajectory Planning for Urban Autonomous Driving with Hierarchical Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:11.735202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:11.735202Z digest=sha256:a60433d61c1d971ab30ec3182487aeb765b035f3983e2ded19126b61dbecc780

Pith citing papers

No inbound Pith citation observations are available.