Pith. sign in

Paper Citation Record · LEDGER

CaRL: Learning Scalable Planning Policies with Simple Rewards

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2504.17838.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.17838 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T15:45:50.027870Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:59:52.270237Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3d9169d7-3c53-4a64-88db-feb5ee5545d1 · inbound

End-to-End Crop Row Navigation via LiDAR-Based Deep Reinforcement Learning cites this paper.

End-to-End Crop Row Navigation via LiDAR-Based Deep Reinforcement Learning CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T15:45:50.027870Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T15:45:50.027870Z digest=sha256:f0b66b6e6665fdc96b2d137b491e225567a39b51158e3c08cd6e1c54e9ad79c5

Observation a6d06704-b62d-4783-bba3-83f3ecb3e0eb · inbound

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer cites this paper.

Zero-Human Demonstration End-to-end Autonomous Driving with Trajectory Scorer CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T07:54:31.437962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:54:31.437962Z digest=sha256:4ff02321ac30d93c923acaa1bcc41b83a0212c9be1ee385b7e4427142cdce7f3

Observation 7f28dc87-b3fd-4a13-b102-b31572796698 · inbound

Goal-Oriented Reactive Simulation for Closed-Loop Trajectory Prediction cites this paper.

Goal-Oriented Reactive Simulation for Closed-Loop Trajectory Prediction CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T00:58:26.219951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T00:55:01.446865Z digest=sha256:9eb927a810d4b158aa4d86d6cf9c39ceea23cd943c99bb0e09e974559b91bbed

Observation ec35f504-c895-4d74-a3db-f860c52472a3 · inbound

Learning Dexterous Grasping from Sparse Taxonomy Guidance cites this paper.

Learning Dexterous Grasping from Sparse Taxonomy Guidance CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:08:00.975998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T17:05:06.300013Z digest=sha256:63bedb4dd5be2f5a87eeb26da0078fc43f5916189f687a21a1954c42f967e65c

Observation 13fe5e58-3f3d-4f22-b7b3-69418b13313e · inbound

On Data Thinning for Model Validation in Small Area Estimation cites this paper.

On Data Thinning for Model Validation in Small Area Estimation CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-13T11:14:37.619203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T11:14:37.619203Z digest=sha256:56b8b53e313113da47ae441b585d872e0f24082c0d31c05ef6397b933a94aee9

Observation 9b6d7338-4464-4b64-8677-b4786c7c01b3 · inbound

Fail2Drive: Benchmarking Closed-Loop Driving Generalization cites this paper.

Fail2Drive: Benchmarking Closed-Loop Driving Generalization CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:06:00.592800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:19:01.704384Z digest=sha256:3088b5d0e8d6cea9e025f0a75e9e1415994d21ba6160134ae15bd8d3a5ced4db

Observation 9d25453d-cff1-442c-88d1-99b20e707ad9 · inbound

Beyond Self-Play and Scale: A Behavior Benchmark for Generalization in Autonomous Driving cites this paper.

Beyond Self-Play and Scale: A Behavior Benchmark for Generalization in Autonomous Driving CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:16:30.361471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:31:25.453515Z digest=sha256:04ccdf00bfe5eee6a2d8f9ebb6b21c1667440a43b6a569112d771eb3cfd333f1

Observation 813d3636-a4bb-4489-ae9d-6ceeb22c4fbb · inbound

MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving cites this paper.

MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:45:01.098432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T04:41:07.402168Z digest=sha256:f508ece1202ae3236b8c2a489f494600d62678c463c7c38320f48a9d0eed5675

Observation e01fa9b1-c60b-4ba6-b6f7-05985bb9b51d · inbound

MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving cites this paper.

MAPLE: Latent Multi-Agent Play for End-to-End Autonomous Driving CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:49:49.990219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T07:48:52.168457Z digest=sha256:1689a55478d17397f49859dc75de4d2a915c468c35b3c033ffc09c635c2b02d4

Observation f7174094-d128-4f74-a8bf-b9a81d72064f · inbound

DriveSafer: End-to-End Autonomous Driving with Safety Guidance cites this paper.

DriveSafer: End-to-End Autonomous Driving with Safety Guidance CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-19T21:47:48.699813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T21:43:25.406524Z digest=sha256:98eef232c4aa1365b57f7d517ccc19b24317fc4d733e131bc613c5effb9cc5df

Observation 29775f65-6292-45d0-95d9-c0df84ef79d6 · inbound

Scaling Self-Play for End-to-End Driving cites this paper.

Scaling Self-Play for End-to-End Driving CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-04T01:29:22.337007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T20:22:03.938518Z digest=sha256:d5f9ec3b82863c17239f9298425a8ea23deb7b61aeffbdf8f7149efb0e6a7445

Observation efe9097f-183a-45b2-9ca0-f806f1d146e0 · inbound

PlanRL: A Trajectory Planning Architecture for Reinforcement Learning-based Driving Experts cites this paper.

PlanRL: A Trajectory Planning Architecture for Reinforcement Learning-based Driving Experts CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:59:52.271864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T04:44:08.377384Z digest=sha256:9195c038b50be17c1cec149c476f0e3ffe1498465842575b25c4132a44d7cdce

Observation e2ac9611-57a9-4f69-8dca-5daa9c4378b3 · inbound

CLEAR: Closed-Loop Reinforcement Learning at Scale for End-to-End Autonomous Driving cites this paper.

CLEAR: Closed-Loop Reinforcement Learning at Scale for End-to-End Autonomous Driving CaRL: Learning Scalable Planning Policies with Simple Rewards

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T06:40:47.686215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T06:40:47.686215Z digest=sha256:49ae1d3cff1b6a856a3f79c4d1118980f09f115d4dc3d4802be48c2e680273b8