Pith. sign in

Paper Citation Record · LEDGER

Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2202.10341.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2202.10341 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:40:41.198683Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T19:56:10.416267Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 904a7ced-b784-4e91-9685-111cab13331b · inbound

Automated Driving with Evolution Capability: A Reinforcement Learning Method with Monotonic Performance Enhancement cites this paper.

Automated Driving with Evolution Capability: A Reinforcement Learning Method with Monotonic Performance Enhancement Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-11T15:40:41.198683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:40:41.198683Z digest=sha256:c36061bb36c0a6046f223705838a35172b04be8c02fbd22b10e315f761825c41

Observation e77de19a-57a6-40fa-875c-cdf0a94b45a5 · inbound

Joint Decision-Making in Robot Teleoperation: When are Two Heads Better Than One? cites this paper.

Joint Decision-Making in Robot Teleoperation: When are Two Heads Better Than One? Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T10:47:59.636135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T10:47:59.636135Z digest=sha256:b4a47049f891f087896697576fc4bca8edb5a23321b1caf0ab1990de1308f51a

Observation db5b28be-2882-4cfd-992e-ae5ffe0228cc · inbound

Bounded Autonomy for Enterprise AI: Typed Action Contracts and Consumer-Side Execution cites this paper.

Bounded Autonomy for Enterprise AI: Typed Action Contracts and Consumer-Side Execution Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:20:11.007289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T11:15:55.974291Z digest=sha256:fed01e18bc81343205c15e3b3347f2c996468ce832d38e1b80cef70639b38fdf

Observation 3dd501a8-d0e7-4bd4-aca7-29c9edb38a61 · inbound

OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation cites this paper.

OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-20T17:38:48.188246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T17:38:15.672066Z digest=sha256:ebaa085ba178414b8d4837947f8bc77392ac3fc245b57379e4cb16d4f281aa4d

Observation c9d30eed-3ef8-44a3-9feb-4865b1296467 · inbound

OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation cites this paper.

OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-06-30T19:45:00.851580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T19:40:19.827541Z digest=sha256:b24f869a38278f2dc46bbb97489ccdccedc34db6d77c56b1d84972cb96c253ac

Observation b2c245e3-e9de-43e1-a95d-67ed23389d15 · inbound

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation cites this paper.

DeMaVLA: A Vision-Language-Action Foundation Model for Generalizable Deformable Manipulation Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.417816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T21:58:48.611343Z digest=sha256:34a82feabdedd18f7a03170b2cce41e46b0c0e848836a61d7553cae125d3f017

Observation 6c394f58-9fb9-4ee9-917a-09f6137ae47d · inbound

WARP-RM: A Warp-Augmented Relative Progress Reward Model for Data Curation cites this paper.

WARP-RM: A Warp-Augmented Relative Progress Reward Model for Data Curation Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:15:51.844603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T03:55:47.911210Z digest=sha256:d11aaf767db8c76be7e2f14ae25bd8b8d086ac5608d883b8d4bb42ed2887e001

Observation 77653c1c-9148-4132-b55d-c3e2a8f08b70 · inbound

WARP-RM: A Warp-Augmented Relative Progress Reward Model for Data Curation cites this paper.

WARP-RM: A Warp-Augmented Relative Progress Reward Model for Data Curation Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T11:25:18.198611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:25:18.198611Z digest=sha256:9d4eaee7b4634cc18bd442349459494551c12a5e5c36b66015e43b53e95b2a62

Observation e991a0d1-5ca7-40ae-af7c-1bcacc5062ea · inbound

WARP-RM: A Warp-Augmented Relative Progress Reward Model for Data Curation cites this paper.

WARP-RM: A Warp-Augmented Relative Progress Reward Model for Data Curation Efficient Learning of Safe Driving Policy via Human-AI Copilot Optimization

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-13T07:17:21.581451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T07:17:21.581451Z digest=sha256:2d33bbfe0e47d5ffbc4dd0acc0a53f2ebe5046d8ffb4b27f818f19ea2d7af123