Pith. sign in

Paper Citation Record · LEDGER

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates

As of 9 August 2026, this Paper Citation Record lists 10 of 10 outbound references and 0 inbound Pith citation observations for arXiv:2602.10430.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2602.10430 v2

Coverage vector

measured 10 of 10 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T01:15:34.171120Z

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

10 of 10 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6ad4a4d7-f1d9-4796-8719-14434d4850b0 · outbound

This paper cites Offline Reinforcement Learning with Implicit Q-Learning.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates Offline Reinforcement Learning with Implicit Q-Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.032799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.032799Z digest=sha256:eb1a18294095d9a8fb823033dd587e36af516aebba1b8fab889e82c445595e30

Observation 83bc6591-6b4e-4530-b722-a8094896a2ec · outbound

This paper cites Crafting papers on machine learning.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates Crafting papers on machine learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.053776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.053776Z digest=sha256:e3bd3552be7d0eecef80f1628ccaba31f7e129e6674add669fda17cdc5af917d

Observation 7b3c89b5-be99-4db9-8d38-498586d393a3 · outbound

This paper cites distance.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates distance

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.171120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.171120Z digest=sha256:a0dc96e91dc9dcc2a48eecd8cf64fab944161e922c19e350131c7d38b7d3f5ea

Observation 6497b1d8-f591-47bf-aa17-ca92aa9af79a · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.114225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.114225Z digest=sha256:f92b3ced094fa5565d77702fac7799a7a32652cc91581a845ec2b111d4597516

Observation b826cf13-c41b-45c7-a675-389fb0854b74 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates Proximal Policy Optimization Algorithms

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.153829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.153829Z digest=sha256:6aab607dfb6ffbd2b2300b5b5026edb1c7fc82e4cee3cbfb54b1a7491d06a6e2

Observation f6fa7d71-6ba3-4732-a4be-74379d801690 · outbound

This paper cites Distributionally Robust Optimization: A Review.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates Distributionally Robust Optimization: A Review

Reference 2010

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.132881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.132881Z digest=sha256:1daf8624427f1e7b39e6ff986f4bb8abe1c70e34448ea6bb3ddd97eb1e5910f4

Observation 9725de26-8f99-4fb8-b5a5-917649045b7a · outbound

This paper cites Off-policy deep reinforcement learning without exploration.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates Off-policy deep reinforcement learning without exploration

Reference 2015

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.015793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.015793Z digest=sha256:4fa76c698af8d1db761ec3615d29d002ebaaa8b514b31e75c757482372191b7b

Observation 5d9f9913-7643-4257-a1b3-559f15af8fc1 · outbound

This paper cites Deep Reinforcement Learning in Large Discrete Action Spaces.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates Deep Reinforcement Learning in Large Discrete Action Spaces

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.005106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.005106Z digest=sha256:eeb3292c86b685cc0c93d8b33de4acd2727d62a94a946908c23843d0f1586366

Observation 2edceb6d-9274-41de-b5e9-dcaf79eb5a4a · outbound

This paper cites AWAC: Accelerating Online Reinforcement Learning with Offline Datasets.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.097009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.097009Z digest=sha256:45f091e01ac73dc01c9cfdfeeee1eb85976246b798451965be2ec675af18f57c

Observation 65a9af51-c76e-4870-9f5f-d7ca59236381 · outbound

This paper cites Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems.

Breaking the Curse of Repulsion: Remoteness-Aware Control of Negative Off-Policy Updates Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-03T01:15:34.075775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:15:34.075775Z digest=sha256:fe1ff45283c756e9df41d7b4039f46e3ed92f7f58ada344bc395613f3349b17d

Pith citing papers

No inbound Pith citation observations are available.