Pith. sign in

Paper Citation Record · LEDGER

Bridging State and History Representations: Understanding Self-Predictive RL

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2401.08898.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.08898 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T23:13:07.349666Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T10:27:14.544741Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0decce05-3a09-4d87-85da-b092720781fe · inbound

Hadamax Encoding: Elevating Performance in Model-Free Atari cites this paper.

Hadamax Encoding: Elevating Performance in Model-Free Atari Bridging State and History Representations: Understanding Self-Predictive RL

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:23:18.249383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:23:18.249383Z digest=sha256:3a801e3e3f69d3f3d090bcf680c2067f62e62e567ad87550b23e50a087090576

Observation c7d1b2e4-9bf3-443a-8c6e-ece296354c84 · inbound

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments cites this paper.

Understanding Behavioral Metric Learning: A Large-Scale Study on Distracting Reinforcement Learning Environments Bridging State and History Representations: Understanding Self-Predictive RL

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:10:04.252390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:10:04.252390Z digest=sha256:011c0df7418a2f7fdf8635673599ca9102218a9312f10d77a871c3ee97e4349c

Observation 5f3ef2cc-45e1-413c-958c-aa0fbc723349 · inbound

Intention-Conditioned Flow Occupancy Models cites this paper.

Intention-Conditioned Flow Occupancy Models Bridging State and History Representations: Understanding Self-Predictive RL

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:27:14.546674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T10:24:52.160209Z digest=sha256:dfab934b63a580644e8f987a75edd1966dbab3aedca7b3a72a647cfeec99ac59

Observation a5e4e9e1-19f5-4833-874f-66ad4f364888 · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Bridging State and History Representations: Understanding Self-Predictive RL

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-19T09:17:14.174785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:0327cffed2604eea72de9e5db9802689936eab17c6637247e5ce39122dbd5e79

Observation e961f44d-52cf-4de8-99ab-6c98522680bb · inbound

Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access cites this paper.

Informed Asymmetric Actor-Critic: Leveraging Privileged Signals Beyond Full-State Access Bridging State and History Representations: Understanding Self-Predictive RL

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T13:43:30.088955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:43:30.088955Z digest=sha256:94536e788ecd26d9bbac29eca7ba5c8c60ac3edf3abeb1e05e6121d75eb0010f

Observation 5e01d52c-1e94-4f20-9bfd-5669015cfaf7 · inbound

Now You See That: Learning End-to-End Humanoid Locomotion from Raw Pixels cites this paper.

Now You See That: Learning End-to-End Humanoid Locomotion from Raw Pixels Bridging State and History Representations: Understanding Self-Predictive RL

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:27:32.259434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T07:23:03.673259Z digest=sha256:5dbdc8b352d1e4f23e8e8593c6eebc7a12788df834dd857f5203864e6bebf376

Observation e95d30f0-92fd-4d0e-a404-9e1d6634045b · inbound

Can We Really Learn One Representation to Optimize All Rewards? cites this paper.

Can We Really Learn One Representation to Optimize All Rewards? Bridging State and History Representations: Understanding Self-Predictive RL

Reference 2023

Resolution
malformed identifier
no resolver link, observed 2026-08-03T00:15:46.425044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:15:46.425044Z digest=sha256:3b0e03c9e4f87e25ca113de41fda6aafac8aaca85c7b4d8f689ad162ef84655d

Observation fc878abf-3964-4754-982b-2fc77cd7285a · inbound

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction cites this paper.

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction Bridging State and History Representations: Understanding Self-Predictive RL

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:50:05.098809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T14:46:18.645009Z digest=sha256:a348cd6142231382cf18dced7ca31acd4b5634a2d44e1fcbe7cda9fce1a168e7

Observation 089e678a-6793-4f26-94a9-74919ce31458 · inbound

Behavior-Constrained Reinforcement Learning with Receding-Horizon Credit Assignment for High-Performance Control cites this paper.

Behavior-Constrained Reinforcement Learning with Receding-Horizon Credit Assignment for High-Performance Control Bridging State and History Representations: Understanding Self-Predictive RL

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-13T19:33:10.250941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T19:30:23.447901Z digest=sha256:91347529897da0a3c49177194ff3feca44f86f29df68e06eb9a6bf6b9e9177b1

Observation 8698c271-a127-42da-be19-1afc2fc38300 · inbound

The University AI Didn't Replace -- Rethinking Universities in the AI Era cites this paper.

The University AI Didn't Replace -- Rethinking Universities in the AI Era Bridging State and History Representations: Understanding Self-Predictive RL

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T17:22:42.591825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T17:22:42.591825Z digest=sha256:0d468e12b4c8d1917efc4be3bcb8c0f4389eb34c6d8bf2ef29d176707ec8c79a

Observation 1c04a257-0296-4fbd-9fdb-2c13687a2507 · inbound

Integrating Causal DAGs in Deep RL: Activating Minimal Markovian States with Multi-Order Exposure cites this paper.

Integrating Causal DAGs in Deep RL: Activating Minimal Markovian States with Multi-Order Exposure Bridging State and History Representations: Understanding Self-Predictive RL

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:15:56.173195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T02:34:12.141437Z digest=sha256:365940b7a6f44a6bf0f5828c5ab1a6317064ff34ebf4858dce546eee9856c00e

Observation 0cb8191a-1a14-4841-9e13-0f12692b0a2c · inbound

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback cites this paper.

LEMUR: Learning to Align with Multi-Objective Reinforcement Learning from Preference Feedback Bridging State and History Representations: Understanding Self-Predictive RL

Reference 257

Resolution
unresolved
no resolver link, observed 2026-08-03T04:39:32.161140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T04:39:32.161140Z digest=sha256:582cf5bf3e1decbde20a022ee4a954eb71780aa6470c55bc6961e5de00dbed36

Observation fe6e40b3-25ba-434b-815b-4e402073984e · inbound

Hierarchical Latent Prediction for Language Models cites this paper.

Hierarchical Latent Prediction for Language Models Bridging State and History Representations: Understanding Self-Predictive RL

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T23:13:07.349666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T23:13:07.349666Z digest=sha256:982d712b06072f9034a297371517a97794339769f67bfdd26562de111240fe56