Pith. sign in

Paper Citation Record · LEDGER

Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2101.07123.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2101.07123 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:20:39.307427Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T06:55:29.516438Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b7f302f2-cda0-4f36-abda-0e42be550535 · inbound

Temporal Representation Alignment: Successor Features Enable Emergent Compositionality in Robot Instruction Following cites this paper.

Temporal Representation Alignment: Successor Features Enable Emergent Compositionality in Robot Instruction Following Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T19:20:39.307427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:20:39.307427Z digest=sha256:7ec4b96abc546b69ba9f0d5ca9c59e1e8271586c934e735f98fc600abd32922c

Observation 0eec7250-8f73-4841-8163-0584315efc57 · inbound

Closing the Gap between TD Learning and Supervised Learning with $Q$-Conditioned Maximization cites this paper.

Closing the Gap between TD Learning and Supervised Learning with $Q$-Conditioned Maximization Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:25.733026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:05:25.733026Z digest=sha256:bb17ff123c511339c3ddd0940f17567378552466359bd18f5fe1c35a5aa43204

Observation 477dd4b7-d9b0-455f-9344-1668513ca4ba · inbound

Intention-Conditioned Flow Occupancy Models cites this paper.

Intention-Conditioned Flow Occupancy Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:27:14.616008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T10:24:52.160209Z digest=sha256:4eed11c8e16e528f86053159ee9bec1de3e199275806977e483a47480e30f9e2

Observation f794af6c-6908-480e-9bcd-2e4b4551201b · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T09:17:14.107283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:b269f9dbf1c470bfb1dff96933b2a171a3765c1290fa3984333c901e94c7999f

Observation aa8d2a68-e130-47e1-8ba9-d8cd7914e5c6 · inbound

Visual Pre-Training on Unlabeled Images using Reinforcement Learning cites this paper.

Visual Pre-Training on Unlabeled Images using Reinforcement Learning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T01:06:16.986525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:06:16.986525Z digest=sha256:2d43b728c5f2bbcaae886209715bfd57b0cab6d50702c72685e939e339d87d3d

Observation 0211dde7-c40f-4a47-b775-6b561654b0ed · inbound

Epistemically-guided forward-backward exploration cites this paper.

Epistemically-guided forward-backward exploration Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:32:32.924634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:32:32.924634Z digest=sha256:9c22063ac3a323c466834dcec14214baf91c786b7ee542b1bfd979d1dc0b6f56

Observation e061b358-255f-4ce7-9a09-2e8fc4a78851 · inbound

Spectral Alignment in Forward-Backward Representations via Temporal Abstraction cites this paper.

Spectral Alignment in Forward-Backward Representations via Temporal Abstraction Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:35:19.258076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T08:32:21.705085Z digest=sha256:226beb836ed3c38b5412234ac9c5ccb9a034a49684a386f33ba728f5aa040161

Observation 0307ed90-5d82-46c7-a7dd-57725bd5f701 · inbound

SVL: Goal-Conditioned Reinforcement Learning as Survival Learning cites this paper.

SVL: Goal-Conditioned Reinforcement Learning as Survival Learning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:46:37.135018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T06:45:54.155848Z digest=sha256:184796271167585322575282f3e631437ad9e3f9fecd738d1ea59ad2a170b203

Observation 6f39062a-6a80-4e0a-ac1a-6d90532133d5 · inbound

Understanding Human Actions through the Lens of Executable Models cites this paper.

Understanding Human Actions through the Lens of Executable Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:10:09.777561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T04:54:29.565609Z digest=sha256:2e31a24904490fd693b0e279064b30e9fa42f77c64ceb3f91fb87c5e99b14736

Observation 0f1edb4e-c796-4a8b-a7ac-8a37a60e40b7 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:57.816224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:f4be751262b50b5e2c24816faddb356cb2dcd26a116791fca8ff2bad934f3dbb

Observation 9f6d1144-0655-4e83-928e-8e6bb24dd369 · inbound

Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning cites this paper.

Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:42:58.552888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T20:26:35.019753Z digest=sha256:7dd1d149658fd2f594b97f8dc1a72400c3529cf51867204f590659f4239f2fcc

Observation 44ae2fa9-89eb-434c-b699-60bf093fbad7 · inbound

Offline Reinforcement Learning with Universal Horizon Models cites this paper.

Offline Reinforcement Learning with Universal Horizon Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:48:57.499849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T19:45:15.458347Z digest=sha256:1cb5f7c45b3c7b2dfe2e8493a9be513330946a035396a770517bea3456dbd91e

Observation fb48cde2-afe1-4053-a90d-38dd781e6fc2 · inbound

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited cites this paper.

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:37:45.250688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T20:37:36.030165Z digest=sha256:9b315ee9bb296bd6ffe8c590926b6cb91973e20467e9cedda3ec24edf6c3de36

Observation 2329491f-01f8-4d0e-8440-266a531305c8 · inbound

Plan, Don't Pose: Long Composite Motion Generation with Text-Aligned BFM cites this paper.

Plan, Don't Pose: Long Composite Motion Generation with Text-Aligned BFM Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:53:15.706839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T08:51:14.809528Z digest=sha256:071ebf0bd98f8d9f4dca156d043812926ddb4bfd2f3d00f410e7de349c64edf7

Observation 678e5fb1-3c4f-4430-b7b8-b626efcb2039 · inbound

Exploration and Online Transfer with Behavioral Foundation Models cites this paper.

Exploration and Online Transfer with Behavioral Foundation Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:55:29.518692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-01T06:50:59.743859Z digest=sha256:43ffd71156fa3a4a8fe368a66a7cae909484371e4b711223ec62619fa119989a