Pith. sign in

Paper Citation Record · LEDGER

Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:2101.07123.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2101.07123 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:20:39.307427Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T06:55:29.516438Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b7f302f2-cda0-4f36-abda-0e42be550535 · inbound

Temporal Representation Alignment: Successor Features Enable Emergent Compositionality in Robot Instruction Following cites this paper.

Temporal Representation Alignment: Successor Features Enable Emergent Compositionality in Robot Instruction Following Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T19:20:39.307427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T19:20:39.307427Z digest=sha256:048e0b910f56f51576670559a996305b8008e5d3bdf4232b0335a209f6ce7647

Observation 0eec7250-8f73-4841-8163-0584315efc57 · inbound

Closing the Gap between TD Learning and Supervised Learning with $Q$-Conditioned Maximization cites this paper.

Closing the Gap between TD Learning and Supervised Learning with $Q$-Conditioned Maximization Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:05:25.733026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:05:25.733026Z digest=sha256:1bae6cb4549d35a97dbd67da3f0319744db91b23d374b607b1d2f637dbae4694

Observation 477dd4b7-d9b0-455f-9344-1668513ca4ba · inbound

Intention-Conditioned Flow Occupancy Models cites this paper.

Intention-Conditioned Flow Occupancy Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T10:27:14.616008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T10:24:52.160209Z digest=sha256:5cec8b3f4d81e0dd9e05f7a0565990617bde577002cd5438de17c22fa7d9727d

Observation f794af6c-6908-480e-9bcd-2e4b4551201b · inbound

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning cites this paper.

Self-Predictive Representations for Combinatorial Generalization in Behavioral Cloning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T09:17:14.107283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T09:15:45.511104Z digest=sha256:97b1e55b0275429b7bf4b2624fe3d4186e15ecedd72419a9d16464242328712e

Observation aa8d2a68-e130-47e1-8ba9-d8cd7914e5c6 · inbound

Visual Pre-Training on Unlabeled Images using Reinforcement Learning cites this paper.

Visual Pre-Training on Unlabeled Images using Reinforcement Learning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T01:06:16.986525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:06:16.986525Z digest=sha256:3c5d9fc5728b4c6ece0623225eb0d9c5d82653ead965eb66656c040d65f1147a

Observation 0211dde7-c40f-4a47-b775-6b561654b0ed · inbound

Epistemically-guided forward-backward exploration cites this paper.

Epistemically-guided forward-backward exploration Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T19:32:32.924634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:32:32.924634Z digest=sha256:fd7755e965621370f7610b6339fe9ef86c267bd5c784cbaa016c73d0af7cdfe2

Observation e061b358-255f-4ce7-9a09-2e8fc4a78851 · inbound

Spectral Alignment in Forward-Backward Representations via Temporal Abstraction cites this paper.

Spectral Alignment in Forward-Backward Representations via Temporal Abstraction Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:35:19.258076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T08:32:21.705085Z digest=sha256:db8e272bad2251f8a61840d9df596dccef5701ecda146029182a9c48cd9a7312

Observation 0307ed90-5d82-46c7-a7dd-57725bd5f701 · inbound

SVL: Goal-Conditioned Reinforcement Learning as Survival Learning cites this paper.

SVL: Goal-Conditioned Reinforcement Learning as Survival Learning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:46:37.135018Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T06:45:54.155848Z digest=sha256:d0be1b45d8b63ea86fa2a5b6e76d712efead1a00484623fc35e2cc5152e323b8

Observation 6f39062a-6a80-4e0a-ac1a-6d90532133d5 · inbound

Understanding Human Actions through the Lens of Executable Models cites this paper.

Understanding Human Actions through the Lens of Executable Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:10:09.777561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T04:54:29.565609Z digest=sha256:d6293c2090edddf2305426b1c6e1d884e0ab8242ba953ce1fafa320b54bf38f1

Observation 0f1edb4e-c796-4a8b-a7ac-8a37a60e40b7 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:30:57.816224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:5941679ddcc9bb5d48395c6a557d14fbf671cbf69be1fa254196258dd4262859

Observation 9f6d1144-0655-4e83-928e-8e6bb24dd369 · inbound

Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning cites this paper.

Switching Successor Measures for Hierarchical Zero-shot Reinforcement Learning Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:42:58.552888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T20:26:35.019753Z digest=sha256:695c5d9c1554f4a30504e860b8ebb2cb5025c7eb9e65bade2e0cc23a38d2d04d

Observation 44ae2fa9-89eb-434c-b699-60bf093fbad7 · inbound

Offline Reinforcement Learning with Universal Horizon Models cites this paper.

Offline Reinforcement Learning with Universal Horizon Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:48:57.499849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T19:45:15.458347Z digest=sha256:a7d933d921993bb45eacca79973350486988aa3dfb3f33fc4ac4375ca706cacf

Observation fb48cde2-afe1-4053-a90d-38dd781e6fc2 · inbound

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited cites this paper.

When Dynamics Shift, Robust Task Inference Wins: Offline Imitation Learning with Behavior Foundation Models Revisited Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:37:45.250688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T20:37:36.030165Z digest=sha256:11aa10a2695be6c710c527bb911dba59671057982aa41aad7be0a3ae9699497f

Observation 2329491f-01f8-4d0e-8440-266a531305c8 · inbound

Plan, Don't Pose: Long Composite Motion Generation with Text-Aligned BFM cites this paper.

Plan, Don't Pose: Long Composite Motion Generation with Text-Aligned BFM Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:53:15.706839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T08:51:14.809528Z digest=sha256:1c4647cd0d14c717685b1f53de5b53d16371e0fac1daba83ddd35c834a2e3f0c

Observation 678e5fb1-3c4f-4430-b7b8-b626efcb2039 · inbound

Exploration and Online Transfer with Behavioral Foundation Models cites this paper.

Exploration and Online Transfer with Behavioral Foundation Models Learning Successor States and Goal-Dependent Values: A Mathematical Viewpoint

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-01T06:55:29.518692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T06:50:59.743859Z digest=sha256:a20a598053ce4cad97429c9172c3c82fa3479e3de5b63aa3cc54793cd5cc30f1