Pith. sign in

Paper Citation Record · LEDGER

Learning Multi-Level Hierarchies with Hindsight

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 15 inbound Pith citation observations for arXiv:1712.00948.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1712.00948 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 15 of 15 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:49:12.947806Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T02:07:34.156817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 81199ecb-c522-41f6-b07a-18363b25401f · inbound

Learning World Graphs to Accelerate Hierarchical Reinforcement Learning cites this paper.

Learning World Graphs to Accelerate Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-25T12:35:49.133772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T12:31:38.848720Z digest=sha256:d26c42c427dd517b63ab4343d22ab3fcf947e1681544c4af69e1a72cd907b5f2

Observation 257f28c5-d6ed-4b06-bfeb-5d0b6e4cb1e6 · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 217

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T12:04:10.833730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:819451159896d9ee8e4b43b505148262f06e1f5b1e75bcdd34846dd05e9ae063

Observation 77b7febc-4a27-4da2-bba3-9c65c30fe74c · inbound

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections cites this paper.

Goal-conditioned Hierarchical Reinforcement Learning for Sample-efficient and Safe Autonomous Driving at Intersections Learning Multi-Level Hierarchies with Hindsight

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T23:49:12.947806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:49:12.947806Z digest=sha256:70438113936278344246e7805ce3e65fd73479dafe981177a8d0c5369265de12

Observation 55873d67-a843-40b5-a060-7d878673b2f0 · inbound

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning cites this paper.

Strict Subgoal Execution: Reliable Long-Horizon Planning in Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:54:31.206433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T00:53:46.002945Z digest=sha256:371d6e8cde9ed97529f0ed0e50ab81556c9e0d2ba9c68a9564c16d4ad3289b06

Observation c4a50341-8965-468d-98c0-0d740f3f11d1 · inbound

Scalable Option Learning in High-Throughput Environments cites this paper.

Scalable Option Learning in High-Throughput Environments Learning Multi-Level Hierarchies with Hindsight

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T20:06:49.738978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T20:04:58.064472Z digest=sha256:292c77e63441d90536b013c59f40062e9a4bb9e5012813477708416dc5cb8ea3

Observation 5e35741b-8585-4856-96b7-a9ebb544b128 · inbound

Combined Constrained Sampling and Reinforcement Learning for Robotic Manipulation cites this paper.

Combined Constrained Sampling and Reinforcement Learning for Robotic Manipulation Learning Multi-Level Hierarchies with Hindsight

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T03:18:43.091575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:18:43.091575Z digest=sha256:b670db4839d6a5658b5bd465ba96e51f3bef634fe2a7d3b2ec2416b61f908dbb

Observation a0c15070-9519-4b6b-864c-67ed56e00bd4 · inbound

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL cites this paper.

QHyer: Q-conditioned Hybrid Attention-mamba Transformer for Offline Goal-conditioned RL Learning Multi-Level Hierarchies with Hindsight

Reference 101

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:30:58.065014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-11T01:17:48.643521Z digest=sha256:0fe88e8e28822e71b91dcc277c530b4b3a49391b817465e4e52f66f3ac1b04bf

Observation 10a79b59-f989-4d6f-9a5f-ff99fc4abf32 · inbound

Delay-Empowered Causal Hierarchical Reinforcement Learning cites this paper.

Delay-Empowered Causal Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:47:21.233473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T05:46:51.659283Z digest=sha256:9a8de59ad118d8ad8cad7abbe97464d1d208d38f5941e982702865a149eca83b

Observation 11ae49c6-7149-4f27-829e-509333980c88 · inbound

Abstraction for Offline Goal-Conditioned Reinforcement Learning cites this paper.

Abstraction for Offline Goal-Conditioned Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:51:16.522569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T07:46:20.289421Z digest=sha256:1f178df2ca0ace003b6b44f3813a6d1b55962e98598a51c78a023f83a552ef4c

Observation 4bd6a974-0da4-402a-b89c-7116527be498 · inbound

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling cites this paper.

Goal Sets, Not Goal States: Queryable Robot Goals through Goal-Set Hindsight Relabeling Learning Multi-Level Hierarchies with Hindsight

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T02:07:34.158320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T16:07:28.361043Z digest=sha256:1298dbba62125a8971cfbbba227a2b9df5bbe421e1ef66fd2aecef99f9ff966d

Observation 99c8ec5e-840c-437a-9ce2-1e0ccfab4028 · inbound

Vision-Based Obstacle Separation for Strawberry Harvesting in Clusters Using Hierarchical Reinforcement Learning cites this paper.

Vision-Based Obstacle Separation for Strawberry Harvesting in Clusters Using Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T03:42:47.680209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:42:47.680209Z digest=sha256:8db5794863948fa94f7034508fee26ea78d98bbc68307e4f95e0971428eefa17

Observation fcd44bf2-ed5b-4e57-bc47-5d60f92935d3 · inbound

S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning cites this paper.

S3: Stable Subgoal Selection by Constraining Uncertainty of Coarse Dynamics in Hierarchical Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T13:10:39.713561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T13:10:39.713561Z digest=sha256:4c4e9075eecef78c8810af7e89e04007c85189f4fdd621c3dd87cc17febfab0b

Observation e82ab8b3-acef-4e7a-a1df-96add91289d3 · inbound

Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning cites this paper.

Hierarchical Soft Actor-Critic for Sparse-Reward Long-Horizon Reinforcement Learning Learning Multi-Level Hierarchies with Hindsight

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-30T14:49:29.163862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T14:49:29.163862Z digest=sha256:0b29924d0361e2037852984f5bb28040ed0d5db78d22abef2daf5b7ff9ac6a33

Observation 460194cc-9184-4098-9d02-d495822f57c1 · inbound

Hierarchical Residual Policy Optimization for Generative Recommendations cites this paper.

Hierarchical Residual Policy Optimization for Generative Recommendations Learning Multi-Level Hierarchies with Hindsight

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T00:28:21.497424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T00:28:21.497424Z digest=sha256:98ae3bbaf918c1401583f3a6298752fd475210c00434c6b79cd408e7676a52c9

Observation 72a1f002-9694-4f28-b2f1-e02318a63af0 · inbound

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills cites this paper.

Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Learning Multi-Level Hierarchies with Hindsight

Reference 132

Resolution
unresolved
no resolver link, observed 2026-08-04T19:45:34.875368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:45:34.875368Z digest=sha256:707bd1f20b22ee3422cf3673b03ce8401b2b394fa31fd6aa33b3d5abcf5493a5