Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T00:46:28.503923Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 0 inbound Pith citation observations for arXiv:2607.09042.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-13T00:46:28.503923Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f00161e5-b403-4de4-9476-b508fe2ac1df · outbound
Learning More from Less: Reinforcement Learning from Hindsight Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 843f21d0-ec44-4fd1-bd04-68ca84bb0645 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Ouyang, J
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8103bb86-da10-4f95-88b2-7d516f4dd7d0 · outbound
Learning More from Less: Reinforcement Learning from Hindsight GPT-4 Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a17b14ad-54de-4652-81fa-ca65d5896356 · outbound
Learning More from Less: Reinforcement Learning from Hindsight $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de8fcbd5-e2dd-4a71-93e2-d42332420541 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Bellemare, S
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dcbeed02-f257-495e-af2e-e7ef0a1bb26d · outbound
Learning More from Less: Reinforcement Learning from Hindsight Pathak, P
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a35f4df-face-414b-9e7c-342ac8ba2f85 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Andrychowicz, F
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f27f1afe-7513-44b5-b258-bb857c5a39c3 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Hindsight policy gradients
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9897961-0b38-4890-950c-2e8118b9f40b · outbound
Learning More from Less: Reinforcement Learning from Hindsight Pathak, P
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7296443-a23c-4cb1-8123-e4d03618e392 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Eysenbach, T
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62dfe183-d33d-4202-8e52-d14672b9a8cf · outbound
Learning More from Less: Reinforcement Learning from Hindsight Sahni, T
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36d24651-2d1f-4ca4-8a93-acf8adbf4371 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Robometer: Scaling General-Purpose Robotic Reward Models via Trajectory Comparisons
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1aa68f0-f66b-4a27-9754-ab748f90b9e8 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Qwen2.5-VL Technical Report
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d0eb24d-430f-4a63-9ba3-99c081bbdb3a · outbound
Learning More from Less: Reinforcement Learning from Hindsight LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43ae3d0d-4a67-4b72-bfa8-56d0de804332 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Autonomous Improvement of Instruction Following Skills via Foundation Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e39e0b9-01c2-4be0-9674-7f933d8fded8 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Learning Instruction-Following Policies through Open-Ended Instruction Relabeling with Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4238627-2b95-455f-8382-181629a0ca7d · outbound
Learning More from Less: Reinforcement Learning from Hindsight Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1f74b4f-9c0e-4037-acb4-f8b45d13a328 · outbound
Learning More from Less: Reinforcement Learning from Hindsight VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c0f97db-bee2-4f1f-837b-37186fefb61d · outbound
Learning More from Less: Reinforcement Learning from Hindsight Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e69b5bc-d6f4-4c2c-8eb8-bf1f51bfefac · outbound
Learning More from Less: Reinforcement Learning from Hindsight Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c41ceef-7621-49ea-aace-af2755895afc · outbound
Learning More from Less: Reinforcement Learning from Hindsight Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaa22542-473a-4093-a51c-0125399e5ac2 · outbound
Learning More from Less: Reinforcement Learning from Hindsight RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 794d7590-b709-4e18-a185-63f9ad7ae119 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Robotic Skill Acquisition via Instruction Augmentation with Vision-Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e90da314-6691-4ee4-b5b0-cd6728142487 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Zhang, K
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19241193-ec53-4c31-aa8b-0e4de2d803c8 · outbound
Learning More from Less: Reinforcement Learning from Hindsight CAST: Counterfactual Labels Improve Instruction Following in Vision-Language-Action Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64dedc25-74bc-476b-8001-6c1344634fca · outbound
Learning More from Less: Reinforcement Learning from Hindsight Unresolved cited work
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff983015-b822-4591-a033-a53a3fbe89de · outbound
Learning More from Less: Reinforcement Learning from Hindsight Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bea7da46-8bc4-4d26-b9db-03bbbde43f2c · outbound
Learning More from Less: Reinforcement Learning from Hindsight OpenVLA: An Open-Source Vision-Language-Action Model
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5b60470-9163-48a8-81cc-8bffad55ac2c · outbound
Learning More from Less: Reinforcement Learning from Hindsight DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15b09afd-8d72-4ae7-8715-efc0d9995fac · outbound
Learning More from Less: Reinforcement Learning from Hindsight Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b798d2e7-5db2-4ea2-8aef-2ace63374622 · outbound
Learning More from Less: Reinforcement Learning from Hindsight Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27992df7-1055-4f59-ae4e-f7eacf1eec31 · outbound
Learning More from Less: Reinforcement Learning from Hindsight GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e7a3819-3aa5-4e81-9e1f-b57d85585d0a · outbound
Learning More from Less: Reinforcement Learning from Hindsight Nothing” option that allows the relabeler to identify uninteresting trajectories, which are then filtered out during training, and (3) an “unsure
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff1f4836-c156-4ca1-a515-42b0589dc212 · outbound
Learning More from Less: Reinforcement Learning from Hindsight pick up the mug and place it in the microwave
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b48d3e7-308a-484a-b1ab-79e69e452a08 · outbound
Learning More from Less: Reinforcement Learning from Hindsight pick up the tape
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.