Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:1910.11956.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:26:30.940459Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T13:39:51.078263Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation b0592225-ea90-47e1-bf98-8704da4577f9 · inbound
D4RL: Datasets for Deep Data-Driven Reinforcement Learning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 198e1864-ad6b-4f8a-b3c7-0b888f799c45 · inbound
VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4defba8b-2764-4aeb-aa63-2183159555b1 · inbound
Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 120
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation de2072f5-0ece-47fb-96a7-13ab1a87fafa · inbound
Diffusion Policy Policy Optimization Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5a0b22ca-5241-4357-9b91-ab026ad64f38 · inbound
Direct Preference Optimization for Primitive-Enabled Hierarchical RL: A Bilevel Approach Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f361a955-fa05-4f92-93bc-d4217624d2d4 · inbound
Block-wise Adaptive Caching for Accelerating Diffusion Policy Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation df18f18a-39af-47af-a11c-a273e36f4d61 · inbound
Steering Robots with Inference-Time Interactions Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85db7f8c-2df1-49c7-836d-3b681e7bcffb · inbound
VQ-VLA: Improving Vision-Language-Action Models via Scaling Vector-Quantized Action Tokenizers Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70a90a27-ee19-4c55-90ab-040b1ee96e71 · inbound
FOUNDER: Grounding Foundation Models in World Models for Open-Ended Embodied Decision Making Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59e65321-a80a-46f8-8a06-056ab68d8390 · inbound
Learning Temporal Abstractions via Variational Homomorphisms in Option-Induced Abstract MDPs Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1c9652e-a846-42bb-b6cc-536c9be01cd3 · inbound
D3P: Dynamic Denoising Diffusion Policy via Reinforcement Learning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17b2e3e5-a7c8-4e3a-ae6b-d8f26c2d3b74 · inbound
Learning Upper Lower Value Envelopes to Shape Online RL: A Principled Approach Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5a808f4-74b4-44be-b620-5437f98b78f3 · inbound
DASIP: Dynamic Test-Time Compute Scaling for Robot Control with Stochastic Interpolant Policies Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bea7f2a-15cc-46e8-a837-adb104d549b1 · inbound
Sparse ActionGen: Accelerating Diffusion Policy with Real-time Pruning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c387d72-f425-4ff5-bcf3-e6344341bcfd · inbound
Efficient Multi-Objective Planning with Weighted Maximization Using Large Neighbourhood Search Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b7d900f5-c284-46da-a984-bf67f9623a63 · inbound
Learning 3D Representations for Spatial Intelligence from Unposed Multi-View Images Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0950a9e9-687e-47f0-a34d-48ea79c64616 · inbound
ScoRe-Flow: Complete Distributional Control via Score-Based Reinforcement Learning for Flow Matching Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 915a78e5-f88f-4878-a8fc-5cca01a41ca9 · inbound
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 147
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 242da48d-3abb-4790-885d-79a84a51b9fd · inbound
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 88d10d64-f776-49cd-babb-95b915500970 · inbound
Learning Reactive Dexterous Grasping via Hierarchical Task-Space RL Planning and Joint-Space QP Control Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e350794-77b1-49ad-9769-58a7a6d14469 · inbound
World Action Models: The Next Frontier in Embodied AI Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 224
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1dfa8652-7124-4f99-a039-7d29d0201766 · inbound
DiLA: Disentangled Latent Action World Models Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a0bcfbde-a451-4380-94f0-5510b9f0adff · inbound
Drift Flow Matching Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f0d5cfea-ba06-45d2-831a-0b89a5816f8f · inbound
Understanding Multimodal Failure in Action-Chunking Behavioral Cloning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 619cd848-645e-45ef-9a86-fcc2a7408db6 · inbound
Automating Potential-based Reward Shaping with Vision Language Model Guidance Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5d93897c-4d3d-4532-b647-8fb3651a43ca · inbound
Learning Task-Sufficient World Models by Synergizing Agentic Exploration and Structured Modeling Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 135
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd3fb91f-ada2-4ef3-9ee7-e3907b692c60 · inbound
Dual-Process Atomic Skill Learning: Decoupling Semantic Reasoning and Real-Time Control Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42ade7d4-095b-4500-ba52-3515edd58d5f · inbound
Relative Value Learning Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 115
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86aeffef-3846-4676-acbd-1de15e085da4 · inbound
Data Pyramid for Embodied Manipulation Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 132
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a92d4533-c64d-43b5-9267-6193ab9d44c5 · inbound
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills Relay Policy Learning: Solving Long-Horizon Tasks via Imitation and Reinforcement Learning
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.