Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T16:45:08.340242Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 19 inbound Pith citation observations for arXiv:2412.09858.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T16:45:08.340242Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-08T19:20:33.558796Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T16:49:58.308235Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 9464cbea-e381-424b-84d0-104b0cf6075e · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Ball, Laura Smith, Ilya Kostrikov, and Sergey Levine
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7d00e518-5cb2-4e1b-a6c1-db8b14f8a353 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d00c172-07b8-4636-b490-b7b01f03a7f9 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6048207-513b-4f7c-a908-5cd37418cd31 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning RT-1: Robotics Transformer for Real-World Control at Scale
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71da2607-7bd9-4001-9cb3-c3a03890cc6a · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning GR-2: A Generative Video-Language-Action Model with Web-Scale Knowledge for Robot Manipulation
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6376c945-8db7-430e-b11b-0a059467daf1 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Diffusion Policy: Visuomotor Policy Learning via Action Diffusion
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea1759bc-13fe-485e-8584-c0ece27fa990 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 56e78ba4-9034-4b60-b157-dcc7cc89dfde · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning PaLM-E: An Embodied Multimodal Language Model
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 133ca315-e9c7-4d60-9efc-19a322f7eb0a · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Divide-and-Conquer Reinforcement Learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45ed9d52-e918-4c2c-b326-4a4ab9434c93 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Zhao, Vikash Kumar, Aaron Rovinsky, Kelvin Xu, Thomas Devlin, and Sergey Levine
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a8b8572-bddb-4b40-8ab5-48cdc27b8151 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Denoising Diffusion Probabilistic Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5d26e27-f1ce-415f-8192-5c006e478ce5 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0dc69b3c-bd94-44e2-bfb6-9bc071348992 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Imitation Bootstrapped Reinforcement Learning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6908ba4-07c1-41ce-a29b-a19df1ac8849 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning REBOOT: Reuse Data for Bootstrapping Efficient Real-World Dexterous Manipulation
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65a46871-d05d-4f35-a1c3-019521b0f284 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Residual reinforcement learning for robot control
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1a588c8-75f0-4b30-8010-055f4d67368c · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning OpenVLA: An Open-Source Vision-Language-Action Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e58e469-aedd-447a-9037-38d835477199 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Learning neural network policies with guided policy search under unknown dynamics
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d5c85c4c-6be1-4d02-a25f-7a66245bdd6e · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning End-to-end training of deep visuomotor policies
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c12b06fb-1ab1-4016-8a7b-10b941794990 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b253e2e0-1fc5-4e49-a82d-6f9958016a5f · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Deep reinforcement learning for robotic assembly of mixed deformable and rigid objects
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 92555464-8fc0-4599-9a8b-ab1bfaabdf61 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Reinforcement learning on variable impedance controller for high-precision robotic assembly
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fa6b2229-d660-429f-bdb8-bb506c4d0fd9 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Robust multi-modal policies for industrial assembly via reinforcement learning and demonstrations: A large-scale study
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation bad56456-90cb-4b0c-a706-99b789b1f465 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning RLIF : Interactive imitation learning as reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 38b03238-9dd2-4bce-b499-7e0e6178ffcf · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Serl: A software suite for sample-efficient robotic reinforcement learning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d1e1726-17ef-4307-a47f-89c887070c2f · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning FMB: a Functional Manipulation Benchmark for Generalizable Robotic Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1e3aadc-b9aa-4feb-82ed-dabeb2145b61 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Precise and Dexterous Robotic Manipulation via Human-in-the-Loop Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fd743c9-b787-40de-9d20-290938471e95 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Actor-Mimic: Deep Multitask and Transfer Reinforcement Learning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9bd3708-d93d-46a1-b73f-699fd238c627 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Learning Complex Dexterous Manipulation with Deep Reinforcement Learning and Demonstrations
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c35a80e-161d-4f9b-8c33-529d13c75907 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Learning complex dexterous manipulation with deep reinforcement learning and demonstrations
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation fd3b14be-c3da-4175-88c6-c592dbe0daee · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Policy Distillation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d732340-a4b3-4f62-a2bd-4b3bafde98bf · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Progressive neural networks
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dea355e5-946f-4da3-bdda-a0af9e9b3586 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Deep reinforcement learning for industrial insertion tasks with visual inputs and natural rewards
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bca4bdb5-3687-4851-ab65-b483ac75c891 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Progress & compress: A scalable framework for continual learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 096f0151-ea13-439f-ad17-5b1229f01b79 · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Octo: An Open-Source Generalist Robot Policy
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9c8f989-c18d-440c-8e3d-26ccbb8333cb · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Distral: Robust Multitask Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f482e84-dc27-43bc-9efe-f8f676abc1bb · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 264b91d2-72dc-4855-80c4-7cb63f9888fc · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning TinyVLA: Towards Fast, Data-Efficient Vision-Language-Action Models for Robotic Manipulation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f25fc5f3-bb27-48ee-9f85-5eefd3aa482b · outbound
RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning Zhao, Jianlan Luo, Oleg Sushkov, Rugile Pevceviciute, Nicolas Heess, Jon Scholz, Stefan Schaal, and Sergey Levine
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b34cab0-a854-47b8-b372-e815dbf85c9e · inbound
ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3364acc7-9aa7-4326-8492-fcdad8c6e926 · inbound
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 360cd47c-54bd-481e-a586-d2ea7a05add6 · inbound
Integrating Diffusion-based Multi-task Learning with Online Reinforcement Learning for Robust Quadruped Robot Control RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 898b5141-124a-41f3-8456-6dccdf659324 · inbound
Arnold: a generalist muscle transformer policy RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 298d2bfe-c995-4e35-9711-a71c32863790 · inbound
$\pi^{*}_{0.6}$: a VLA That Learns From Experience RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1fe7b4aa-83b4-4e27-a0c9-959c76ea977d · inbound
ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6bac5bf-d19d-453b-b1a2-575a2262b002 · inbound
${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1a33a8ec-78ce-4ec8-acf0-89a077ad6181 · inbound
Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8f56a2a4-f713-460f-b1f0-517b52dff45c · inbound
Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6008e264-f2bc-4cc8-b535-58f40c74c76b · inbound
Multi-Objective Learning for Diffusion Models: A Statistical Theory under Semi-Supervised Learning RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 17a9adad-e6d9-44dc-9c31-3d559f5d3484 · inbound
Flow-based Policy Adaptation without Policy Updates RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8270259e-b914-4de0-851f-481fa9787007 · inbound
DexPIE: Stable Dexterous Policy Improvement from Real-World Experience RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cfbea7c4-f093-4a13-9c56-293551becfd1 · inbound
AllDayNav: Lifelong Navigation via Real-World Reinforcement Learning RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dc49a3a3-079f-48c7-82cc-22ea3731d402 · inbound
Improving Robotic Generalist Policies via Flow Reversal Steering RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 62d40101-1ed8-4d44-9e87-b271bf9dba30 · inbound
JoyAI-Sim: A Simulation-Enabled Interconversion Toolchain for the Embodied Data Pyramid RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e2a28978-94d0-46ca-aa7e-c48aa2622140 · inbound
Scalable Multi-Task Data Generation via Reinforcement Learning for Language-Conditioned Bimanual Dexterous Manipulation RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 0fb83220-c451-4f5b-8c9c-c6527a00d5d1 · inbound
Scalable Multi-Task Data Generation via Reinforcement Learning for Language-Conditioned Bimanual Dexterous Manipulation RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e64a8b7e-6135-4813-bd8b-490bde4e5899 · inbound
InSight: Self-Guided Skill Acquisition via Steerable VLAs RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 6f15f21a-c235-49f2-ad1e-e3fc6bf16095 · inbound
RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation RLDG: Robotic Generalist Policy Distillation via Reinforcement Learning
Reference 189
Source-reported events for the cited work
Unavailable: canonical work link unavailable.