Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 46 inbound Pith citation observations for arXiv:2411.19309.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T19:38:41.768698Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 83a3fcea-f4b4-4d05-886d-741da72770f3 · inbound
DexVLA: Vision-Language Model with Plug-In Diffusion Expert for General Robot Control GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d799988c-429d-4e5f-9812-7004b9454f9e · inbound
SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2fe91573-9106-4706-b7c4-aa0c4edc2f88 · inbound
AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a6898c71-0575-4ffd-8d6e-122c1bd2b242 · inbound
UniVLA: Learning to Act Anywhere with Task-centric Latent Actions GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a50457bd-f4cd-4f9a-a775-a7e4dfb9c6c6 · inbound
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4bface85-924a-42ee-a04b-36476e85f515 · inbound
Adversarial Attacks on Robotic Vision Language Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29e67213-daad-427c-b70d-6acbf6f1b574 · inbound
Continual Learning for Generative AI: From LLMs to MLLMs and Beyond GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 258
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8878e613-9cc9-4eda-923c-404de6112c0d · inbound
Large VLM-based Vision-Language-Action Models for Robotic Manipulation: A Survey GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 177
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 02898c68-a010-4aa3-adf2-0cfe4243bf80 · inbound
Balancing Signal and Variance: Adaptive Offline RL Post-Training for VLA Flow Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd5ff45e-a0ba-4ba4-952a-a685a8f4638b · inbound
SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 54f664df-f197-40ff-89df-f5497b457f57 · inbound
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40ef77c0-734c-44db-9fd0-8c8f4fb483ea · inbound
Reflection-Based Task Adaptation for Self-Improving VLA GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a64082d7-3ec8-4d36-bd0f-2c30aa744d6b · inbound
RESample: A Robust Data Augmentation Framework via Exploratory Sampling for Robotic Manipulation GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1dd654af-0a47-4061-a394-7a1a3f3bf613 · inbound
$\pi^{*}_{0.6}$: a VLA That Learns From Experience GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation eb701436-3068-4bac-a2b7-ca52cf2f6edc · inbound
DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd542b69-2fa3-45a5-a2df-22e01d8028ab · inbound
VLA-Arena: An Open-Source Framework for Benchmarking Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62cb33a6-30c1-48ce-b3d1-b892d91f022f · inbound
TwinRL: Digital Twin-Driven Reinforcement Learning for Real-World Robotic Manipulation GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 904c8a16-d0a8-4b19-8bae-3e0284a62a0e · inbound
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42866686-c264-452a-a320-63705cc5313b · inbound
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e7e21ad5-2002-4d86-9cfb-3157990b3a32 · inbound
LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5605a0c1-a662-42a1-acba-56c0b3227373 · inbound
Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8446b9d6-3128-4dda-b251-182de97f9328 · inbound
Learning While Deploying: Fleet-Scale Reinforcement Learning for Generalist Robot Policies GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 14c156f9-c290-4a3b-b2ff-a42e8914e864 · inbound
ProcVLM: Learning Procedure-Grounded Progress Rewards for Robotic Manipulation GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 406737e0-8593-4ee5-98f8-7b550cf92ba6 · inbound
RePO-VLA: Recovery-Driven Policy Optimization for Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5e16325d-038d-4ed9-998a-680fd90a8f0c · inbound
From Imagined Futures to Executable Actions: Mixture of Latent Actions for Robot Manipulation GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c38a1916-b160-4388-893a-8d8bf55445c7 · inbound
DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 191
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5114bd6a-3171-4846-b807-df20d8a2dfa4 · inbound
PAPO-VLA: Planning-Aware Policy Optimization for Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9a8982b7-5c89-4284-bcae-25bde5ea9884 · inbound
Position: Good Embodied Reward Models Need Bad Behavior Data GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 02a91f91-65be-4ab1-a6f7-ca0aaa2d4ece · inbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7f8b034d-70c4-4bbb-b528-7fe26894c452 · inbound
UniIntervene: Agentic Intervention for Efficient Real-World Reinforcement Learning GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 627851db-f62c-4f08-9b3e-9bb97f8af871 · inbound
Foresight: Iterative Reasoning About Clues that Matter for Navigation GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f7277805-4317-4e28-9e61-fc005ca76c40 · inbound
Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62cfe2bf-a65b-4d10-9fe0-c2e626669ab7 · inbound
SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dd067f52-1652-4806-8e97-4d554a3ac16d · inbound
SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3e82a913-f3bb-4323-ad4c-b10b8512e9dc · inbound
HiL-ResRL: A Model-Agnostic Finetuning Adapter via Human-in-the-loop Residual Reinforcement Learning GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 37132ba4-8cb9-454d-87ad-233ba7680559 · inbound
Learning Process Rewards via Success Visitation Matching for Efficient RL GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 276ca8c9-6be5-404b-8e94-6c71abf6c499 · inbound
ROAD-VLA: Robust Online Adaptation via Self-Distillation for Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f7e59c4a-1b10-4957-ab3e-fd326e41c4b5 · inbound
Trust Your Instincts: Confidence-Driven Test-Time RL for Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5c19d9a5-74d1-4f82-8f75-e53ecebe2fc2 · inbound
Rethinking Foundation Model Collaboration: Enhancing Specialized Models through Proxy Task Reasoning GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 534908e9-3bd8-4b9e-a92e-584e4d8aa8a2 · inbound
Freeform Preference Learning for Robotic Manipulation GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2d48f57a-1bc6-4b95-bcde-dd18adfd100a · inbound
Freeform Preference Learning for Robotic Manipulation GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b5e5b47-3b10-4bcf-8972-443f065ab236 · inbound
Freeform Preference Learning for Robotic Manipulation GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f0f7346-005b-42ac-ba88-c2df11dc3a3b · inbound
RL$^2$-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8247ecf9-282f-4cea-99d9-e682e690fe3f · inbound
RL$^2$-VLA: Adaptive RL Latent Compositional Steering with Test-Time Scaling for Vision-Language-Action Models GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a95f753a-b662-479e-bbb8-80748f72cebb · inbound
RedFlow: Redirect Failure into Action-Level Corrections for Flow-matching VLA Policy GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97963ac0-26fe-4a7e-b568-b59535017128 · inbound
Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.