Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:07:05.277615Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2504.21530.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:07:05.277615Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:57:26.612635Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T00:57:31.251745Z
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f074e8e6-36cd-4847-9400-191526bc441e · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d461be6-308f-4049-b4c4-d4bf31ce7dd9 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors RT-H: Action Hierarchies Using Language
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ef83400-73fb-493f-aebc-6736262fa895 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Track2act: Predicting point tracks from internet videos enables diverse zero-shot robot manip- ulation, 2024
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4836559d-3557-4069-804f-920f4084e771 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Zero-Shot Robotic Manipulation with Pretrained Image-Editing Diffusion Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ce6648c-e311-409c-9a1d-44ef3561d056 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2824a6d8-0541-4ffe-89e4-b040ff4cfe62 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Closed-Loop Visuomotor Control with Generative Expectation for Robotic Manipulation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46e141c7-077d-4b04-9206-4332712dbd01 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Shikra: Unleashing multimodal llm’s referential dialogue magic, 2023
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 69db985e-c967-4495-bffa-78c239e509c5 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Diffusion policy: Visuomotor policy learning via action dif- fusion
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 61ba1ce3-9d6c-4b66-b236-ad5704926dfb · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Instructblip: Towards general- purpose vision-language models with instruction tuning,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb84a066-c793-414a-9616-46a80439ccf0 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Imitating Task and Motion Planning with Visuomotor Transformers
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fb40c13-a3c4-425a-9999-85db878aed17 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Objaverse: A universe of annotated 3d objects
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee4c90f3-db52-4217-9d5d-df1c49c4340b · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors FlowBot3D: Learning 3D Articulation Flow to Manipulate Articulated Objects
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db5e5724-441e-4e18-a801-d1e2dfe57d50 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Anygrasp: Robust and efficient grasp perception in spa- tial and temporal domains
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67075254-6d1b-4dfb-b3a4-b71e2eac5973 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Moka: Open-world robotic manipulation through mark- based visual prompting
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dce6ae1b-891a-4f92-acf5-519d726b0f15 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors RT-Trajectory: Robotic Task Generalization via Hindsight Trajectory Sketches
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c440137-15a8-49d7-8e4e-808e75f09171 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Masked autoencoders are scalable vision learners
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94e26d11-60ca-461b-9501-8ae73a4335c9 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Perceiver: General perception with iterative attention
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2064a1ad-e8e5-4fd1-90f3-81836ef34b23 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Rlbench: The robot learning benchmark & learning environment
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6dcf33ae-c757-4c84-9011-f240efd95410 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors OpenVLA: An Open-Source Vision-Language-Action Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a032ea2-8a3c-4f81-9829-da414ccef191 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Segment Anything
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb65e6f4-5702-4e09-908b-9d2c06834f18 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Lisa: Reasoning segmentation via large language model
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8b745074-d37b-41b5-bb85-14f9fa64895e · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors BLIP-2: bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e331e974-d1ec-41bc-b1a0-a22077e7d56e · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Visual instruction tuning, 2023
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4e6f5ae-d174-4aec-aec3-fd7c21ea9676 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Groma: Localized visual tokenization for grounding multimodal large language models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 03b09d70-72d5-49f7-8dcf-899593fbd439 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors What Matters in Learning from Offline Human Demonstrations for Robot Manipulation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7ad62f4-ff13-4623-9aaf-bdb8c8051f59 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5354ceb-51cd-440b-9a55-78753364ebba · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Calvin: A benchmark for language- conditioned policy learning for long-horizon robot manip- ulation tasks
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 837e1179-462f-4f19-a917-b055c634894f · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1715922-d1dd-4aab-9682-659c4505358a · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Kosmos-2: Ground- ing multimodal large language models to the world
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ddd45b25-6d60-4b09-a78c-13707d575483 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Film: Visual reasoning with a general conditioning layer
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b2365aef-51a9-4f86-b86b-8a121f1917be · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Learning transferable visual models from natural language supervi- sion
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0eefa914-6846-4363-95fe-df26598a7180 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Glamm: Pixel grounding large multimodal model
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4775d94b-b1f0-4892-a54a-ec08ba8c47ff · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Toolflownet: Robotic manipulation with tools via predicting tool flow from point clouds
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ea02a72-57ef-4be3-876c-634121302f77 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Cliport: What and where pathways for robotic manipulation
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c4303925-de3e-4ff7-b20f-e2d8fda771b3 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Open-World Object Manipulation using Pre-trained Vision-Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06adf63a-fb02-466e-917e-62a3edf1ff75 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors KITE: Keypoint-Conditioned Policies for Semantic Manipulation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99ce5b5c-6f34-4606-8ad6-b484b9192006 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Rt-sketch: Goal-conditioned imitation learning from hand-drawn sketches
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 3c951082-1103-41b6-a1fd-93714512b2d4 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa50c0e3-b952-419d-8fb3-7bb6e17d6376 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Robotap: Tracking arbitrary points for few-shot visual imitation
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1f055d2a-0859-4feb-8b4d-d4d2e90921cb · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors GenSim: Generating Robotic Simulation Tasks via Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4c29b91-de0c-400b-b238-62856187c9ae · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors The All-Seeing Project V2: Towards General Relation Comprehension of the Open World
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c87fe91-406e-482f-9f33-dd3d4e553154 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Any-point Trajectory Modeling for Policy Learning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38f1035b-a4ea-4dd2-91ed-9445a9c65e77 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Unleashing Large-Scale Video Generative Pre-training for Visual Robot Manipulation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f07915e-3277-4d46-a68e-a438cba7f124 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Florence-2: Advancing a unified representation for a variety of vision tasks
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 029ea9f6-7685-4cf3-9bce-2e865967315b · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Flow as the cross-domain manipulation interface
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ecbbc05a-5278-44fc-8150-7e04daaf905a · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Ferret: Refer and Ground Anything Anywhere at Any Granularity
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7e61a4b-e04a-4e2f-8d5c-e7aee350d623 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors General Flow as Foundation Affordance for Scalable Robot Learning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 169caaa6-595d-4751-b87d-a9bd1b76cd8c · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Ferret-v2: An Improved Baseline for Referring and Grounding with Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c18a2de-149d-4851-8d80-6174614df944 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Sprint: Scalable policy pre-training via language instruc- 10 tion relabeling
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6872fe4e-93ad-48de-b552-228b094bfe3d · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Learning Fine-Grained Bimanual Manipulation with Low-Cost Hardware
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ae6efb6-9784-4b30-86a4-c8cf4450b3b9 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c314359-1979-466b-9693-1c318cd8b848 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c5833f51-b78b-47a2-9536-ba3c34ab632e · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Unresolved cited work
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 71b8116e-2fb6-44ad-bc13-402a9bbaa1d2 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a133c291-10b0-4dfa-b7c8-e906210a3e78 · outbound
RoboGround: Robotic Manipulation with Grounded Vision-Language Priors yellow surface with brown spots
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7c29c10a-ade1-4c72-91e6-45a7d04c8a88 · inbound
AntiGrounding: Lifting Robotic Actions into VLM Representation Space for Decision Making RoboGround: Robotic Manipulation with Grounded Vision-Language Priors
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.