Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T21:16:50.651526Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 36 inbound Pith citation observations for arXiv:2508.09071.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-05T21:16:50.651526Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:19:46.803171Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T17:38:43.940925Z
12 of 12 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 044b3559-17b0-4d81-81da-b5245e869a54 · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models Lift3D Foundation Policy: Lifting 2D Large-Scale Pretrained Models for Robust 3D Robotic Manipulation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f75680b-eb5e-4213-93c0-4d7397703433 · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19dc6192-36bc-4fbf-995e-a2ea50decd90 · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models Flow Matching for Generative Modeling
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0953712d-9267-4135-882f-c3bfcbf75c5f · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a3bf52f-d29f-42f7-a82d-f45e792260a3 · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cc1f00d7-649d-4957-920c-df764718e823 · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models 3D-VLA: A 3D Vision-Language-Action Generative World Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fb6f571-f63e-4466-985d-9a52c58cb6d2 · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models LLaVA-3D: A Simple yet Effective Pathway to Empowering LMMs with 3D-awareness
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8779decc-64d5-427b-a144-f8ce6795cb08 · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models Dita: Scaling Diffusion Transformer for Generalist Vision-Language-Action Policy
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 845a691d-df31-4f89-9d08-7449379b87df · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models Qwen Technical Report
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75752dd1-39ad-493c-888b-7bad0658224e · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models DEGround: An Effective Baseline for Ego-centric 3D Visual Grounding with a Homogeneous Framework
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9c6fa5d-81b5-4426-873a-e0eae9c2790e · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models PointVLA: Injecting the 3D World into Vision-Language-Action Models
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48171647-2b2b-4126-b3c1-14667dfd248c · outbound
GeoVLA: Empowering 3D Representations in Vision-Language-Action Models GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62c69a0d-4cd7-467b-b9d5-3b8d2080edb3 · inbound
MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a014a169-5699-4e65-b9f4-7d9f9f2c16aa · inbound
LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1e1e29c0-0145-4860-8940-ad6e07cf9bed · inbound
LIBERO-PRO: Towards Robust and Fair Evaluation of Vision-Language-Action Models Beyond Memorization GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d85dd0d-8aee-4bdf-875b-303f0b8d2231 · inbound
QDepth-VLA: Quantized Depth Prediction as Auxiliary Supervision for Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71d97c6e-4e7d-4b10-b9ff-f83eedfed65a · inbound
A Pragmatic VLA Foundation Model GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 83ff0ae9-ccf5-42c4-9dee-05e79c17ddda · inbound
E-VLA: Event-Augmented Vision-Language-Action Model for Dark and Blurred Scenes GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7d9bab16-e82e-41f8-93dc-9dcfc1d2a85d · inbound
CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 48bd2dfc-665b-498f-8670-62cde3646a6f · inbound
Robotic Manipulation is Vision-to-Geometry Mapping ($f(v) \rightarrow G$): Vision-Geometry Backbones over Language and Video Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3c4e72ad-2694-47f3-8c08-22e5abc0a37c · inbound
R3D: Revisiting 3D Policy Learning GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e8d9e6ef-2437-443e-abcf-a7a7f98caced · inbound
PokeVLA: Empowering Pocket-Sized Vision-Language-Action Model with Comprehensive World Knowledge Guidance GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 19e63f52-6734-438b-a66a-4cbef4048519 · inbound
Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d57c6f33-91f7-46fa-97f4-0e8527a07f2e · inbound
Unified 4D World Action Modeling from Video Priors with Asynchronous Denoising GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fa371804-0d82-4488-a3c2-42452d784ade · inbound
ConsisVLA-4D: Advancing Spatiotemporal Consistency in Efficient 3D-Perception and 4D-Reasoning for Robotic Manipulation GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 27f97e89-9f07-455e-b991-4f3b9dbc9460 · inbound
VEGA: Visual Encoder Grounding Alignment for Spatially-Aware Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4fcfcc17-bca0-4983-90e1-8e69f46bb3db · inbound
Learning Action Manifold with Multi-view Latent Priors for Robotic Manipulation GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 230cea23-6c8d-4fe1-8421-db1277ed9a20 · inbound
X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation af347835-7fd8-4fa8-b65a-10715cf232bb · inbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 02fd77d4-1ee8-499e-9500-af46074635e6 · inbound
GuidedVLA: Specifying Task-Relevant Factors via Plug-and-Play Action Attention Specialization GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ff736a69-2844-4053-9402-17fddd5d9379 · inbound
PointACT: Vision-Language-Action Models with Multi-Scale Point-Action Interaction GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cdff4942-d152-4ee1-bdc1-672ff08d89e0 · inbound
ELAN4D: Embodiment-Centric 4D Supervision for Vision-Language-Action Models via Plug-and-Play Adaptation GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 774ff56d-cb86-4531-ac0a-e37f7093b0b8 · inbound
Dexterity-BEV: Aligning 3D World and Actions for Generalizable Robot Policies Learning GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f163582c-b0d7-43c9-905a-83044c62c510 · inbound
GeoAlign: Beyond Semantics with State-Guided Spatial Alignment in VLA Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3fcff667-0757-4581-994a-7d30ee88679a · inbound
3DThinkVLA: Endowing Vision-Language-Action Models with Latent 3D Priors via 3D-Thinking-Guided Co-training GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3921b61e-cb8c-4e63-85b0-ca7ec9126dd7 · inbound
Robots Need More than VLA and World Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8e4205c0-1ddb-4ee1-8255-a97cdf6b1ef5 · inbound
MotionVLA: Injecting Geometric Motion into Vision-Language-Action Model GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9143f34f-a725-4658-85f3-38421e58cb5e · inbound
MemoryVLA++: Temporal Modeling via Memory and Imagination in Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation deb79e72-96fd-4fb6-a5db-a938de7c19f2 · inbound
VeriSpace: Spatially Grounded Action Verification for Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a0f12d5f-9ade-463e-ac63-4ee602e7259a · inbound
GeoHAT: Geometry-Adaptive Hybrid Action Transformer for Mobile Manipulation GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 51bb3c66-f216-4640-a45b-a17b4c375d58 · inbound
Geometric Action Model for Robot Policy Learning GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d36d26c3-0f0d-41e6-8210-ae9cac506e7e · inbound
Learning 4D Geometric Priors for Inference-Efficient World Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7dcc86c-a1b0-4b2b-a6fb-29884bb9807e · inbound
See like a Robot: Robot-Centric Pointmaps for Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 856e1487-93bc-4916-9ee1-2bf9a0884165 · inbound
VistaVLA: Geometry- and Semantic-Aware 3D Gaussian-Grounded VLA for Robotic Manipulation GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 049d2ea7-780b-4e41-95b4-8a945fbce450 · inbound
SAM3D-Guided Object-Centric Representation Alignment for Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dd3cd96-9da4-4c72-b559-74a2723c3c96 · inbound
VR3D: View-Robust 3D Representation Learning for Aerial-Ground Person Re-Identification GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db459a88-3610-46fd-9aaa-6187dbaf2234 · inbound
Mind-VLA: Instruction-Aware Spatial Representation Alignment for Vision-Language-Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7df52d0-2bea-422c-a38b-49f06c97be77 · inbound
DreamWAM: Beyond RGB Future Prediction for World Action Models GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.