Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 73 inbound Pith citation observations for arXiv:2406.10721.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:28:05.321365Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation f1362d0b-4ad5-4183-9415-315b15b6b608 · inbound
A Survey on Vision-Language-Action Models for Embodied AI RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a86b71ee-5115-4234-9f7b-5b5ec4fca855 · inbound
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 112
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2de53a4-3b18-4735-991b-48d8882f0416 · inbound
On the Dual-Use Dilemma in Physical Reasoning and Force RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8be2bea4-2890-496c-bca7-06d60f75503d · inbound
PartInstruct: Part-level Instruction Following for Fine-grained Robot Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation effe07d9-5415-4556-af8c-7c6f59977368 · inbound
OWMM-Agent: Open World Mobile Manipulation With Multi-modal Agentic Data Synthesis RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd915169-45e4-4cce-b1e1-7e8450643794 · inbound
VideoMolmo: Spatio-Temporal Grounding Meets Pointing RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27e4b751-6c31-4737-b6b3-7cb5d2275871 · inbound
Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d130ec4d-70b9-4aee-89bd-8c518839ea01 · inbound
UAD: Unsupervised Affordance Distillation for Generalization in Robotic Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 104
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 822ca968-e8f6-4d95-9ddf-8cea73989b70 · inbound
GENMANIP: LLM-driven Simulation for Generalizable Instruction-Following Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97df0da4-788d-40b8-96b0-9f6b05992343 · inbound
Gondola: Grounded Vision Language Planning for Generalizable Robotic Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2c9f0f7-54d6-4314-9f05-e85f520a24dd · inbound
CodeDiffuser: Attention-Enhanced Diffusion Policy via VLM-Generated Code for Instruction Ambiguity RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 919de7e8-553f-4b0f-a847-4d1c298511a7 · inbound
Hierarchical Vision-Language Planning for Multi-Step Humanoid Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3fd2d4a8-16fb-4e9f-8587-06bfe8fdcf83 · inbound
RoboBrain 2.0 Technical Report RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2920025b-eab7-4974-8b61-b7e3e2db957a · inbound
DreamVLA: A Vision-Language-Action Model Dreamed with Comprehensive World Knowledge RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 137
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c2ef52c-cfd3-4b4d-a73a-02f4766f676c · inbound
Being-H0: Vision-Language-Action Pretraining from Large-Scale Human Videos RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 144
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 799b436f-3e57-409b-abd4-28772b3e6b03 · inbound
ThinkAct: Vision-Language-Action Reasoning via Reinforced Visual Latent Planning RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a47d01f-2053-4c84-9239-7b43e393ea94 · inbound
PASG: A Closed-Loop Framework for Automated Geometric Primitive Extraction and Semantic Anchoring in Robotic Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1883ef15-e4ae-4ef4-a52b-c9eb240603b1 · inbound
AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84fa15c7-d9e8-47b4-a224-0d2076156adc · inbound
Robix: A Unified Model for Robot Interaction, Reasoning and Planning RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edd14926-7aa5-46c8-882f-ee5306c133e8 · inbound
Weakly-Supervised Learning of Dense Functional Correspondences RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47f571ff-d220-4d07-a286-9e0699cbc58e · inbound
O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46605936-c719-4104-b1f6-212c958e9f39 · inbound
FailSafe: Reasoning and Recovery from Failures in Vision-Language-Action Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faddb885-6fc6-4180-a6fc-477c51cb433b · inbound
GAE: Unleashing Physical Potential of VLM with Generalizable Action Expert RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c60c20fb-9034-4501-9924-567e7ebe634a · inbound
InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e2afd202-39a6-4391-aed0-81760dca456f · inbound
LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8e378c4c-a10a-4695-a5af-ca05f8850fde · inbound
SpatialThinker: Reinforcing Scene Graph-Grounded Spatial Reasoning via Dense Rewards RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ff12f55-fc60-4a0c-b496-23efb6822fd4 · inbound
MiMo-Embodied: X-Embodied Foundation Model Technical Report RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2804e13d-11be-4ea6-8a0c-eacf649fd7f2 · inbound
SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa24a293-4403-4269-b033-1a30ab6cf798 · inbound
Towards Spatial Trace with Reasoning in Vision-Language Models for Robotics RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 113
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cd5c50c-af50-4cf1-8e1a-614f4bb1962a · inbound
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 391defdb-eab7-48c5-806a-e3101db63784 · inbound
VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43e46557-d8fa-4ca2-8f09-9299bcc4f620 · inbound
PALM: Progress-Aware Policy Learning via Affordance Reasoning for Long-Horizon Robotic Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 135
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9274b740-af13-46e0-9de9-c52b97d3638b · inbound
Structured Observation Language for Efficient and Generalizable Vision-Language Navigation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 254beb79-cf5c-45dc-af76-9e0c1b5988d7 · inbound
Structured Observation Language for Efficient and Generalizable Vision-Language Navigation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de868af6-064d-48f7-8e51-4de660d3900e · inbound
LAST: Leveraging Tools as Hints to Enhance Spatial Reasoning for Multimodal Large Language Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9e1b85eb-f5c9-4993-a106-14ab9545ac34 · inbound
AnySlot: Goal-Conditioned Vision-Language-Action Policies for Zero-Shot Slot-Level Placement RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6e8dcd6a-1d33-4205-a770-2014147731f1 · inbound
Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation da05135e-e967-4443-9587-68188f6ea59c · inbound
Assessing VLM-Driven Semantic-Affordance Inference for Non-Humanoid Robot Morphologies RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3db3868-3eeb-4fb3-845e-da793759f331 · inbound
Exploring Spatial Intelligence from a Generative Perspective RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5e866b2c-037e-410b-925d-368ac96618d4 · inbound
PRTS: A Primitive Reasoning and Tasking System via Contrastive Representations RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 12f7b1f8-18d8-401a-8b25-6097b7f9a629 · inbound
Robotic Desk Organization: A Multi-Primitive Approach to Manipulating Heterogeneous Objects via Environmental Constraints RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 97c48c6d-4df8-4c1a-86fd-39b946f8bd03 · inbound
VEGA: Visual Encoder Grounding Alignment for Spatially-Aware Vision-Language-Action Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 90163eae-dc20-4ea3-8123-6f0d3136c618 · inbound
X-Imitator: Spatial-Aware Imitation Learning via Bidirectional Action-Pose Interaction RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fb92b2c-e9b2-4941-86a3-45e559b548f3 · inbound
SceneParser: Hierarchical Scene Parsing for Visual Semantics Understanding RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e03f028b-ed1b-44dd-b163-cecab73588cb · inbound
Humanoid Whole-Body Manipulation via Active Spatial Brain and Generalizable Action Cerebellum RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9d8fe995-50e9-4862-8ea5-8197db99b03d · inbound
Afford-VLA: Action-Aligned Visual Planning via Internalized Affordance RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e447a59a-8567-4e38-890b-11f3f06a3558 · inbound
Extending Embodied Question Answering from Perception to Decision RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 16fc1a9d-0a96-405d-833a-d9b5e6981623 · inbound
GEM: Generative Supervision Helps Embodied Intelligence RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 88e654cc-355e-462e-8b12-7f2678198020 · inbound
Wall-OSS-0.5 Technical Report RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f8729099-9003-4783-9eae-7ff1f06f14fa · inbound
Where to Look: Can Foundation Models Reach a Target Viewpoint Through Active Exploration? RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ce2fd4cf-eb14-47a6-b0e6-06898011e70d · inbound
GeoAlign: Beyond Semantics with State-Guided Spatial Alignment in VLA Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a2989cb6-a0c0-4749-8013-d6925f1b1801 · inbound
AxisGuide: Grounding Robot Action Coordinate System in RGB Observations for Robust Visuomotor Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 814a907f-5180-4204-9aa0-d918b211eebe · inbound
Two Bridges, One Pathway: From VLMs to Generalizable VLAs with Embodied Trajectory-Coupled Data RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7e2dfe1e-ff0c-453d-802c-5d0b80d94d4b · inbound
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b3fff410-d358-4348-b8e6-ec8766235b83 · inbound
Embodied-R1.5: Evolving Physical Intelligence via Embodied Foundation Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1361bfc4-c107-491a-9f80-94ee524203d0 · inbound
SVoT: State-aware Visualization-of-Thought for Spatial Reasoning via Reinforcement Learning RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 58227ffd-37e0-4898-a3bd-91655f024069 · inbound
MaskWAM: Unifying Mask Prompting and Prediction for World-Action Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34ca37b2-dfa9-485e-933e-ebf12aacaae3 · inbound
RoboPIN: Grounded Embodied Reasoning via Pinned Chain-of-Thought RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 491565ac-848c-4828-8859-870777043396 · inbound
RoboPIN: Grounded Embodied Reasoning via Pinned Chain-of-Thought RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca5517f8-f184-4605-aa09-f9fb89a051c2 · inbound
ZeroDex: Zero-Shot Long-Horizon Dexterous Manipulation via Multi-View 3D-Grounded VLM Reasoning RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 13ac652e-62d1-4a4f-ac54-99b95a57b51c · inbound
Vesta: A Generalist Embodied Reasoning Model RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 148
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e56303d-cfbf-4606-852e-e181b9b08c8f · inbound
CVSBench: A Comprehensive Benchmark for Cross-view Spatial Reasoning and Dreaming RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1df1a1d4-2a42-43d7-9246-91b017b79155 · inbound
CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 12c2c2d0-9b1c-4ad9-9380-92a2adc8ed2e · inbound
CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c3eeb3aa-87ef-4a5c-bf9a-778eb1f90d17 · inbound
TAP-VLA: Tactile Annotation Prompting for Vision Language Action Models RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 15278caf-c70f-4e39-af8f-f2e00244729b · inbound
Enhancing Part-Level Point Grounding for Any Open-Source MLLMs RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34828b7b-75a3-440e-bbb2-385c40c67ae5 · inbound
OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 95305c74-c76c-43a8-b68f-eff34a19beb5 · inbound
EgoSteer: A Full-Stack System Towards Steerable Dexterous Manipulation from Egocentric Videos RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa6146c2-7270-49b2-a86e-7391c98c9509 · inbound
Action Map Policy: Learning 3D Closed-loop Manipulation via Pixel Classification RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58d9205b-0127-4f60-bcdb-574a6b05700a · inbound
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2accf09d-3eb9-4550-a3da-5f47d33ff0b6 · inbound
RoboInter1.5: A Holistic Intermediate Representation Suite for Embodied World Modeling and Robotic Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99762039-ea60-47ac-85e5-667f7911b654 · inbound
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 295
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7340192-801b-4d8f-bdc4-a939dbc5ee48 · inbound
BridgeVLA++: A Data-Efficient, Generalizable, and Memory-Augmented Vision-Language-Action Framework for 3D Manipulation RoboPoint: A Vision-Language Model for Spatial Affordance Prediction for Robotics
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.