Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:50:09.623638Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2507.12391.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T16:50:09.623638Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T11:40:21.327031Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T11:40:25.831314Z
23 of 23 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 08cbf44a-4b00-4086-b09c-742698c10219 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning A survey on large language model based autonomous agents
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9986f59e-4f4c-475a-965f-7ee8a08b25c0 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Large language models for robotics: Opportunities, challenges, and perspectives
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c65b190b-166d-4fcc-a5f3-a3722221d776 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Human-Robot collaboration in surgery: Advances and challenges towards autonomous surgical assis- tants
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 57623a7e-2b68-4061-abe0-f2ee19db3928 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Language models as zero-shot planners: Extracting actionable knowledge for embodied agents
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cbc1c722-8cdf-4b41-b90a-9a502e50e7a8 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3f2e3fba-f0c5-4623-88d5-2bbb2c9d94e3 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Progprompt: Generating situated robot task plans using large language models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 19a8e19c-e42c-4bdf-ab88-6750cce49d7a · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning V oice con- trol interface for surgical robot assistants
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a94424e7-7caa-4cee-a8cf-f67ecc0d3653 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning LLM-based ambiguity detection in natural language instructions for collaborative surgical robots
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1a42df65-51b0-4524-ba32-23ca0dd7a425 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Toward autonomous robotic minimally invasive surgery: A hybrid framework combining task-motion planning and dynamic behavior trees
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4a39828d-d7e5-453d-88f5-03b56020e6eb · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Exploring embodied mul- timodal large models: Development, datasets, and future directions
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 47faf2df-fb07-430d-ad0a-1957377fb85c · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Multimodal Fusion and Vision-Language Models: A Survey for Robot Vision
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16c1735e-0042-4033-abad-8427410ddd74 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning PaLM-E: An Embodied Multimodal Language Model
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07c40737-2ab5-4210-8b4e-34d859072622 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning RT-2: Vision-Language-Action Models Transfer Web Knowledge to Robotic Control
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c1d1cae-4389-435f-a7d6-0e4ba6a3e83e · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning VIMA: General Robot Manipulation with Multimodal Prompts
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1d0268a-a77f-4a04-9f93-2964c0b7155d · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning LLM-A*: Large Language Model Enhanced Incremental Heuristic Search on Path Planning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ccd6026-280f-407f-892c-740167a13455 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning LLM-Advisor: An LLM Benchmark for Cost-efficient Path Planning across Multiple Terrains
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 622f095f-1e64-4971-a277-1bd0feddc52a · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning LLM-Enhanced Path Planning: Safe and Efficient Autonomous Navigation with Instructional Inputs
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 235df9a0-fdba-4ab2-a74f-909b7a894a4e · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Guide-LLM: An Embodied LLM Agent and Text-Based Topological Map for Robotic Guidance of People with Visual Impairments
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation acc50030-12b0-4839-a227-482856a712ca · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning From Text to Space: Mapping Abstract Spatial Models in LLMs during a Grid-World Navigation Task
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeccf834-057c-4c69-be2d-ae4783029d6c · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Scaling and Beyond: Advancing Spatial Reasoning in MLLMs Requires New Recipes
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6367c078-f26e-4924-bb48-ed3c51739195 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Mitigating Cross-Modal Distraction and Ensuring Geometric Feasibility via Affordance- Guided, Self-Consistent MLLMs for Food Prepara- tion Task Planning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f6edbf4-224d-4902-ae90-0d0f32c2d5d7 · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning Can Large Language Models be Good Path Planners? A Benchmark and Investigation on Spatial-temporal Reasoning
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20a633f0-d5a2-4c3f-93ba-942cbab5df6d · outbound
Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff00f254-9fce-4018-87cd-dad413f06d7e · inbound
Why Do MLLMs Struggle with Spatial Understanding? A Systematic Analysis from Data to Architecture Assessing the Value of Visual Input: A Benchmark of Multimodal Large Language Models for Robotic Path Planning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.