Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 33 inbound Pith citation observations for arXiv:2401.06209.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:32.061891Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation acac3d1d-5e78-403e-9081-b9765ca49f05 · inbound
DeepSeek-VL: Towards Real-World Vision-Language Understanding Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 939e1045-4e5a-4cc2-8d49-572279dcc5e2 · inbound
MM1: Methods, Analysis & Insights from Multimodal LLM Pre-training Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 108
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 62c29b44-7eb0-47e4-91b2-6be587e15c2a · inbound
How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 111
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6c02f745-d645-4819-9829-623ccc82d368 · inbound
Hallucination of Multimodal Large Language Models: A Survey Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 156
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e4e60157-5b8f-4138-bf26-7e5b1f2d07bd · inbound
Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91e87091-4d6a-46ae-b4f2-9f348ae857a9 · inbound
AMIA: Automatic Masking and Joint Intention Analysis Makes LVLMs Robust Jailbreak Defenders Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 514944c9-b9b9-427f-99fb-7d2e0502644a · inbound
Humanoid World Models: Open World Foundation Models for Humanoid Robotics Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bafe4089-8d7e-40e5-aae8-f71239d64ec9 · inbound
MoDA: Modulation Adapter for Fine-Grained Visual Grounding in Instructional MLLMs Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4d2c4f5-7971-4722-9fd1-74b1712a1143 · inbound
Unfolding Spatial Cognition: Evaluating Multimodal Models on Visual Simulations Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03aea609-f8a5-465a-bd29-24afa8a79559 · inbound
Backdoor Attack on Vision Language Models with Stealthy Semantic Manipulation Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3698515-14c1-45aa-a30f-59ac2fb2f97a · inbound
A Good CREPE needs more than just Sugar: Investigating Biases in Compositional Vision-Language Benchmarks Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffedd751-f1f1-40a0-92bd-419d0354ae2e · inbound
How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation fa75da25-bfe4-4e13-822e-b46c35599431 · inbound
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1480ea88-014d-41e3-91a9-1e552c22266b · inbound
Understanding Space Is Rocket Science -- Only Top Reasoning Models Can Solve Spatial Understanding Tasks Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd397fc2-4ae1-42b8-ba90-65c061c5ff18 · inbound
Agentic Learner with Grow-and-Refine Multimodal Semantic Memory Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 61285ced-81ef-45f4-aa01-93533de21893 · inbound
Kimi K2.5: Visual Agentic Intelligence Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d7fc9b8a-1cc8-4058-8cb7-618ac66d9d06 · inbound
Visual Para-Thinker: Divide-and-Conquer Reasoning for Visual Comprehension Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8d35bc86-3a3c-4cfc-be9f-1b1a390b3d28 · inbound
Revisiting Model Stitching In the Foundation Model Era Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0d5d9e9-0238-4fb9-bc1b-1f2f903be8a2 · inbound
Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 88b072a1-0dd7-4ab6-9610-fcb7454daba0 · inbound
Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52b91b38-6d68-4b2f-a4a7-ca77ac00eb1f · inbound
Gaslight, Gatekeep, V1-V3: Early Visual Cortex Alignment Shields Vision-Language Models from Sycophantic Manipulation Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1cea5ba8-9582-406c-91f6-83972d727325 · inbound
SinkRouter: Sink-Aware Routing for Efficient Long-Context Decoding in Large Language and Multimodal Models Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b574e902-cc1f-4634-8f37-eccd69075fc3 · inbound
The Expense of Seeing: Attaining Trustworthy Multimodal Reasoning Within the Monolithic Paradigm Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5bf38efe-c0b3-4175-ae13-d0f1cb80751c · inbound
The Expense of Seeing: Attaining Trustworthy Multimodal Reasoning Within the Monolithic Paradigm Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4ad5867c-504a-44f8-be22-bb2dda432282 · inbound
HypEHR: Hyperbolic Modeling of Electronic Health Records for Efficient Question Answering Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 251
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e12fcc5e-a825-45fb-be72-40741c6758aa · inbound
VisualNeedle: Benchmarking Active Visual Search in Information-Dense Scenes Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 589070f9-af55-4bd1-94c9-a01ecd1e9b50 · inbound
VESTA: Visual Exploration with Statistical Tool Agents Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0d379c7f-0032-4471-934d-5751010ed5f5 · inbound
Fine-tuning Multi-modal LLMs with ART: Art-based Reinforcement Training Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 72f3f772-e9ff-400a-af7d-525651eb08f5 · inbound
Bridging the Usability Gap: Lessons from Interpreting Studies for Machine Interpreting Design Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 79206ccc-5d08-40bb-bc08-6dacc2b386d8 · inbound
Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1c4639ee-9939-43cf-8533-2a3508e43a68 · inbound
DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c1a1b0b-26a0-4323-8fdc-0a01a5581cfa · inbound
VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e51ee67-2932-4e6c-a1ff-4b3e0843eb3e · inbound
The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping Eyes Wide Shut? Exploring the Visual Shortcomings of Multimodal LLMs
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.