Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 39 inbound Pith citation observations for arXiv:2312.14135.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:29:19.034618Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
1
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation c4752a49-937d-4f37-8801-70f4bc348d93 · inbound
A Survey on Multimodal Large Language Models V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 208
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5d2ab2cf-19a8-45c1-8e0d-6a1e5c44ce3f · inbound
PaliGemma: A versatile 3B VLM for transfer V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 147
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1d2ef04f-9865-412b-9b35-1307004e607b · inbound
Seed1.5-VL Technical Report V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 149
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5684eb9b-6dd6-4071-a5a2-23d709b86867 · inbound
Grounded Reinforcement Learning for Visual Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 71f3e515-41de-4d7a-b0bf-ad75ebc445e9 · inbound
Unifying Language Agent Algorithms with Graph-based Orchestration Engine for Reproducible Agent Research V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7599ee8d-24d9-4414-ba22-ad3aa2d74bf5 · inbound
MiMo-VL Technical Report V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 858d4b93-a6c5-4dbe-9a08-d68c203ae2d6 · inbound
Unfolding Spatial Cognition: Evaluating Multimodal Models on Visual Simulations V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3c446d9-6703-4203-825c-4baea7897aba · inbound
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e64e0d55-fe91-403e-8dc5-24610c1afd0b · inbound
Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a052efc3-2ef7-4196-8175-27680ac2b9d1 · inbound
HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e7edd0ef-5bc5-42cd-8a8d-baaa308ee84a · inbound
HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bead9c2-52c2-44ea-9125-0979bab87762 · inbound
Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 978dc477-32ba-4afb-8e4c-f600cbfa318b · inbound
Blink: Dynamic Visual Token Resolution for Enhanced Multimodal Understanding V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92a875fa-dc60-4fdb-91e4-ed1ceac05f94 · inbound
Visual Para-Thinker: Divide-and-Conquer Reasoning for Visual Comprehension V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1c20257b-e448-47cf-bad4-28b558dd828a · inbound
DLEBench: Evaluating Small-scale Object Editing Ability for Instruction-based Image Editing Model V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 1e1dfc9c-74da-4b71-a0dc-5347b54124fa · inbound
SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc98f7b9-99a6-49ed-a866-90b9366f1cf1 · inbound
LanteRn: Latent Visual Structured Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40ed16f7-83f2-455c-a1fc-4d451d3d4b24 · inbound
Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 7eb14478-c1d8-4d8c-9be8-ebd448c8b100 · inbound
Multimodal Latent Reasoning via Predictive Embeddings V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d99eb290-4d03-4448-8e02-44676820aed7 · inbound
Entropy-Gradient Grounding: Training-Free Evidence Retrieval in Vision-Language Models V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 83b8a0dd-9add-4223-8527-8a5659dfeeff · inbound
MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d3c74dd9-a762-4006-bd93-04f769f35297 · inbound
Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 983d9ea9-53f7-41d6-abae-b8d9662b0c16 · inbound
Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 814f6ec7-7150-4d8d-9bf2-18e47e7f8447 · inbound
AnchorSeg: Language Grounded Query Banks for Reasoning Segmentation V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 213
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ea694479-978c-4604-9b46-9c67c71b600d · inbound
SketchVLM: Vision language models can annotate images to explain thoughts and guide users V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d44ab193-c2c8-4f94-9ecf-b6b415bbd7b6 · inbound
Improving Vision-language Models with Perception-centric Process Reward Models V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4d1a14b4-c409-4d27-8502-8d53df71a759 · inbound
GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ccad97df-3074-475e-b73f-7de3b8ff1d3e · inbound
What's Holding Back Latent Visual Reasoning? V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e89677b3-b613-488c-9467-a4394c408300 · inbound
Starve to Perceive: Taming Lazy Perception in VLMs with Constrained Visual Bandwidth V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a867bb2b-1d00-4b86-9775-0b198c98424d · inbound
Starve to Perceive: Taming Lazy Perception in VLMs with Constrained Visual Bandwidth V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5408f8a-eebe-4ad9-a88a-24da1b98dd3f · inbound
Self-Prophetic Decoding to Unlock Visual Search in LVLMs V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 22009e46-56ed-46c6-ab6a-b1cf6c9a5d9a · inbound
ToolGate: Token-Efficient Pre-Call Control for Tool-Augmented Vision-Language Agents V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 79c455c0-6e7c-4123-848c-947c636d6fd2 · inbound
MOSS-Video-Preview: Toward Real-Time Video Understanding via Cross-Attention V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 02a06411-4ac5-46f5-9891-ef32db2a3b91 · inbound
Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8497a1cf-e8ba-4d0d-a171-6b6a1eee502d · inbound
Look Before You Zoom: Adaptive Routing for the Resolution-Context Trade-off in Visual RAG V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3b1d8caa-5deb-486b-8c29-6611a5ad14b9 · inbound
Kamera: Unified Position-Invariant Multimodal KV Cache for Training-Free Reuse V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e17be238-b3bf-460a-8f63-3713f499ddcd · inbound
ST-Veto: Spatio-Temporal Token Veto for Diffusion MLLMs via Taylor Prediction and Visual Grounding V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54c1de77-47be-4752-be6a-b354061bc3c3 · inbound
VAD: Attributing Visual Evidence for Target Reconstruction in Multimodal On-Policy Distillation V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c555d85-f35f-4299-851a-8fcf9e14a703 · inbound
What Transfers from Text to Vision? Capability Scaling Laws and Transfer Dynamics for VLMs V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.