Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 22 inbound Pith citation observations for arXiv:2503.12799.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:34:36.765233Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation f2d0cd6a-a6de-4a82-aeb9-aee55ccb2dbc · inbound
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 2ea0c3d6-bc3a-43b0-b421-951261aefd76 · inbound
Ego-R1: Chain-of-Tool-Thought for Ultra-Long Egocentric Video Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33aafc2c-50ed-4f2a-8985-4d80ee31dcc9 · inbound
Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2774bf3-af99-407d-af29-21f65c5c471d · inbound
Why Do MLLMs Struggle with Spatial Understanding? A Systematic Analysis from Data to Architecture Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51891bc4-500e-4224-9a61-e537c1f404e1 · inbound
A Comprehensive Survey on Trustworthiness in Reasoning with Large Language Models Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17b940df-2077-4fa0-ac82-5b632675a6f1 · inbound
Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 59ee0f4e-8e78-40bf-9884-17d6dddf6adb · inbound
Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0c874784-5226-47e3-b452-a4842a504355 · inbound
VisReason: A Large-Scale Dataset for Visual Chain-of-Thought Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8edea675-28c1-49e9-8a8c-c67c7668c18f · inbound
Boosting RL-Based Visual Reasoning with Selective Adversarial Entropy Intervention Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d3f4706-9dff-476e-8931-6d15122eac37 · inbound
SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ab30d83a-a8e2-433e-87cb-8c060130f687 · inbound
SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4a50d179-76ff-46a8-bce4-4ffcff692ba5 · inbound
CAVE: A Structured Credit Assignment Approach for Fragmented Visual Evidence Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 05970f97-2491-4dfa-be92-e26319b89bde · inbound
Vision Harnessing Agent for Open Ad-hoc Segmentation Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c014bab5-6996-423b-b930-9bf9c6509f2e · inbound
Rethinking Visual Attribution for Chest X-ray Reasoning in Large Vision Language Models Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b4dbc6f8-3b17-445e-92ae-457a0d065b00 · inbound
ProSR: Process-Shaped Spatial Reasoning for Reliable Chain-of-Thought in VLMs Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 85f17c59-249d-408c-acc2-0d3fe13fe54e · inbound
ROVER: Routing Object-Centric Visual Evidence for Grounded Multi-Image Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c3d27d2b-00ff-46ed-bcdf-653d8dd7df31 · inbound
Mags-RL: Wearing Multimodal LLMs a Magnifying Glass via Agentic Reinforcement Learning For Complex Scene Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 65dd6ab2-5237-4ff1-9f72-0d32f2290ac7 · inbound
Grounded 3D-Aware Spatial Vision-Language Modeling Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d9c0e8e4-f5d4-4704-9801-13007a9e5f77 · inbound
V-Zero: Answer-Label-Free On-Policy Distillation with Contrastive Evidence Gating for Fine-Grained Visual Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d07fa59e-6f8d-4f08-aef1-664a84ff5b83 · inbound
Visual Access Boundaries in Vision-Language Model Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c1400db-a6ca-4bb0-afa2-ed604e0e375d · inbound
OPLD: On-Policy Latent Distillation for Multimodal Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23fadc11-0a1a-41f7-9d4d-9088468e9795 · inbound
Balancing Efficiency and Efficacy: Training-Free Attention-Guided Switching Between Explicit and Latent Thoughts for MLLMs Grounded Chain-of-Thought for Multimodal Large Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.