Pith. sign in

Paper Citation Record · LEDGER

Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2305.02317.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.02317 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:10:11.266165Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T23:49:02.294828Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8d2a8bfd-bce7-4e65-abcb-f1cb399801d9 · inbound

A Survey on Multimodal Large Language Models cites this paper.

A Survey on Multimodal Large Language Models Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 187

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:56:42.155815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T02:56:41.658658Z digest=sha256:1ee73a956f4888a87622ccc5012a2ba08341367923535515e733a5b85689f77f

Observation 6a9d2b29-e7e5-4a5f-aade-acd9df9586fa · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.689340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:965659b43fbd59184f254caad6b4c0cc31147b7f5849b8a945a13f78982b211f

Observation 01fe981e-0f3a-4d7e-97af-0a6f31164355 · inbound

CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models cites this paper.

CoT-VLA: Visual Chain-of-Thought Reasoning for Vision-Language-Action Models Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:21:44.988534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T05:21:44.903048Z digest=sha256:aca3d43a9ccec24f0c7fad1fee7f1b518778f9b49b9105a0bb1364a4277f3782

Observation 824d62e9-3f98-438e-9cc0-bbd5e9de4f34 · inbound

VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought cites this paper.

VLM-R$^3$: Region Recognition, Reasoning, and Refinement for Enhanced Multimodal Chain-of-Thought Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:10:11.266165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:10:11.266165Z digest=sha256:2b575d57039c3d9835b8161f269a6e6909d21c7b56617779b072caca248dbd34

Observation 2747b5cd-31f1-4a9a-a7d0-9b27b197983b · inbound

ReFineVLA: Reasoning-Aware Teacher-Guided Transfer Fine-Tuning cites this paper.

ReFineVLA: Reasoning-Aware Teacher-Guided Transfer Fine-Tuning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:23:17.734824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:23:17.734824Z digest=sha256:e33eebfb8f8c86144002a733e2019095ab17ece35a104f384c6d40f99aadf3e8

Observation 251461e2-d61c-43f5-9dff-f1df33622f8f · inbound

Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning cites this paper.

Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:05.316112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:05.316112Z digest=sha256:75357891bb6e82790429b40834a4c6b374d9de1a9bf0c75d5c7a1f398065974c

Observation 991f54e3-8231-4670-bdee-c8beff1d6ab8 · inbound

Argus: Vision-Centric Reasoning with Grounded Chain-of-Thought cites this paper.

Argus: Vision-Centric Reasoning with Grounded Chain-of-Thought Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:25.021502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:25.021502Z digest=sha256:47105b1dc66ab6f90bbf29d8e403ff4e4202d280ffd3906605269b3b9e22d85c

Observation d6973e5d-3d08-4214-bbe4-a5764c3ead37 · inbound

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs cites this paper.

ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T04:40:10.790536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:40:10.790536Z digest=sha256:29cc4787f01f2758f4cd9f338057195ab8c60a3fa37f9ecd1a21e2513a436818

Observation 9515bdd8-7253-490f-8bd9-b5d53f00a0c8 · inbound

Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? cites this paper.

Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T11:19:02.751761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:19:02.751761Z digest=sha256:20c611d56abfb3d3fa6506f9cfc269d01dcb56ab429d0fdb3d0746415335c96e

Observation e1e41564-69ac-402c-b641-dd779435f174 · inbound

MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning cites this paper.

MagiC: Evaluating Multimodal Cognition Toward Grounded Visual Reasoning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:05.057773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:51:05.057773Z digest=sha256:5ff11605e01efd66d0dfc2367b82bb1389accf748dbd6c5ae13a5500599c2554

Observation 1d9134d9-2d67-4ef6-8865-6454a3103638 · inbound

WebWatcher: Breaking New Frontier of Vision-Language Deep Research Agent cites this paper.

WebWatcher: Breaking New Frontier of Vision-Language Deep Research Agent Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T18:56:23.987203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T18:56:23.817544Z digest=sha256:7e1b765c97d678b140d2ee202e04060e2bb7ff44afbb20a40a89f008122ae42f

Observation ad075480-7c56-48d2-b986-c70603077360 · inbound

Less is More Tokens: Efficient Math Reasoning via Difficulty-Aware Chain-of-Thought Distillation cites this paper.

Less is More Tokens: Efficient Math Reasoning via Difficulty-Aware Chain-of-Thought Distillation Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T05:34:40.348235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:34:40.348235Z digest=sha256:6c675e7c6e371754c108bdb1832453bc49ca798b5b77d6a5569e94798109af1e

Observation bcaae467-e363-4982-b444-879b2eaf533f · inbound

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework cites this paper.

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:32:36.347236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-18T12:31:25.257879Z digest=sha256:b7e99a04073c130ac10f9d9187f7b7f6ff6b8e824163cab69e4dc41e50202098

Observation c9501840-9dc0-4a57-8174-2ed738536823 · inbound

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning cites this paper.

ReFineVLA: Multimodal Reasoning-Aware Generalist Robotic Policies via Teacher-Guided Fine-Tuning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:48:48.348511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T05:06:38.517652Z digest=sha256:e00a4bb84846e0aaaaca881b38659f7d746dccc9d0b9725d588d9996f6c868a1

Observation ab011063-c40e-4571-a5aa-9352f7f6effe · inbound

R-CoV: Region-Aware Chain-of-Verification for Alleviating Object Hallucinations in LVLMs cites this paper.

R-CoV: Region-Aware Chain-of-Verification for Alleviating Object Hallucinations in LVLMs Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:04.240674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T01:28:14.039910Z digest=sha256:246a6edd33f7cb74a18601789f87fdb653cc454ba7b060c0294c5db93fdc6816

Observation 595f3cef-e4df-4b0c-95c9-73f5a1ade0d6 · inbound

Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning cites this paper.

Visual-OPSD: Cross-Modal On-Policy Self-Distillation for Efficient Unified Multimodal Reasoning Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T23:49:02.296271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T21:47:51.437284Z digest=sha256:5f0ef6cab40f3ee16607e205b5b78cf94dc4c7f5d1f202a2d577ed9e4ba2f600

Observation 1e0eeb18-b51d-4fbc-9a48-fd58cc5f6a4c · inbound

ReShift: Aha-Moment-Driven Reasoning-Level Backdoor Attacks on Vision-Language Models cites this paper.

ReShift: Aha-Moment-Driven Reasoning-Level Backdoor Attacks on Vision-Language Models Visual Chain of Thought: Bridging Logical Gaps with Multimodal Infillings

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T11:56:54.961436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-02T11:48:35.972372Z digest=sha256:3f3a240782f9eb6d659ebe082df8796b4eb52a583e89803d59dcd730b790bdba