Pith. sign in

Paper Citation Record · LEDGER

Visual Access Boundaries in Vision-Language Model Reasoning

As of 21 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.12815.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.12815 v2

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-02T06:24:29.865289Z

measured 22 of 22 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4ab637ef-2e2c-4309-8082-436d893e5348 · outbound

This paper cites Hidden in plain sight: VLMs overlook their visual representations.

Visual Access Boundaries in Vision-Language Model Reasoning Hidden in plain sight: VLMs overlook their visual representations

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:27.325903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:27.325903Z digest=sha256:342c79472275cd3a0d934f77f012cfe4af7aa5ca15afe7cfbedb870e0da3d04a

Observation 86b99209-9e10-414b-ada0-8b73ec0fbf6f · outbound

This paper cites The Llama 3 Herd of Models.

Visual Access Boundaries in Vision-Language Model Reasoning The Llama 3 Herd of Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:27.684742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:27.684742Z digest=sha256:64aef2c4fd1c313f0858e4ded462992ff8d5267a245873efd69daff531526ac3

Observation 2f069232-2fc6-490e-bc39-ed4e1bfda74f · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Visual Access Boundaries in Vision-Language Model Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:28.082474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:28.082474Z digest=sha256:4fd7ac5c6a1c1c3efa48ae5ec624e7747117dfb2d4c2b88c671df8a158abd6af

Observation 5e7db585-cffd-40c5-bfc0-4fc0913ee509 · outbound

This paper cites Investigating Inference-time Scaling for Chain of Multi-modal Thought: A Preliminary Study.

Visual Access Boundaries in Vision-Language Model Reasoning Investigating Inference-time Scaling for Chain of Multi-modal Thought: A Preliminary Study

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:28.174746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:28.174746Z digest=sha256:83d5aca5c4181fb570130497a3bedd1d96ba6b3754b8469e9942d844ff32de18

Observation 2d690a40-c96d-402f-a155-3180290068c3 · outbound

This paper cites Rethinking the Role of Prompting Strategies in LLM Test-Time Scaling: A Perspective of Probability Theory.

Visual Access Boundaries in Vision-Language Model Reasoning Rethinking the Role of Prompting Strategies in LLM Test-Time Scaling: A Perspective of Probability Theory

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:28.324742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:28.324742Z digest=sha256:de49a976032fa9a5c7048f18fd9d0c12c11a44e5f80763bdd06609d4db5614d4

Observation 3d0856dd-7956-4a5a-b24f-4a218373a933 · outbound

This paper cites Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting.

Visual Access Boundaries in Vision-Language Model Reasoning Language Models Don't Always Say What They Think: Unfaithful Explanations in Chain-of-Thought Prompting

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:28.514743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:28.514743Z digest=sha256:9b277e89c953d70a84befd2b9983737246818e468bd6cb9fcce0cf5bc5e6290b

Observation 2ee309ca-7383-4054-bc26-01581fd9f8c8 · outbound

This paper cites an unresolved cited work.

Visual Access Boundaries in Vision-Language Model Reasoning Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:28.634750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:28.634750Z digest=sha256:129b94b0c83df528a4c7116cab200aa98d123bb4010e5601dc59df6665d6a05d

Observation d07fa59e-6f8d-4f08-aef1-664a84ff5b83 · outbound

This paper cites Grounded Chain-of-Thought for Multimodal Large Language Models.

Visual Access Boundaries in Vision-Language Model Reasoning Grounded Chain-of-Thought for Multimodal Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:28.750079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:28.750079Z digest=sha256:a18b4e23c1ac33233604175b0b1263ec0d57b40807a23a3f03cdbb664e35675e

Observation 1cccbf46-ee10-4923-897e-93f11be2c53e · outbound

This paper cites IKOD: Mitigating Visual Attention Degradation in Large Vision-Language Models.

Visual Access Boundaries in Vision-Language Model Reasoning IKOD: Mitigating Visual Attention Degradation in Large Vision-Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:28.914749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:28.914749Z digest=sha256:70c25010fd3154717937403b50bb065755541bfe3e67a1503c9ea4ed1b1070b7

Observation c3b6a2db-9f62-4656-9d02-1af7581f551f · outbound

This paper cites Introducing Visual Perception Token into Multimodal Large Language Model.

Visual Access Boundaries in Vision-Language Model Reasoning Introducing Visual Perception Token into Multimodal Large Language Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:29.245780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:29.245780Z digest=sha256:d43e0a9d02c3098f10216d23c77e407ea43d5edc553dfd0438e804acd1219bdf

Observation e7c1a778-abdf-4e96-b4c4-da4ec616abda · outbound

This paper cites an unresolved cited work.

Visual Access Boundaries in Vision-Language Model Reasoning Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:29.393679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:29.393679Z digest=sha256:89b746cbf8dffd904ca2f950379183ad9280527ec9e3b3a2973693a86f249561

Observation 8748d5d4-adf1-4b04-80ef-6a9e769d54bd · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

Visual Access Boundaries in Vision-Language Model Reasoning InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:29.537189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:29.537189Z digest=sha256:2173b0c4eebdbdcc013ecc1258cf27adf27393301e3f960971fc9150cc3e099a

Observation 460f5360-033a-4eba-911c-8e7c45aee9cd · outbound

This paper cites Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models.

Visual Access Boundaries in Vision-Language Model Reasoning Look Twice Before You Answer: Memory-Space Visual Retracing for Hallucination Mitigation in Multimodal Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:29.678069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:29.678069Z digest=sha256:03f059aa5e1647b027df083916c16c88bcc3c377e5c1a35e16a3b849f58c5ff1

Observation 6c7f837c-c7e0-4bfb-ad80-8590771416c0 · outbound

This paper cites We denote the corresponding CoT-side boundaries by ℓ∗ CoT and ℓ∗ DA, respectively, and report shifts relative to the Direct boundaryℓ∗ D following Eq.

Visual Access Boundaries in Vision-Language Model Reasoning We denote the corresponding CoT-side boundaries by ℓ∗ CoT and ℓ∗ DA, respectively, and report shifts relative to the Direct boundaryℓ∗ D following Eq

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:29.865289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:29.865289Z digest=sha256:4a7d64e8772d0675642959798fbc30296de4a163175dce659711d93688354764

Observation cd31ca40-85cc-4072-a324-a434d63366c2 · outbound

This paper cites Understanding intermediate layers using linear classifier probes.

Visual Access Boundaries in Vision-Language Model Reasoning Understanding intermediate layers using linear classifier probes

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:26.784742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:26.784742Z digest=sha256:bc6c9515610b6d11cd47ecdfe494d29984f8fb67ec4f779e7f8d4186b02ac324

Observation 11a64650-03dd-4d43-bce7-7c9b4a507792 · outbound

This paper cites CLEVR: A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning.

Visual Access Boundaries in Vision-Language Model Reasoning CLEVR: A Diagnostic Dataset for Compositional Language and Elementary Visual Reasoning

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:27.792154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:27.792154Z digest=sha256:af30e0549bb9bacc0a85014c3c8864beca2ef061f4a8385f847af3a0c2d7be53

Observation e9573295-1f5c-4545-a22b-ab3d0cd93718 · outbound

This paper cites an unresolved cited work.

Visual Access Boundaries in Vision-Language Model Reasoning Unresolved cited work

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:27.434878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:27.434878Z digest=sha256:3d36f059db6668b0903e6a0f9c710112fdcc84ac95a1eb4616334f45f12d1faf

Observation c4a84307-c14f-4b21-a92a-f3c73faa86c2 · outbound

This paper cites Locating and Editing Factual Associations in GPT.

Visual Access Boundaries in Vision-Language Model Reasoning Locating and Editing Factual Associations in GPT

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:28.404744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:28.404744Z digest=sha256:7f087974c2e28916694a0f3fe14ef3517a83a1cf7f612c0135f654a478907f39

Observation 54940c80-306f-4087-8e4c-3938e73a57e5 · outbound

This paper cites Eliciting Latent Predictions from Transformers with the Tuned Lens.

Visual Access Boundaries in Vision-Language Model Reasoning Eliciting Latent Predictions from Transformers with the Tuned Lens

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:27.188310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:27.188310Z digest=sha256:b70cacab274080459a4bfc0296a01ab8b7d901503d979c2675f3a8cbc9caa65d

Observation d8984636-5460-480a-ae89-64911d16dc79 · outbound

This paper cites Qwen2.5-VL Technical Report.

Visual Access Boundaries in Vision-Language Model Reasoning Qwen2.5-VL Technical Report

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:26.894776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:26.894776Z digest=sha256:2336183b96d16d2fe80fd7b80dec276021da6dbad10ec4f8a2b78c9f0afe1c25

Observation 95b3dbb1-3004-40aa-98b3-a93a6e567c4a · outbound

This paper cites Qwen2.5-VL Technical Report.

Visual Access Boundaries in Vision-Language Model Reasoning Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:27.024869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:27.024869Z digest=sha256:d4d4f372965836b93c6eaaf9e59b2f58eae22e56b49ba8cbd8dec819f272d471

Observation 4adfd617-66cd-4bb3-870a-37657d71a07b · outbound

This paper cites Chain-of-Thought Degrades Visual Spatial Reasoning Capabilities of Multimodal LLMs.

Visual Access Boundaries in Vision-Language Model Reasoning Chain-of-Thought Degrades Visual Spatial Reasoning Capabilities of Multimodal LLMs

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-02T06:24:27.922484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T06:24:27.922484Z digest=sha256:46fbd655ff641b9dce602a4b0a09958160d501e1cc01092443eb45a72bf03358

Pith citing papers

No inbound Pith citation observations are available.