Pith. sign in

Paper Citation Record · LEDGER

Where do Large Vision-Language Models Look at when Answering Questions?

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2503.13891.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.13891 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:28.465442Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:58:58.571629Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0cfe0c03-bf00-4cff-95a9-3fae186b43fb · inbound

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations cites this paper.

Seeing It or Not? Interpretable Vision-aware Latent Steering to Mitigate Object Hallucinations Where do Large Vision-Language Models Look at when Answering Questions?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:28.465442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:28.465442Z digest=sha256:6a95cd8e0c98cab907e7cd2c42ed1bd6f845bb824a2a17a54bba3b0f85d2734a

Observation 3d47b392-f26b-49fe-8ba2-fc7ef011a951 · inbound

LaSM: Layer-wise Scaling Mechanism for Defending Pop-up Attack on GUI Agents cites this paper.

LaSM: Layer-wise Scaling Mechanism for Defending Pop-up Attack on GUI Agents Where do Large Vision-Language Models Look at when Answering Questions?

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:12:02.136028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-19T04:10:57.882345Z digest=sha256:ab65a0fb8fbc82ea5aa8d4e52cf0803e0c57e652251cfab0a15e3f7e84b14ee0

Observation 92142dd0-425e-426d-a1ee-7514f56f1e1e · inbound

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models cites this paper.

V-SEAM: Visual Semantic Editing and Attention Modulating for Causal Interpretability of Vision-Language Models Where do Large Vision-Language Models Look at when Answering Questions?

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:21:36.561269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-18T16:21:20.463222Z digest=sha256:c9c857a04561b26f8170b7878915702c89fcefbec13b606f691052f11efd8c1a

Observation 630cfa6c-1d13-473c-b4b5-4b452a1cc883 · inbound

Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making cites this paper.

Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making Where do Large Vision-Language Models Look at when Answering Questions?

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T06:28:45.699007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:28:45.699007Z digest=sha256:a64c096fefb54a209c81e0800716a2861ad037fe1713849828cdd26b5b44a596

Observation 751baa03-d9ea-44f1-878a-a470a9521866 · inbound

Measuring Cross-Modal Synergy: A Benchmark for VLM Explainability cites this paper.

Measuring Cross-Modal Synergy: A Benchmark for VLM Explainability Where do Large Vision-Language Models Look at when Answering Questions?

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:06:08.717537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T06:05:34.039507Z digest=sha256:44f3eec67e3a3b6464f7aa0407a5cccc8ea6251ea0ac75a4ecf1637e6d26c377

Observation 172c0798-7cf1-47eb-a108-0a9df966d1f9 · inbound

PhaseWin: An Efficient Search Algorithm for Faithful Visual Attribution cites this paper.

PhaseWin: An Efficient Search Algorithm for Faithful Visual Attribution Where do Large Vision-Language Models Look at when Answering Questions?

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:58:58.573488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-06-27T00:57:24.855665Z digest=sha256:4864066527f219db47c6e92673f07bc16cb796b6833570650204b4a72084190d