Pith. sign in

Paper Citation Record · LEDGER

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering

As of 15 August 2026, this Paper Citation Record lists 26 of 26 outbound references and 0 inbound Pith citation observations for arXiv:2507.21335.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21335 v1

Coverage vector

measured 26 of 26 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T12:58:01.385743Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

26 of 26 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved26
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ac221131-ecee-4acd-abb9-c6c86238947d · outbound

This paper cites online" 'onlinestring :=.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.302775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.302775Z digest=sha256:fc998ab4e3e757ba1bc81d3338256c2bc7746e4b78669f10b26ca0000a8274af

Observation 7b2211c4-390d-4cbe-ab4b-633b2ec7f258 · outbound

This paper cites write newline.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.308197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.308197Z digest=sha256:5a5a1350b8f936dbf2f3368359349b502f2a3a764ae1c12faf6e6f86fb349bc3

Observation c5b4d3c0-1a2c-419d-9779-3934d2920ca3 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.644291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.312144Z digest=sha256:012216b132ff057c025025bd46483aa4e3e90ff6502c3a77069cb5e0b6834376

Observation 3c9df4a3-3254-4c11-981a-b60aa734bb20 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.315514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.315514Z digest=sha256:df92dc18cab5c77234cfd50b64d5b348270d6d9c6474c3a1d566b322468208b6

Observation 1aa4f239-c6eb-4355-ac41-6ca8b2f175a7 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.623753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.318678Z digest=sha256:857446435f7fc0051b53710d68560384eb2a849527b7a8b455a6bfa3fda427a3

Observation 5199c0c7-3b31-4143-83d5-265b25749c74 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.611618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.321753Z digest=sha256:0e05b8a9808cb43eeaf465d99046df6d2c6ada7a02b0e2513f2a474e1e1285f3

Observation d36ac070-4fe3-4450-89a3-a821fb33626f · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.325361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.325361Z digest=sha256:82f8fa6f98206fb9f9e16259386ee3141abeac0c2d4099e93a4295495fff6883

Observation 59166c72-65e1-463d-8dd8-a3576ed1365e · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.328290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.328290Z digest=sha256:fc45f0775f04adbba00d42cfc49521f9313f1f9d5d2aefc64a842b30afea1c62

Observation 65ed488e-0ecb-4c9e-bdf8-521923c9701b · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.588246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.331903Z digest=sha256:77d7ee7d7f8920a506bdc4fff240de39e0b3644e747566bb9d8c4f9bf8404fe1

Observation fb94189d-49ff-47d9-9a4a-c8a1f4284759 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.576461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.336014Z digest=sha256:8207a02ca45d019ec98604f84b7a04c8a5d858ccdf0c5a915eec5504e3697b0d

Observation f248bbce-b326-4b73-bf50-53d02b89b04a · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.566583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.339423Z digest=sha256:cbae72c7dde970334d4826b87d3d217ba6c35a9d096edbb68c4431028fbd3439

Observation 4a014d6d-249b-42d1-809f-50a756d54f61 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.342128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.342128Z digest=sha256:42f7002489eeb5e796d1040ed6679e3b6adb8ced5461180b1d7fc99e8d58a98b

Observation bb6cf1e3-54a8-4f17-9fd0-22aa9bb54271 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.345600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.345600Z digest=sha256:7813731466a7e92615619ecffa067642b2fac46e279f4f67f11682642b94a62d

Observation ba70bf8f-c936-463e-ac3e-6320ef72f1d7 · outbound

This paper cites GPT-4o System Card.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering GPT-4o System Card

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.348309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.348309Z digest=sha256:aa92021730bd01b44960537454835d2307b14d6ff28f1ea081b6a9bb1f093434

Observation eb506337-5b20-4aba-8996-9426aa66869e · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.538651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.351351Z digest=sha256:b76ed53544ea7a016a0b4c17effd2aa6c54db8fbd44f2f522142875a037595dd

Observation 62763ddb-5de0-4798-877e-a8973016e09a · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.527054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.355186Z digest=sha256:8b65e5435f2db61ecbb560d38864c91adc05488206c83be38ba5bd7315b9a1d3

Observation 87b83ab0-41ce-4e1f-bc9f-6cdf1806f9be · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.358224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.358224Z digest=sha256:8dc0e69827b61269028a11b537360be80dab1c793403ff8cf52718bd5e2bd0cf

Observation 38ceacf5-03cf-4ce2-8712-f53b785b74fd · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.508515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.361372Z digest=sha256:63a41eb46673b502e41793b402fc632fafb04b833d4bfd877c70a4c9ab87aeca

Observation 0707af02-b80a-44d0-b2ac-46492a695fe9 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.497499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.364808Z digest=sha256:d170c92136af0a2f117bacc242ba1e40b8c6a7eee5a20c190ce53a879dd3e032

Observation 706d9c05-8131-4f5d-8f69-22fa2424b87b · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.485368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.367374Z digest=sha256:e2f6acebee8fad316b66f331f4e0fbb6a1a71ef34a5ddcf7aef6d846418c608a

Observation 91065e8e-bc80-41f5-afc7-13d83357015c · outbound

This paper cites LXMERT: Learning Cross-Modality Encoder Representations from Transformers.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering LXMERT: Learning Cross-Modality Encoder Representations from Transformers

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.370083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.370083Z digest=sha256:d230524f130066f82ee12708c0e5f02b1e738b7281d581ab0c7c77fe2e7c57aa

Observation 7d6d7e5d-b647-4e1e-a79f-434632a8a03c · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.372850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.372850Z digest=sha256:64cc71190b1efcfbc12121328d901c41c5a39c16b8975e500605bda5b02822ad

Observation 50b925df-0b02-4a8b-a625-8c0a8355c21b · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.473907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.376225Z digest=sha256:e4a871ecc42f28f5ee0865ec9adf0a03508fdb02eb6a9fd324c345f7748af16f

Observation b985f387-fc3d-4c59-aa04-4a9677e8d458 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.463085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.379266Z digest=sha256:47491c88525ff102767722e0ce9cebe720ca295a83cf6f9967983fb15ca5ae51

Observation 5668cf93-4b5c-459e-ad80-fedab7362c86 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T12:58:01.450808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-06T12:58:01.382308Z digest=sha256:8d5b4a219090e33b73ed9b73e4942c861139d67be658b7e647edf3de9c2d2866

Observation cb497a8c-0093-4d53-8c1a-60f62f672650 · outbound

This paper cites an unresolved cited work.

Analyzing the Sensitivity of Vision Language Models in Visual Question Answering Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T12:58:01.385743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:58:01.385743Z digest=sha256:f457cc1591e3bfc4c9463b62a228dcdb8ac589457b1776167f9f721c149649ac

Pith citing papers

No inbound Pith citation observations are available.