Pith. sign in

Paper Citation Record · LEDGER

Revisiting the Role of Language Priors in Vision-Language Models

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2306.01879.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.01879 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:05:37.515954Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T07:31:13.950276Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04b7d7ec-d4f9-402a-ae47-7cdbe860d0e6 · inbound

Decomposing Complex Visual Comprehension into Atomic Visual Skills for Vision Language Models cites this paper.

Decomposing Complex Visual Comprehension into Atomic Visual Skills for Vision Language Models Revisiting the Role of Language Priors in Vision-Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:37.515954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:37.515954Z digest=sha256:f97aeff63ae7c920f5b3f7ed9876112fa6a22d2eaaacdaf0cec947e8578ec140

Observation ecbdd126-2c17-4950-8c3c-f65d8f2cea17 · inbound

ARGUS: Hallucination and Omission Evaluation in Video-LLMs cites this paper.

ARGUS: Hallucination and Omission Evaluation in Video-LLMs Revisiting the Role of Language Priors in Vision-Language Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:41:39.695176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:41:39.695176Z digest=sha256:826cc11f335edbcc8a08425495d43e6f892c00496a153bed2a3d8b51646e52a0

Observation 3cea00c3-9b4f-4d96-9dd4-cab80c24a988 · inbound

Trade-offs in Image Generation: How Do Different Dimensions Interact? cites this paper.

Trade-offs in Image Generation: How Do Different Dimensions Interact? Revisiting the Role of Language Priors in Vision-Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T12:08:01.756106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:08:01.756106Z digest=sha256:3e53a33e5320a8d0bd671d789ca6590773a979f22bf31da369a4f58e7a75a84d

Observation da03bf5f-cf85-4386-a535-6bc73e606300 · inbound

TokenSwap: Backdoor Attack on the Compositional Understanding of Large Vision-Language Models cites this paper.

TokenSwap: Backdoor Attack on the Compositional Understanding of Large Vision-Language Models Revisiting the Role of Language Priors in Vision-Language Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-04T13:52:15.074817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:52:15.074817Z digest=sha256:8f40a39330be62f617430cad51d738f03dd543b263b623df05e1eaa2fa9978fc

Observation ed0e9454-7441-4148-89dd-c43c9464f95d · inbound

Building a Precise Video Language with Human-AI Oversight cites this paper.

Building a Precise Video Language with Human-AI Oversight Revisiting the Role of Language Priors in Vision-Language Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:46:04.507451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T00:37:31.858728Z digest=sha256:982ab2aec766a114ff529d9eecf1552218c2db0abdf33a0bd67a185c34c069e1

Observation 27270a0a-9306-4d62-a338-67bc4726689e · inbound

SDGBiasBench: Benchmarking and Mitigating Vision--Language Models' Biases in Sustainable Development Goals cites this paper.

SDGBiasBench: Benchmarking and Mitigating Vision--Language Models' Biases in Sustainable Development Goals Revisiting the Role of Language Priors in Vision-Language Models

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-22T07:31:13.953141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-22T07:29:49.950084Z digest=sha256:c16b2baab4dfb78a059d4f047dc5bc5598ddb17ac88a36d2f4b663f2aebd7f5e

Observation db7c25f4-b9aa-461a-8b8f-5d7f5e3f644f · inbound

Prior Bias in Vision Language Models on UML Diagram Interpretation cites this paper.

Prior Bias in Vision Language Models on UML Diagram Interpretation Revisiting the Role of Language Priors in Vision-Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-12T06:36:12.761601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T06:36:12.761601Z digest=sha256:aa87fc971623319122a59936993e8ad5e39509c16cdcd3db6c76cb368d986e87

Observation d20ccc5c-43f6-487e-aaf1-5b1d8fa8e22a · inbound

Scalable Visual Pretraining for Language Intelligence cites this paper.

Scalable Visual Pretraining for Language Intelligence Revisiting the Role of Language Priors in Vision-Language Models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-02T07:33:08.314581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:33:08.314581Z digest=sha256:5482997c803e7f44d994220ff6ecb4c654f0b5ed86b43f7c6b464aebbc725c39

Observation 02261088-172b-430b-bb3d-1085ad446991 · inbound

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO cites this paper.

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO Revisiting the Role of Language Priors in Vision-Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T18:44:46.262146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:44:46.262146Z digest=sha256:15262ea6bae9204019aeba26f052a0f8e31cfeeb72e7367578397272bafa9120