Pith. sign in

Paper Citation Record · LEDGER

Quantifying Cross-Modality Memorization in Vision-Language Models

As of 14 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 1 inbound Pith citation observation for arXiv:2506.05198.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05198 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:27:49.753134Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-07T17:58:11.469707Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T23:11:18.189304Z

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f57a6b85-dfd0-4eed-80f8-080d04ad8f25 · outbound

This paper cites Physics of Language Models: Part 3.2, Knowledge Manipulation.

Quantifying Cross-Modality Memorization in Vision-Language Models Physics of Language Models: Part 3.2, Knowledge Manipulation

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:48.299716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:48.299716Z digest=sha256:a0e278dd390676e450392fe48f4fae1832a40b8a6aad268b161b89dc1617de20

Observation aafb8524-eb26-4d8a-a1bc-bc867eb3b7aa · outbound

This paper cites Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?.

Quantifying Cross-Modality Memorization in Vision-Language Models Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:48.584237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:48.584237Z digest=sha256:68bfd3f89268d637a309f983c52a3ab7ca36cb38e5dc6fdaeff3da962782a009

Observation 42a76126-2361-49a8-8ddf-f2b2256c0f6f · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Quantifying Cross-Modality Memorization in Vision-Language Models Measuring Massive Multitask Language Understanding

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:48.692794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:48.692794Z digest=sha256:e2cb775fdf32e929966832973bcfcea2e61382e514afd52ba89a9e3a198c6113

Observation c9dfc881-582b-45ce-a82e-62168f112854 · outbound

This paper cites Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick.

Quantifying Cross-Modality Memorization in Vision-Language Models Tsung-Yi Lin, Michael Maire, Serge Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C Lawrence Zitnick

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:27:50.466596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T10:27:48.918246Z digest=sha256:8790ba2ee39bc11284ee2e839981fbd687f2d4da0d2f76bf97efa1282837a628

Observation b5cd2a40-f0c0-4b9c-bc4a-34f1a8a42fb9 · outbound

This paper cites Scalable Extraction of Training Data from (Production) Language Models.

Quantifying Cross-Modality Memorization in Vision-Language Models Scalable Extraction of Training Data from (Production) Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.155819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.155819Z digest=sha256:412f68956c92a0aae1c3442421d3c3b64a2089eec02f4d7e8f23ffe01c2060da

Observation 51609f8b-3d48-4062-b23c-5cd2169d5732 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Quantifying Cross-Modality Memorization in Vision-Language Models GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.223478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.223478Z digest=sha256:ad52bb2565e30128956c5f2adc44ef58ba2345abc848384c92e9d2dac29e13b0

Observation aad3d495-db8d-4758-bb98-00dd1e0ed42d · outbound

This paper cites Measuring Style Similarity in Diffusion Models.

Quantifying Cross-Modality Memorization in Vision-Language Models Measuring Style Similarity in Diffusion Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.301075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.301075Z digest=sha256:7f3e16c35cb218ed1160317861d63c93286b3db07393b9223731ce9780061a9e

Observation cb20fe76-8122-4f4f-8618-bbf24535a16f · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Quantifying Cross-Modality Memorization in Vision-Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.389145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.389145Z digest=sha256:9c9a5df2185e93622c89369585f29b0f5ae9bb04b7004396fc195ca36a043580

Observation 9cba103c-ab3c-4040-ba27-2fefc0378a5f · outbound

This paper cites Gemma 3 Technical Report.

Quantifying Cross-Modality Memorization in Vision-Language Models Gemma 3 Technical Report

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.493850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.493850Z digest=sha256:282d7646652b2ae496a5e7835f8f1658f24ddeab44b5a53af029de3d1aeca59d

Observation b0e4bb2c-45b7-4e5d-8659-41979ecf42d5 · outbound

This paper cites Measuring short-form factuality in large language models.

Quantifying Cross-Modality Memorization in Vision-Language Models Measuring short-form factuality in large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.586262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.586262Z digest=sha256:fe43aa32d32a05f98b13be1ae261891c32ec4955111e83bbed5c58bd25f57a86

Observation bc4530e6-94f6-4ced-878f-07503ae1f9d4 · outbound

This paper cites Privacy Backdoors: Enhancing Membership Inference through Poisoning Pre-trained Models.

Quantifying Cross-Modality Memorization in Vision-Language Models Privacy Backdoors: Enhancing Membership Inference through Poisoning Pre-trained Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.683237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.683237Z digest=sha256:bbd86460bc47d6c02bbb864b800a9ede8b2853ec7e15fd1d978f248a07bd3810

Observation 0f4f4cc9-899a-41e1-8089-17635ccbe180 · outbound

This paper cites Large Language Models Are Not Robust Multiple Choice Selectors.

Quantifying Cross-Modality Memorization in Vision-Language Models Large Language Models Are Not Robust Multiple Choice Selectors

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.753134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.753134Z digest=sha256:e4603efd2820b9e8dc6975702f9d616291374850d251d29bc16e5892e3ccbfaa

Observation 9919505c-1bbb-47e4-b8db-1f85e8d4cefa · outbound

This paper cites Decoupled Weight Decay Regularization.

Quantifying Cross-Modality Memorization in Vision-Language Models Decoupled Weight Decay Regularization

Reference 2014

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.013233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.013233Z digest=sha256:d66c9f077bb42e4e66b90365a0633bb580089d9ebc51e574b1628bb8a0f1d784

Observation 44fcc08c-757c-4119-8e82-0c1ba43040ec · outbound

This paper cites TOFU: A Task of Fictitious Unlearning for LLMs.

Quantifying Cross-Modality Memorization in Vision-Language Models TOFU: A Task of Fictitious Unlearning for LLMs

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:49.098394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:49.098394Z digest=sha256:73294b136da0f1180f3513691707c14bdcbd4455895feddd0e9dfd4f833cbb80

Observation 2824d38a-4f12-4b92-a6ec-e990b3ef09a6 · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

Quantifying Cross-Modality Memorization in Vision-Language Models LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:48.851135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:48.851135Z digest=sha256:1e2464744e0ff1f4a53682aebf2264a9fa84c41dcff14aa45448043bcc759e93

Observation ff8c600c-0673-41e4-918d-fdc6bb831460 · outbound

This paper cites Training Data Leakage Analysis in Language Models.

Quantifying Cross-Modality Memorization in Vision-Language Models Training Data Leakage Analysis in Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:48.763185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:48.763185Z digest=sha256:0da8070635c467d7f8019ea0f31c6055b5beb8b4fce3f0a013bd6cbc829edf16

Observation 447ab3df-50fd-4c0f-9dba-199a1ba88968 · outbound

This paper cites Imagen 3.arXiv preprint arXiv:2408.07009,.

Quantifying Cross-Modality Memorization in Vision-Language Models Imagen 3.arXiv preprint arXiv:2408.07009,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:48.402343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:48.402343Z digest=sha256:c1ff8287b8e777eccba9376daf9b7de6428f4706af98b8aec38525949eb83ff1

Observation 90cedb31-bfa6-44d4-9360-8724dd32b562 · outbound

This paper cites Fantastic Copyrighted Beasts and How (Not) to Generate Them.

Quantifying Cross-Modality Memorization in Vision-Language Models Fantastic Copyrighted Beasts and How (Not) to Generate Them

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:48.630517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:48.630517Z digest=sha256:579f9b6e15e461606fd173f7e50e940e743e7959b7684eef251cbb8e85fb7b82

Observation 83fed1bc-4803-45cd-ae98-ff2c68d4f35e · outbound

This paper cites The Reversal Curse: LLMs trained on "A is B" fail to learn "B is A".

Quantifying Cross-Modality Memorization in Vision-Language Models The Reversal Curse: LLMs trained on "A is B" fail to learn "B is A"

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T10:27:48.477387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:27:48.477387Z digest=sha256:cb439772996a45a9f5b5922f8a6a108013fbb2a5a32fb5d0cd2f065f030a473b

Pith citing papers

Observation 67bde97d-8577-4dcc-ae3d-5d98f920875f · inbound

Before Forgetting, Learn to Remember: Revisiting Foundational Learning Failures in LVLM Unlearning Benchmarks cites this paper.

Before Forgetting, Learn to Remember: Revisiting Foundational Learning Failures in LVLM Unlearning Benchmarks Quantifying Cross-Modality Memorization in Vision-Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:11:18.193547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-07T17:58:11.469707Z digest=sha256:4493aa8b8c38c47c6b276b36f417edaab6c9c108bf593a1fdabe78d03638de6a