Pith. sign in

Paper Citation Record · LEDGER

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation

As of 9 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 2 inbound Pith citation observations for arXiv:2506.17664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.17664 v1

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:34:56.649817Z

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-11T01:49:15.136031Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T16:01:22.762740Z

Reference resolution

18 of 18 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a01c9e40-12ba-4f38-aed7-c5675210c69f · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.081734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.081734Z digest=sha256:2201e2c82e9c0934ae519020e978bbe59dd9159e4916b27d38f12f08edc3d6e4

Observation 5381a6ff-955e-46b2-a247-d81895bd4159 · outbound

This paper cites Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Seeing is Believing: Mitigating Hallucination in Large Vision-Language Models via CLIP-Guided Decoding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.164743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.164743Z digest=sha256:264e78d5877ce472eff9069b90453e15f95128416eba10b645cf082521a2497b

Observation 3acb9a3f-0ac5-4fc1-9016-dfa90bb4677d · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.278265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.278265Z digest=sha256:89e6c73139098dae320a3d2ebf96dc2d2e5f547c409d9fff522ed8b919287030

Observation 8b2b0a8f-af0a-4306-ae4a-dd4706c9bad8 · outbound

This paper cites GPT-4o System Card.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation GPT-4o System Card

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.465653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.465653Z digest=sha256:f2c2afb45996f2c465cf5b162858ac526d18eb4e9ed452904420f37f6d3a5851

Observation 58ef4792-b290-4307-9c93-f8f8618fade0 · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.674752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.674752Z digest=sha256:d72bb0e7a2f4b599b5e96ae7e572f344cbd4aae46201cb5721d919b1b0607df9

Observation f5974394-d72d-41d7-8774-134f78e52d20 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.735935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.735935Z digest=sha256:6139303f5c9b9068df667dc558145739ff77cad27d3ae910886688aa3c8d5c20

Observation 6876b6ee-b62b-4f54-bc7e-060ab3694c6a · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.885669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.885669Z digest=sha256:440f14ef8744df69946d42451ae42de286cea31c14100572dd3a654f1257170e

Observation b08c0987-0469-4868-ad41-5cf26c981d82 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Gemini: A Family of Highly Capable Multimodal Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.934743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.934743Z digest=sha256:62e3ebf2d4c94f36f89c0592ace56160abe399e6b994a7754d15cfecc6853c91

Observation b6bd49ea-46fc-4c65-836e-cf82d369b442 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.992900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.992900Z digest=sha256:3ad50d699519de508f867199d59abcae61fa81fbc4bb41aaf42729bb2edff4be

Observation ded3c166-299f-491a-8d8a-cf7f8be290b8 · outbound

This paper cites Caption Anything: Interactive Image Description with Diverse Multimodal Controls.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Caption Anything: Interactive Image Description with Diverse Multimodal Controls

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.110735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.110735Z digest=sha256:a23a56d95d2c10dddd8db37f8b81f38505deb4313791d647cf2e670228770589

Observation 0bdd5568-3306-4e1c-9e9b-31e075c5403e · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.275773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.275773Z digest=sha256:386f7fcc9ca5c77fc5366188d5b0a582ecf5c5db9608137e6bdd62fdc4a63e78

Observation abe0cc92-da89-4367-a161-33631ac7882e · outbound

This paper cites mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.417716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.417716Z digest=sha256:933a4d99db20232d88a9e6bd4be3559ffa546e957e6094bf77db85b20aaf297a

Observation a95efcf4-84f7-4b69-8af0-693b4a982f1e · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.541827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.541827Z digest=sha256:f7604b88e81f1f5d0133e8e0527547588ab656c02ec5eb93b719185f037d5ad1

Observation 921ef3b2-bdf5-4337-85df-c2d8e04ff6f1 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:56.649817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:56.649817Z digest=sha256:2a536d5a86b0ca97c299811596508fd0996919df11a7c76d00020c66f506c3ca

Observation 39332d3d-a380-49ed-9974-bd42dc128f82 · outbound

This paper cites Object Hallucination in Image Captioning.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Object Hallucination in Image Captioning

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.813461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.813461Z digest=sha256:35aab0a7ab54ee09d562ae6eb485ab07771b99752f3b4021b714f36b5f43789d

Observation 751df6f5-9b7c-486d-b97a-77f3358bb160 · outbound

This paper cites VisualBERT: A Simple and Performant Baseline for Vision and Language.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation VisualBERT: A Simple and Performant Baseline for Vision and Language

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.544749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.544749Z digest=sha256:591cc753f3d745ad40b69de3bed15c5c87d1016b6a1ee9e7dbb0ae6d6063c2c5

Observation 400765dc-0cac-40bc-a2aa-1021cce2aab4 · outbound

This paper cites MultiModal-GPT: A Vision and Language Model for Dialogue with Humans.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation MultiModal-GPT: A Vision and Language Model for Dialogue with Humans

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:55.367724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:55.367724Z digest=sha256:01d8e6dfaae76497a07c746c0a2aa6c79e2871b98bef5d83104ad52df8c09895

Observation 629c7d3e-52de-4b5f-9070-f98089e618c2 · outbound

This paper cites Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention.

MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation Mitigating Object Hallucinations in Large Vision-Language Models with Assembly of Global and Local Attention

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T23:34:54.937267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:34:54.937267Z digest=sha256:01e0854ca3e5f8d7ba3380f1ef2ba22b34e25eec4fb0e187ffcddcd322ff56e3

Pith citing papers

Observation a6efc5ab-82e8-46db-9ee7-616d25da5476 · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:01:22.777348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T18:53:06.494640Z digest=sha256:93cef37243cb01a96304c5a0a69b492d82ace5c6fc4b0de79b0b39660b447887

Observation c1436332-d86a-4610-9ff0-277a199108ba · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs MDSAM:Memory-Driven Sparse Attention Matrix for LVLMs Hallucination Mitigation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:50:51.532867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:49:15.136031Z digest=sha256:df7a977c92ba7a7bc2f230048c977367c1dab68772f539849b4818100550393f