Pith. sign in

Paper Citation Record · LEDGER

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization

As of 15 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 1 inbound Pith citation observation for arXiv:2508.20181.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.20181 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:18:38.304903Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T13:29:33.614965Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T13:33:28.064115Z

Reference resolution

63 of 63 outbound references displayed

  • verified exact0
  • verified fuzzy42
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 86fa58ad-9fef-468b-8454-9981aceb40a5 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:33.237590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:33.237590Z digest=sha256:09715845ee4c89a47660e131bbaae5da8bd59c54ebeaf92b68f637bf209bf1ae

Observation 30c65b2a-8c2a-4420-9b05-6500010b9090 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Hallucination of Multimodal Large Language Models: A Survey

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:33.324613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:33.324613Z digest=sha256:e45041f52ca8437832ee5d9713ebd29b86f26839acc94df7b83686f9386dd084

Observation 4ae0cf45-190a-474b-87e3-2f62f86ec140 · outbound

This paper cites With a Little Help from your own Past: Prototypical Memory Networks for Image Captioning.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization With a Little Help from your own Past: Prototypical Memory Networks for Image Captioning

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:46.651646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:33.368982Z digest=sha256:6c2a52c9cf137a9cef4a763fe631d0b9997e8bbfa2b9fe0e27ffe0def908b305

Observation 1659d00d-a5ac-46c0-9207-c1b612038898 · outbound

This paper cites The Revolution of Multimodal Large Language Models: A Survey.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization The Revolution of Multimodal Large Language Models: A Survey

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:46.443655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:33.438837Z digest=sha256:99805357a195af689c7bbdb712f7fafb99ec5d0a69825dbaa8e3c46349ea93f1

Observation 9bfb76dc-749f-4a03-abd1-e54c319ce65d · outbound

This paper cites End-to-End Object Detection with Transformers.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization End-to-End Object Detection with Transformers

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:46.285354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:33.537393Z digest=sha256:cd6834b29d0f95b09dd71292cac8116c7289dee1750bce8f947b269b67bbb9c0

Observation b9ab9a6f-c19c-4e72-832d-1e18dfcb6fda · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:33.646628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:33.646628Z digest=sha256:29ab0f4d03a86aae055061495facf2defb53641a6f814713a21eb7cb3c71b26d

Observation de52eb4e-5928-4687-a62e-82d6be2a150c · outbound

This paper cites Gonzalez, Ion Stoica, and Eric P.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Gonzalez, Ion Stoica, and Eric P

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:46.139959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:33.725432Z digest=sha256:95d374b4478300f254d1c74511a2fd2859eaca2ad1ff9d07125389a76858f825

Observation a864a7fd-65c0-4e08-a6ee-f8308d2cb009 · outbound

This paper cites DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:45.818209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:33.892444Z digest=sha256:c6c47c18779e0b10f18389efdf6d81773fd462071e5e451b867c9a332a2564f2

Observation 409edc50-62c0-4faf-b294-d9d658422371 · outbound

This paper cites LLaV A-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization LLaV A-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:45.699335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:33.975071Z digest=sha256:c54273faad63aebe7cb3f5a18b2c6ecbed87d20307d10f95684b55227ac72e4b

Observation 85874f6f-1ea4-42e8-988f-1fd23bfac7d8 · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:34.027368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:34.027368Z digest=sha256:50e21db3f380d439588480ea0a0c7430a6c5d66d7a5425af8610e042f16a585e

Observation bbfcb825-8c4e-4243-9022-cddcb6cf1bfb · outbound

This paper cites The Llama 3 Herd of Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:34.097678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:34.097678Z digest=sha256:16b81e8c4d559242e83befb4becf8931cf427649e0e4223993078440529d6fd5

Observation 26692736-0744-4ec6-ab86-8691a51495b5 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:34.147013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:34.147013Z digest=sha256:d84cc275d11ef60f2fdcbbd34d5dc751c749af4d83a6ecd04a8ec8752cec6f32

Observation 975ac317-3c2a-4179-bac7-21f7b5ddf0c3 · outbound

This paper cites OneLLM: One Framework to Align All Modalities with Language.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization OneLLM: One Framework to Align All Modalities with Language

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:45.558840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.201914Z digest=sha256:92709c426d9a956c3ca6414c0d69ac035a7338794e6c4c4c40599a2b40eec3d9

Observation 48bc36a0-c4c1-4021-a769-7251e75154b0 · outbound

This paper cites ORPO: Monolithic Preference Optimization without Reference Model.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization ORPO: Monolithic Preference Optimization without Reference Model

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:45.415876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.264082Z digest=sha256:4f317cafbd963ce87446791ca2701eb35e42a0dbbaad39c744cde070fd485c82

Observation 71a0e014-6121-40c8-bc06-53f4ad4b5d69 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization LoRA: Low-Rank Adaptation of Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:34.315829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:34.315829Z digest=sha256:36ab25faeae60b88161dd049f985b2fb6aec58643070b752ea98fda621f38bf3

Observation 7a7bfd4f-cb9f-4325-bd30-aaf09ad67403 · outbound

This paper cites A Survey on Hal- lucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization A Survey on Hal- lucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:45.197572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.371606Z digest=sha256:94c2ec7313ea1baa93020b017e96bb20d3192089c2aeb60739a01ddddb2c493c

Observation c14090ae-40f0-470b-ab88-20f9de405246 · outbound

This paper cites OPERA: Alleviating Hallucination in Multi- Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization OPERA: Alleviating Hallucination in Multi- Modal Large Language Models via Over-Trust Penalty and Retrospection-Allocation

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:45.013860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.415801Z digest=sha256:b3a09c5b0ef86eb67bccecac528d591bbfc8e1c3fe4c45e04d89b362a9f3c19b

Observation 5e7976ae-1078-4768-951f-4cc87b0400a0 · outbound

This paper cites SymDPO: Boosting In-Context Learning of Large Multimodal Models with Symbol Demonstration Direct Preference Optimization.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization SymDPO: Boosting In-Context Learning of Large Multimodal Models with Symbol Demonstration Direct Preference Optimization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:34.487538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:34.487538Z digest=sha256:cf58e5543258734a4354f8ff5fcc0206d5d88f0e4b7bdbb04116c267631ec039

Observation e846edc9-7c40-4f77-b1ce-f3caa3706ea2 · outbound

This paper cites Modality-Fair Preference Optimization for Trustworthy MLLM Alignment.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Modality-Fair Preference Optimization for Trustworthy MLLM Alignment

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:34.573801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:34.573801Z digest=sha256:61190484d5b84d32af9c42946d2d8a86bc89ea3c23b7df089a5f637cdc831827

Observation a42f421e-1b5d-415d-9ed2-ff3e785bc210 · outbound

This paper cites Deep Visual-Semantic Alignments for Generating Image Descriptions.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Deep Visual-Semantic Alignments for Generating Image Descriptions

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:44.831257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.627243Z digest=sha256:2b630fdda4f38ea14505be41e9f8f4fe99a29fcc138bbc9f31c41dd4131c9486

Observation 59680675-abe2-464c-94f4-a16dfb6e9bb0 · outbound

This paper cites A Diagram is Worth a Dozen Images.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization A Diagram is Worth a Dozen Images

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:44.684805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.679487Z digest=sha256:212a9e94c4398958d7f588c37fd10ba443c36b675582061cdd5b2c87a9b4bc63

Observation 5e912978-ee41-49a9-9984-22d5bd633da6 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Adam: A Method for Stochastic Optimization

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:44.544165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.760053Z digest=sha256:0f69b253591c601d87323417f60f95de471bbf2697e236dfe3e4743b95f893d3

Observation 1052199a-a17c-4420-88d2-7892d7a4e89e · outbound

This paper cites Building and Better Understanding Vision-Language Models: Insights and Future Directions.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Building and Better Understanding Vision-Language Models: Insights and Future Directions

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:44.355978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.809548Z digest=sha256:7587159a893aa99361adc870507fc59c2c94643e018e4f912fd103d6968daf57

Observation 782f36dd-a6c6-4514-9b0e-130490a86fe2 · outbound

This paper cites Mitigating Object Hallucinations in Large Vision-Language Models through Visual Contrastive Decoding.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Mitigating Object Hallucinations in Large Vision-Language Models through Visual Contrastive Decoding

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:44.204406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.877521Z digest=sha256:34d27fec81c10bc29a80c33415e66ef4cb6079d9c406cc908299771c914c32ad

Observation 40f8d7e9-5703-4601-bb7d-a57e3398ae68 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:34.966231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:34.966231Z digest=sha256:023203d9dae744d69dea544cc1385102125c09b267f32ef8ad84155a180d1d3a

Observation 58dee9f8-daed-449c-9788-e2ce6244e4a2 · outbound

This paper cites Microsoft COCO: Common Objects in Context.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Microsoft COCO: Common Objects in Context

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:44.061394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:34.986672Z digest=sha256:772fd779eed4f319917c35e6fef9df43440618015ae0e7ada6c66600243d444a

Observation 8a29cf57-fccb-4d96-804a-05a41d34fda5 · outbound

This paper cites Visual Instruction Tuning.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Visual Instruction Tuning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:43.906662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:35.104916Z digest=sha256:154de9e36ac5ab4c3f9718878991dece6a7561bc6c0323805eb173e65ac81938

Observation c781c242-1977-4ea8-9c01-e67a1ebf7b27 · outbound

This paper cites Improved Baselines with Visual Instruction Tuning.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Improved Baselines with Visual Instruction Tuning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:43.781369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:35.168310Z digest=sha256:29dbbd15bd37cc96ab8e50ac13fd928ad6115d402c0f6754d830175c98e01862

Observation 8f50a13a-764c-48c3-80ec-3fa938bc28b8 · outbound

This paper cites Neural Baby Talk.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Neural Baby Talk

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:43.575627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:35.235925Z digest=sha256:dd6130cbd14033e3a7b84cd170b2e37d90406499317392ddec97a06dc981d2f2

Observation 77d1b174-dcb7-4de3-9546-555040288061 · outbound

This paper cites Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:43.406314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:35.338300Z digest=sha256:e36fe46b2237b5c4e0a6f51fc1fce5e77c746ec64cecdc9ebbeafb0b0acdfe40

Observation 8dc5e733-1208-4222-ab15-4a45185e77ea · outbound

This paper cites Revisiting Image Captioning Training Paradigm via Direct CLIP-based Optimization.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Revisiting Image Captioning Training Paradigm via Direct CLIP-based Optimization

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:43.299250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:35.418133Z digest=sha256:611d1d574dc90739a6783665cb6750d52162e57ed0ffc98d64848776ad74f4a2

Observation 832d9554-25a1-458b-aafb-d7903e8b18df · outbound

This paper cites Training Language Models to Follow Instructions with Human Feedback.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Training Language Models to Follow Instructions with Human Feedback

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:43.120956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:35.522414Z digest=sha256:babedb9a43acd9f7188ac137362dd4b48be3fb1c620e090165ae96e1f33c49a5

Observation bee4125a-b824-4a19-a701-e96c733a7aa2 · outbound

This paper cites X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization X-InstructBLIP: A Framework for aligning X-Modal instruction-aware representations to LLMs and Emergent Cross-modal Reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:35.610779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:35.610779Z digest=sha256:84aa60640bdc23e447578cae743cb6b073f9d1f44e35cefd45ffd34a74c033a7

Observation 9544c7ab-c9dc-49a6-a6ee-5bbee819651c · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:35.706620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:35.706620Z digest=sha256:6a8ea9833868b7440fbfd568aa6c3fbc18b631094464472595c2564db29f1912

Observation 20eb7156-1f02-471e-9caf-07b94a86a2ec · outbound

This paper cites ALOHa: A New Measure for Hallucination in Captioning Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization ALOHa: A New Measure for Hallucination in Captioning Models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:42.905433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:35.805621Z digest=sha256:c1ebb3174fc03ed36f969a89b2bb66375388348c8617e5137cbd3a8e63af5c5d

Observation 07fdf4d2-ab81-4fa0-b523-3151a51980cf · outbound

This paper cites Learning Transferable Visual Models from Natural Language Supervision.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Learning Transferable Visual Models from Natural Language Supervision

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:42.755521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:35.930386Z digest=sha256:5f803596206cce1fc749a70bb5a6bc9f9c7d8013fa2956a9c090d73741257ddf

Observation 843fec70-3497-42ef-b395-0745e3657dc9 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:42.550010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.047228Z digest=sha256:b77632a42c9a9f2f7685c16f884b8529fb297108c615f1044f5ce530faca3389

Observation 2ab81ef8-0b8d-4628-a278-a6ef25256dbf · outbound

This paper cites Zero: Memory optimizations toward training trillion parameter models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Zero: Memory optimizations toward training trillion parameter models

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:42.384739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.126068Z digest=sha256:4a9c0b40307da2e1db1fffad34930df286a77047a23d58c7dd592028c25e2688

Observation 95965ad1-59e2-46e3-8621-f1a702492dc4 · outbound

This paper cites an unresolved cited work.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:18:42.192143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.231373Z digest=sha256:7a9e930887fe1cd7a4f719cc062eb2a3945723e260252cbef7ffec4de94ad069

Observation aeffee8b-6a26-47bc-a105-d830f14030d7 · outbound

This paper cites GLaMM: Pixel Grounding Large Multimodal Model.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization GLaMM: Pixel Grounding Large Multimodal Model

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:41.986421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.346112Z digest=sha256:49679981f2337e6ef7de88baf1215ed0d1a02f3076a782cc132f4a4c7168d5cf

Observation f21c23fc-3b63-43d6-b3f3-1ab37cb59282 · outbound

This paper cites Object Hallucination in Image Captioning.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Object Hallucination in Image Captioning

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:41.766647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.415263Z digest=sha256:ec2d4b92aa976d1f0d9c1fe3c077a7e9f6ee921aa705ee444e39bda09a1c7239

Observation e2148b3b-839c-4bc9-b88e-1885fa84595a · outbound

This paper cites A Comprehensive Survey of Hallucination in Large Language, Image, Video and Audio Foundation Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization A Comprehensive Survey of Hallucination in Large Language, Image, Video and Audio Foundation Models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:41.518595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.509687Z digest=sha256:0c7bf9db5bab4794b3fc50e70fce88048786b2bfcd3a6ff6184a38722467b9bd

Observation 18cd79d6-89c2-4614-9fd1-0edbf04ae8c5 · outbound

This paper cites Mitigating Object Hallucination in MLLMs via Data-augmented Phrase-level Alignment.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Mitigating Object Hallucination in MLLMs via Data-augmented Phrase-level Alignment

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:41.320547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.579679Z digest=sha256:1c6666c4f679459f6c2c304c3d032bd2ed3e1cd44285d4a1295b1a9ff3f79431

Observation 2039ac83-8e54-4728-b8a7-577a9767572f · outbound

This paper cites Image Captioning Evaluation in the Age of Multimodal LLMs: Challenges and Future Perspectives.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Image Captioning Evaluation in the Age of Multimodal LLMs: Challenges and Future Perspectives

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:41.081923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.705610Z digest=sha256:d09945960be6e575fd1a5fc142a0de6d97c722290d9213d49d42a47da43b24e0

Observation cbd03fb7-f0da-4845-9c97-0303d3c985e1 · outbound

This paper cites Preference Ranking Optimization for Human Alignment.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Preference Ranking Optimization for Human Alignment

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:40.899088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.810488Z digest=sha256:5bac6cbb4c6329bf8eb59eaec7a3a658ea0d8c81c568bc7d810ee9d7b84e9765

Observation 1421c234-b11e-4b39-b548-1af87161d091 · outbound

This paper cites Generative Multimodal Models Are In-Context Learners.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Generative Multimodal Models Are In-Context Learners

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:40.659037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.909666Z digest=sha256:a61f3c4539b72c0a449a764eda657aace277beddbdfede785ce3a7b3e1963886

Observation 8bfa37b4-121b-4936-9a27-13d83b38e546 · outbound

This paper cites Emu: Generative Pretraining in Multimodality.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Emu: Generative Pretraining in Multimodality

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:40.408757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:36.985220Z digest=sha256:5cfc1302cdf9eb23683df403dab50475dc1520a2eddcd5cfa216258fadc17ac4

Observation a9ea855e-2585-4e7c-b925-5d19ae5679d1 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:37.090202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:37.090202Z digest=sha256:77fb3a02f10b3c4de5b4197b2589d80eb21715cf1ebeb547068fd33a2d47bb32

Observation b496aa58-812d-4983-9b89-7980a5cfb16d · outbound

This paper cites Diffusion Model Alignment Using Direct Preference Optimization.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Diffusion Model Alignment Using Direct Preference Optimization

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:40.143255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:37.175875Z digest=sha256:c133ea27ddfc8f36a47ebafb6e40a7e7041ada1d57ce6f2aba61a52e6c1996b0

Observation 542df41f-67b2-4837-b2cd-d153c98b5043 · outbound

This paper cites mDPO: Conditional Preference Optimization for Multimodal Large Language Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization mDPO: Conditional Preference Optimization for Multimodal Large Language Models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:39.954334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:37.237229Z digest=sha256:6443dceede86989ebc30c846c87692d0e8e72d7c620eb5514d15654d992d875a

Observation e47e4548-3d66-4c7c-812c-0ee946ed0445 · outbound

This paper cites AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization AMBER: An LLM-free Multi-dimensional Benchmark for MLLMs Hallucination Evaluation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:37.343758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:37.343758Z digest=sha256:cabcdb95e318951df96076358cea308662d680bce91d357fa2cd0e33ee357c51

Observation 763ebf0d-5362-40ed-bb2b-a54dc70e86a7 · outbound

This paper cites β-DPO: Direct Preference Optimization with Dynamic β.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization β-DPO: Direct Preference Optimization with Dynamic β

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:39.747243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:37.450314Z digest=sha256:4a5a75116d9b351dcd84656428c23ed6c2832d01a5e0bcb0b032bd2b4ded893e

Observation 71201094-c262-4c15-a072-894163de4499 · outbound

This paper cites Generate, but Verify: Reducing Hallucination in Vision-Language Models with Retrospective Resampling.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Generate, but Verify: Reducing Hallucination in Vision-Language Models with Retrospective Resampling

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:37.545387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:37.545387Z digest=sha256:3db576994e68f7f8ac1219d5b65534a3b7beaf50edde4b09d1a19264924cb4a4

Observation 68748eee-39d8-4884-970b-7d04b4f7f28e · outbound

This paper cites Hallucination is Inevitable: An Innate Limitation of Large Language Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Hallucination is Inevitable: An Innate Limitation of Large Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:37.645387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:37.645387Z digest=sha256:b46f64a1d11f00957b7e4d3948feb9d7ac5f7f4095c74636b8d06231a6d8d739

Observation da24e78b-5ea9-48d6-b792-84db0f1e00c9 · outbound

This paper cites mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:39.562614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:37.730768Z digest=sha256:4ab9a0c0c8dbdcabccce624a09cd672fe44c1c7eb7eec34cbbbf4ecc0e978c28

Observation 54aecb68-9571-4f58-aad0-03e267d4e7cb · outbound

This paper cites Woodpecker: Hallucination Correction for Multimodal Large Language Models.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Woodpecker: Hallucination Correction for Multimodal Large Language Models

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:39.284926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:37.868533Z digest=sha256:161b7572610215e0a8769988550c5d5e5bbb9e4da37fd89099bd0690e15f398e

Observation 59fcc504-01a8-47f3-ac35-cb77c76f39af · outbound

This paper cites RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization RLHF-V: Towards Trustworthy MLLMs via Behavior Alignment from Fine-grained Correctional Human Feedback

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:39.157806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:37.961289Z digest=sha256:d3d97789c4458625ce4d6c6d098679c5d10a7c708cf3d51236fe17de5e4f78c4

Observation 313a7679-acee-4046-a3cd-d430e7fe5070 · outbound

This paper cites MMMU: A Massive Multi- discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization MMMU: A Massive Multi- discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:38.972271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:38.030782Z digest=sha256:b1b99ebd198d52f87e77b43e85af8a951649a66ab97819cf6bbc9dc4e5afaab3

Observation 546fe4ea-ce4f-4fe7-b9f4-1bd76adecadc · outbound

This paper cites Less is More: Mitigating Multimodal Hallucina- tion from an EOS Decision Perspective.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Less is More: Mitigating Multimodal Hallucina- tion from an EOS Decision Perspective

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T15:18:38.798240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:38.089116Z digest=sha256:cd107a985805d7e4de606c4934242b06e86661e5289b2b25113bfaa5750c8f12

Observation 1661428f-ebf7-4c65-8919-adb87ceb1eb5 · outbound

This paper cites Direct Preference Optimization of Video Large Multimodal Models from Language Model Reward.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Direct Preference Optimization of Video Large Multimodal Models from Language Model Reward

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:38.154826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:38.154826Z digest=sha256:2460d6700e6050da41ad5c27fe51e9c54421817ea947bf9c20e65a18e4c46eec

Observation cd5e3884-6b94-458c-8444-5aad718bfc37 · outbound

This paper cites Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:38.238296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:38.238296Z digest=sha256:9bcbcc492ee077353f6097b97fa489ac6d96626a291dbed7355d3db6649ded18

Observation 434d2507-e21f-4685-9e3a-59f1de608687 · outbound

This paper cites Aligning Modalities in Vision Large Language Models via Preference Fine-tuning.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Aligning Modalities in Vision Large Language Models via Preference Fine-tuning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-05T15:18:38.304903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:18:38.304903Z digest=sha256:f514c0e76ede1b6416f1bb8a37fc3dbca6d31597a39d7201ceb9b0fbc681207f

Observation 58115e0f-2772-4f35-9bbf-efa8d4f24aa3 · outbound

This paper cites an unresolved cited work.

Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-05T15:18:45.992589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-05T15:18:33.768882Z digest=sha256:7da60664adc74c6d035f9a078df2df0ca2269984446992ef616e5dcb50c2309a

Pith citing papers

Observation 52162736-4518-4fe8-a41c-cfd720d0e08f · inbound

Rethinking Visual Neglect: Steering via Context-Preference for MLLM Hallucination Mitigation cites this paper.

Rethinking Visual Neglect: Steering via Context-Preference for MLLM Hallucination Mitigation Mitigating Hallucinations in Multimodal LLMs via Object-aware Preference Optimization

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:33:28.065432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-29T13:29:33.614965Z digest=sha256:67b39d501fc25dad43ee6d3e33286f1b3cd73eeb0a9ff843bfd5dd9680b34915