Pith. sign in

Paper Citation Record · LEDGER

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models

As of 8 August 2026, this Paper Citation Record lists 75 of 75 outbound references and 1 inbound Pith citation observation for arXiv:2505.19474.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19474 v1

Coverage vector

measured 75 of 75 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:18:07.338777Z

measured 76 of 76 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T22:38:33.963054Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T22:44:01.550820Z

Reference resolution

75 of 75 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved34
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 014a6b4e-5543-4ca1-b076-21b0c270a20b · outbound

This paper cites Vqa: Visual question answering.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Vqa: Visual question answering

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:57.911238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:57.911238Z digest=sha256:cee57bfa4f4e3f28fbb7c82bdeca8faf786b8c0f4916b7c4617244bf82ebf097

Observation 10842e3c-6d3a-4439-b75b-b3045ac97e33 · outbound

This paper cites Qwen-vl: A versatile vision-language model for understanding, localization, text reading, and beyond.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Qwen-vl: A versatile vision-language model for understanding, localization, text reading, and beyond

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:17.864780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:17:58.010741Z digest=sha256:2b7ccb3de2ec3b4169c08235d8ad12241d86a292dbd288bf867c466bf71c898b

Observation 165fadfc-068b-4381-9f8b-f0418c780c06 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Hallucination of Multimodal Large Language Models: A Survey

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:58.221526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:58.221526Z digest=sha256:d54d55966d22f1a1d48e4091d2a7106fdc7b3be0ef771eff1ebc7b688f691804

Observation f23aa11e-ea2d-4a04-bcb2-280671559ec8 · outbound

This paper cites Causal feature learning: an overview.Behav- iormetrika, 44:137–164, 2017.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Causal feature learning: an overview.Behav- iormetrika, 44:137–164, 2017

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:17.675881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:17:58.385633Z digest=sha256:18e8e88cd7e76ac89304cce5bb4aa64ea1816fb887e38213539d134f9389b04d

Observation 34ac98be-077b-4666-8259-1b97513f6430 · outbound

This paper cites Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:58.499296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:58.499296Z digest=sha256:f994ccbd54687fe5b3fd834007e3ee6629644dac00229af3c173d418fc2e2cf3

Observation b655ec47-1d66-4009-aaed-2c680c5fc13b · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:58.625849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:58.625849Z digest=sha256:44063e28a418325f937f053f5c28b29e3dbb9fb22dccdd3eda77e39f651d96f4

Observation 08463434-275a-44f0-8a15-921f2f369b28 · outbound

This paper cites Neural Modular Control for Embodied Question Answering.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Neural Modular Control for Embodied Question Answering

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:58.749892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:58.749892Z digest=sha256:12e3d445b1214d65f332ab61b1d11cf8a4c513d1d1a788e040f29b800db0e7a1

Observation b2fbccac-c66d-4628-be79-e878396c9193 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:58.895427Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:58.895427Z digest=sha256:a8833f11676260175a8aaddd53872018cccffe487ef7d6b7a5246e1986533723

Observation c9e7988c-42e5-43fa-b732-83a94c0e3604 · outbound

This paper cites Making the v in vqa matter: Elevating the role of image understanding in visual question answering.International Journal of Computer Vision, 127:398 – 414, 2016.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Making the v in vqa matter: Elevating the role of image understanding in visual question answering.International Journal of Computer Vision, 127:398 – 414, 2016

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:17.377920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:17:59.038400Z digest=sha256:4d4ce51cf9a64bde5097b09623bd91ef029df6c41396116b3c220f78c7fad907

Observation 4c0e50e2-16d7-4d1d-abe8-98e1c48e207f · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:17.149581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:17:59.189959Z digest=sha256:fab8e37dc547686e1dadaa8cae36806eb082e219053e0243f6ae2ea5939acc0c

Observation 743e38ca-d838-4589-a5b6-fac6d159fec8 · outbound

This paper cites Unbiased classification through bias-contrastive and bias-balanced learning.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unbiased classification through bias-contrastive and bias-balanced learning

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:16.884156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:17:59.352255Z digest=sha256:c8119d5c42cc4fc257d6f2ca42e0f01104f17241f72a6312cb680a18c1abb9a3

Observation 6f844add-5e84-4dbf-966b-b0ddcb69709e · outbound

This paper cites CIEM: Contrastive Instruction Evaluation Method for Better Instruction Tuning.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models CIEM: Contrastive Instruction Evaluation Method for Better Instruction Tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:59.471288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:59.471288Z digest=sha256:d288c648f64d8c05d7f6dcae406bd1862928348cd78571e1a167930813f9279f

Observation 1b550aec-47b5-4911-bf1f-78e635313428 · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:16.685909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:17:59.606631Z digest=sha256:800d43d1dc9946a9deb59fddcaf658b1431b6eda780f2aee79da3b58e01b8343

Observation 371ec215-5d3f-4af2-9653-411bd6ef6134 · outbound

This paper cites Hudson and Christopher D.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Hudson and Christopher D

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:16.494678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:17:59.786528Z digest=sha256:2da9471a5e3d8c7c2823989d73f5ea71e83a9393fc7ad8c25e1980991183d393

Observation 6b824612-57b4-4b5d-97b2-6a50423c7b47 · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and compositional question answering.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Gqa: A new dataset for real-world visual reasoning and compositional question answering

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:17:59.923714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:17:59.923714Z digest=sha256:96b7793b52029507d10ebf00ab5d0a21fb6a9f62f2ba1fa8f9a96298a0d54478

Observation 23edb86b-9426-4220-910f-9df8a7b8e2f2 · outbound

This paper cites Introducing idefics: An open reproduction of state-of-the-art visual language model.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Introducing idefics: An open reproduction of state-of-the-art visual language model

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:16.267082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:00.095956Z digest=sha256:5255007df21ea38bbf08c533e7c6f8336c6835590b4ceb988745d8cd08acec20

Observation b199e25a-bed3-414e-b672-dba534049efa · outbound

This paper cites Vcoder: Versatile vision encoders for multimodal large language models.2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 27992–28002, 2023.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Vcoder: Versatile vision encoders for multimodal large language models.2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 27992–28002, 2023

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:16.062074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:00.231087Z digest=sha256:88ca9a63d23d97cfe1725895885c69565abc92d6e73e3f3367077dd17af19b18

Observation 08d6b5e4-04de-42e6-b611-c6b08f6e1937 · outbound

This paper cites Causal inference meets deep learning: A comprehensive survey.Research, 7, 2024.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Causal inference meets deep learning: A comprehensive survey.Research, 7, 2024

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:15.820195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:00.334616Z digest=sha256:62b48a9af498cc674589824c2cc58937326ed53b7de9dc33420630cb93c01ea4

Observation 7c7091d6-4d2d-46bd-a51a-954f60465837 · outbound

This paper cites Unbiased learning-to-rank with biased feedback.Proceedings of the Tenth ACM International Conference on Web Search and Data Mining, 2016.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unbiased learning-to-rank with biased feedback.Proceedings of the Tenth ACM International Conference on Web Search and Data Mining, 2016

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:15.592963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:00.483452Z digest=sha256:f9567cf18502c0eb72a2b4b1affa7bc5dee2f7bee6ef704b4f49000de66f942b

Observation 6ec84d65-83d9-479c-a740-01987c29c820 · outbound

This paper cites Deep visual-semantic alignments for generating image descriptions.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Deep visual-semantic alignments for generating image descriptions

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:00.636529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:00.636529Z digest=sha256:70a11e1baf64e68bf8918cd96bf10b0f26d3eb052171182d43b38732e5524a1e

Observation 374a87b5-da22-49ce-9e0c-32b9bc6c8c25 · outbound

This paper cites Learning not to learn: Training deep neural networks with biased data.2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 9004–9012, 2018.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Learning not to learn: Training deep neural networks with biased data.2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 9004–9012, 2018

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:15.414896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:00.758308Z digest=sha256:005196adcfac2ce10f527138a716c06b781e910b1cea7da01725d2910665493c

Observation 15cc1d48-831d-4b2b-815f-8922bf6ca451 · outbound

This paper cites Sophia Koepke, Cordelia Schmid, and Zeynep Akata.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Sophia Koepke, Cordelia Schmid, and Zeynep Akata

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:15.204904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:00.933381Z digest=sha256:7bee649d391d931934fadfe5f38bd922939cfc634ab8c1e420514900afc893ab

Observation 8017d1f0-3ddf-4639-928e-61290230a9f3 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Adam: A Method for Stochastic Optimization

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:01.073175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:01.073175Z digest=sha256:c73a828c59c3737d0a7ff7ca17caac6052e35a4114953d0075b486ce0ac762a9

Observation 057b404f-ab45-4d77-a9ce-4786f8262b62 · outbound

This paper cites Shamma, Michael S.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Shamma, Michael S

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:14.795871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:01.333009Z digest=sha256:5828672d6ab905a855b9edd1d4d3afae229d7839c6f29b4a6a4176daac996e91

Observation 88185505-8ad4-4524-8633-415616b4eaba · outbound

This paper cites V olcano: Mitigating multimodal hal- lucination through self-feedback guided revision.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models V olcano: Mitigating multimodal hal- lucination through self-feedback guided revision

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:14.624966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:01.512276Z digest=sha256:5058bbbce0474ed2559d40ab3989571726cb55cb09a4216c30271e514b0a4f12

Observation 7659c444-2a8f-4bfb-bdff-119f9446e7ba · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:14.451860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:01.654660Z digest=sha256:73ce7fe3b52c231ee7fa3a6879b3adcf3261ae44b74dfbcd3354983e4f1a9dfc

Observation 6ae6f190-5140-4792-902b-f0acf5125dc7 · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:14.227695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:01.793027Z digest=sha256:54555d577e71b602af242a5f814d3c5da06f68b33a88a4db9a5c303597cc004d

Observation bd09b2c2-edf5-4bdd-8762-fa1094b9ac36 · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:14.038951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:01.909532Z digest=sha256:4c1e3dcaa2ff5a63c60d948aec7ef3f1268094291eaba3c8dadc6bff789694e2

Observation fec8f324-f7ed-4747-838f-b80befc809d8 · outbound

This paper cites Towards deconfounded image-text matching with causal inference.Proceedings of the 31st ACM International Conference on Multimedia, 2023.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Towards deconfounded image-text matching with causal inference.Proceedings of the 31st ACM International Conference on Multimedia, 2023

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:13.877823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:02.026813Z digest=sha256:0760ecc7aed7b87c382b77e4ac806cd276762df9cbf935a599442d1cea23e223

Observation 24454eb6-d078-405a-b4d0-26cb107c77ce · outbound

This paper cites Evaluating object hallucination in large vision-language models.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Evaluating object hallucination in large vision-language models

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:13.663560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:02.119604Z digest=sha256:f3b729e7d30bb416a4b5743732e70deab45e8a9c8cb093e110081f8ad8e6fa11

Observation 2e2e3bee-a71e-4ca4-af16-0469f0d9b3ca · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:13.475187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:02.216921Z digest=sha256:151e3ce87f90f93391656cfdf7740e8372bf8b695f4b81deb24914119a6ddca7

Observation 54b96082-8efe-4679-822c-235ba30b0b2b · outbound

This paper cites Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Dollár, and C

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:13.292514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:02.336970Z digest=sha256:f9a1cb0b158d5b20efcdd367540db265b90f6117c95db5ea043254903f69e6c3

Observation eb905292-61e0-4dab-b002-2e2cddee5023 · outbound

This paper cites Show, deconfound and tell: Image captioning with causal inference.2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 18020–18029, 2022.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Show, deconfound and tell: Image captioning with causal inference.2022 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 18020–18029, 2022

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:13.146836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:02.439173Z digest=sha256:0555189dac0be4744803ad1acecc1b462a67bc70c99446e590a6a8363f1dbcf5

Observation a1c74763-284a-4a2e-8092-48a08a69eeca · outbound

This paper cites Mitigating hallucination in large multi-modal models via robust instruction tuning.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Mitigating hallucination in large multi-modal models via robust instruction tuning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:12.977450Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:02.569384Z digest=sha256:32d9b5517ca6a0194b730aead8e6bdfa1578ba80b6ef790057faaff69f27aa36

Observation 8983dbec-8768-4db9-ba5e-b0e487b5f1a6 · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models A Survey on Hallucination in Large Vision-Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:02.648285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:02.648285Z digest=sha256:a8865663126cc90fb137260905aa90cef8c253c938081c9c5ac65b15a984afcf

Observation 941f5dd5-4539-49b9-a280-cf10f8299786 · outbound

This paper cites Improved baselines with visual instruction tuning.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Improved baselines with visual instruction tuning

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:12.676120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:02.899244Z digest=sha256:7e66be73fd1a11e139ef3c29b3cb1c1230513b4f4190e1d81e418f359493a24b

Observation 5e7a841c-fdb8-4c1f-a050-953bc72af9d7 · outbound

This paper cites Visual Instruction Tuning.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Visual Instruction Tuning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:03.192413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:03.192413Z digest=sha256:1dfe0fce2d0c2650bf0cfab69cce533ea3ac3e70a5b240a953667d0c91475594

Observation 0d077064-b255-47b8-ac09-99c7c253a416 · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player? InEuropean Conference on Computer Vision, 2023.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Mmbench: Is your multi-modal model an all-around player? InEuropean Conference on Computer Vision, 2023

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:12.302595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:03.317090Z digest=sha256:95c256f14abf043607bca41850d8f920c327f0233d3db2ed312fd093f25ec0bc

Observation b7588ef9-9cb8-436b-a37c-800b87776e0a · outbound

This paper cites Discovering causal signals in images.2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pages 58–66, 2016.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Discovering causal signals in images.2017 IEEE Conference on Computer Vision and Pattern Recognition (CVPR), pages 58–66, 2016

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:12.118295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:03.453521Z digest=sha256:52295c1bff4d97c2f95b39ebd8f33a486a6af9239fd45de343f73e9c4711d220

Observation 523091e8-a24f-4750-ae8b-f8182bb2be3b · outbound

This paper cites Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Learn to Explain: Multimodal Reasoning via Thought Chains for Science Question Answering

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:03.561844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:03.561844Z digest=sha256:d5aca37232bfc969345978b74c27cf5878d7888b09926cd9693881ecf03434e5

Observation 302123be-ccd8-441d-95c2-aa4eaeb76a51 · outbound

This paper cites The deep regression bayesian network and its applications: Proba- bilistic deep learning for computer vision.IEEE Signal Processing Magazine, 35:101–111, 2018.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models The deep regression bayesian network and its applications: Proba- bilistic deep learning for computer vision.IEEE Signal Processing Magazine, 35:101–111, 2018

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:11.947909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:03.681270Z digest=sha256:2f95bf4558fedb5e6da3daa94f1e5bc30d53254a8c296ce58af794a0b9073e08

Observation 006add3d-a315-4f53-a516-c8a0a11441ad · outbound

This paper cites Counterfactual vqa: A cause-effect look at language bias.2021 IEEE/CVF Conference on Computer Vision and Pattern Recog- nition (CVPR), pages 12695–12705, 2020.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Counterfactual vqa: A cause-effect look at language bias.2021 IEEE/CVF Conference on Computer Vision and Pattern Recog- nition (CVPR), pages 12695–12705, 2020

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:11.777390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:03.808113Z digest=sha256:b1072cc585ca1f52a3dda3e89ada7a012ce60d2fdbc53660399bfde0f274f75c

Observation 5ecf661f-6ea1-435b-926d-75e954dee91a · outbound

This paper cites Basic books, 2018.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Basic books, 2018

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:03.907896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:03.907896Z digest=sha256:be80522c011d4d63577988beda2d39fe554e50f77990a2fb2d708b4544fbf171

Observation 8dfdcc2a-a1fe-4914-a4c7-dc3190657254 · outbound

This paper cites Two causal principles for improving visual dialog.2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 10857–10866, 2019.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Two causal principles for improving visual dialog.2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 10857–10866, 2019

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:11.617036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:04.004376Z digest=sha256:b30acb51222e64189e04752be10db8227be4c8c8abe31139b306c33be202fc29

Observation d55702c7-ff6b-4c49-a756-de8213488948 · outbound

This paper cites Object hallucination in image captioning.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Object hallucination in image captioning

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:11.469998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:04.087000Z digest=sha256:28a1135df4a30a340763f0c3042aadf7cf6a7fd19b6dbee4de00b77932e100e8

Observation 3bbb0d25-af3a-423d-8185-218fc55a05ac · outbound

This paper cites Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encoders.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Eagle: Exploring The Design Space for Multimodal LLMs with Mixture of Encoders

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:04.245081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:04.245081Z digest=sha256:08fdda0dfbfa3d581a600e5a3eaa217a590c557ee067a32f00568071587d31ca

Observation 8bb3660c-308d-4a13-8354-25b943a7e2e1 · outbound

This paper cites Towards vqa models that can read.2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 8309–8318, 2019.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Towards vqa models that can read.2019 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 8309–8318, 2019

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:11.302421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:04.364659Z digest=sha256:1ee6a71d1007e536a19cf566a8b81b67fa67c6624b1876a8eb86733d0a992932

Observation 65608bd0-3c4c-4d26-8627-33705a2fcae3 · outbound

This paper cites Heikkila, and Li Liu.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Heikkila, and Li Liu

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:11.130832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:04.470123Z digest=sha256:150215e68a77bb8823d4b23fce7b619a47dd7d1a1287202e5b9c7a446f634059

Observation edac3404-1493-48f5-80d1-2653097f8606 · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:04.556684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:04.556684Z digest=sha256:171efb396343e7bae71495c92f23c7494249d6c8769c4c5fc469be90d7b4d97d

Observation f1f9e1e8-9e64-434b-a8e2-fa11d12d3a42 · outbound

This paper cites Unbiased scene graph generation from biased training.2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 3713–3722, 2020.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unbiased scene graph generation from biased training.2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 3713–3722, 2020

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:10.936087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:04.696834Z digest=sha256:77ffb742b9365867bff55a849f97692546bed191270707e4d3587f887a61fb19

Observation 919477ab-3f62-44cc-977e-1c958d98f2c6 · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:10.780403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:04.828456Z digest=sha256:13acd88962675298323582c23baad215d4d244ea757dfaa9acd9419149c4e54b

Observation 745b2ab8-e68d-4476-b6fd-ab9083935f9b · outbound

This paper cites Vigc: Visual instruction generation and correction.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Vigc: Visual instruction generation and correction

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:10.604240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:04.944643Z digest=sha256:fed067e2142bcfd2a572a1aa37783e1961313ea64618ec96eb1a27683ee75684

Observation 46414d78-f54d-4b3b-be4d-1f5f1f83e100 · outbound

This paper cites Visual commonsense r-cnn.2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 10757–10767, 2020.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Visual commonsense r-cnn.2020 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 10757–10767, 2020

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:10.410650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:05.074887Z digest=sha256:65e69a2d9c3d31f7c5551883bf7ced78a2e67b32cf1682464fac97876654fe4d

Observation a8223c6f-90d4-4f43-8f1e-5880c854b88c · outbound

This paper cites Advancing Fine-Grained Visual Understanding with Multi-Scale Alignment in Multi-Modal Models.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Advancing Fine-Grained Visual Understanding with Multi-Scale Alignment in Multi-Modal Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:05.180098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:05.180098Z digest=sha256:368fab8b07e45bae9dec3002148a63c35c4cda7c691f7dc8bd5871fedc814b4a

Observation d1959433-08a9-49a0-93d0-1a067cdbb920 · outbound

This paper cites Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Mitigating Hallucinations in Large Vision-Language Models with Instruction Contrastive Decoding

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:05.285816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:05.285816Z digest=sha256:eaf538ed5a40af29dc91a29b0ba2a4136c9e8023a6ba827bc40d9b0bb60df9b0

Observation fa113eba-7ed1-415b-8f0d-49af9320eeb5 · outbound

This paper cites NoiseBoost: Alleviating Hallucination with Noise Perturbation for Multimodal Large Language Models.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models NoiseBoost: Alleviating Hallucination with Noise Perturbation for Multimodal Large Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:05.380471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:05.380471Z digest=sha256:4b7bb48f5cd443f9db0206b5a9d901291884351eabd8d7c45d3eefd8f8d6b101

Observation 253f180e-74c8-4f1e-ad21-5d10357f2318 · outbound

This paper cites Courville, Ruslan Salakhutdinov, Richard S.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Courville, Ruslan Salakhutdinov, Richard S

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:10.277841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:05.525323Z digest=sha256:0e93d782905f7dd031aa5a20774b6f9f7ff4044c7cc27500dd673241689ebbac

Observation 20d51a02-14c0-45ff-8692-a18202018126 · outbound

This paper cites Deconfounded image captioning: A causal retrospect.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Deconfounded image captioning: A causal retrospect

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:10.087615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:05.638942Z digest=sha256:d8053b1c04c81a14d50935c63864865abaf091b77ac442335ed2ac3cccb38a70

Observation 71dbf5a5-7195-4ddb-bfb1-09dabccc5f90 · outbound

This paper cites Causal attention for vision-language tasks.2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 9842–9852, 2021.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Causal attention for vision-language tasks.2021 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pages 9842–9852, 2021

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:09.901506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:05.745609Z digest=sha256:d3f4c237500a3a793ca0730b196e38fc87cafe4e90f2252def4bdc5b72d02b86

Observation 7575dd97-3d66-4c89-bb73-6b0f483b71fe · outbound

This paper cites A survey on causal inference.ACM Transactions on Knowledge Discovery from Data (TKDD), 15:1 – 46, 2020.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models A survey on causal inference.ACM Transactions on Knowledge Discovery from Data (TKDD), 15:1 – 46, 2020

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:09.718877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:05.901162Z digest=sha256:002acaf8d34799ec2e93efe5cd558c9f564be142bcd2d8cfa5136d465bac9835

Observation 41f45e62-096e-4bac-b892-b359a868553e · outbound

This paper cites Woodpecker: Hallucination correction for multimodal large language models.Sci.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Woodpecker: Hallucination correction for multimodal large language models.Sci

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:09.490957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:06.067611Z digest=sha256:2bc919e0568f3226ec494f43ac764bfdf74613618180d2f3d3d6d91ce4579d4f

Observation 224f1f59-cb6d-497f-bdf7-87ef6f198797 · outbound

This paper cites Ferret: Refer and Ground Anything Anywhere at Any Granularity.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Ferret: Refer and Ground Anything Anywhere at Any Granularity

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:06.180366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:06.180366Z digest=sha256:2f76ed83526580b76b01f7d2c418b9f053c06cb85749b8fb7b0b50339b70907f

Observation 80659cfb-199e-4c0f-9523-42f07edff6c6 · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:06.339518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:06.339518Z digest=sha256:cb03d8e07c6abadd0c2a9abebbba6bc5c07b94410afbfa1d045ff0d0ba99ddd3

Observation c82cd4c3-1388-4a71-b80b-5dcc43ae2439 · outbound

This paper cites Interventional Few-Shot Learning.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Interventional Few-Shot Learning

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:18:07.619481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:06.631806Z digest=sha256:736183a9c32565412959b12051e61dcc1a3b2e3a9a0e3990c9ba01aa9f185108

Observation 5e3c6167-066a-43bc-ab62-1856b4a4871d · outbound

This paper cites Halle-control: Controlling object hallucination in large multimodal models.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Halle-control: Controlling object hallucination in large multimodal models

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:08.887598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:06.759675Z digest=sha256:03627b1667d50aa9edfc168ebba97138c47b6b4fe9230eb2e0bcb0a876da849a

Observation f5f9ddb8-4be1-4c37-b9d3-3b73313b8254 · outbound

This paper cites Analyzing and Mitigating Object Hallucination in Large Vision-Language Models.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Analyzing and Mitigating Object Hallucination in Large Vision-Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:06.855970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:06.855970Z digest=sha256:ac1a1149cd30970637675138735eaa6ab56be276be06a7aef3b07c31c95d0abc

Observation 79e8e05a-e626-4d00-9df2-dd4a61db20ba · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:09.114842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:06.505004Z digest=sha256:9ae0e38b28b06230ce491a8ca4d7be41ecb5d2df6772198fbf578112a3f5e14f

Observation f0f35041-5b4b-4063-bf88-62415a9fb8ee · outbound

This paper cites dining table.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models dining table

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:08.731952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:07.028689Z digest=sha256:ad3e2d8cc43a77131f8f51faa5c6216100cc8a8ecc4190c8245db0f1b1ec7405

Observation bde4ca91-50f9-43c9-a11c-74f6db19fcf9 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T14:18:06.946545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:18:06.946545Z digest=sha256:463b913b8eb0e94645724527506064ee7af92b73ac3519916e2cf1661fb523bd

Observation 1ff641f8-205a-4537-af09-3406dcfcdb44 · outbound

This paper cites 7 and 8).

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models 7 and 8)

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:08.514660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:07.102654Z digest=sha256:4cd52bcf8bd8d890605bce3f16a555c5f8e519309b1478db64e7f40ea5e6bd21

Observation 8e766066-1fec-4c4a-80cb-c71c96755af1 · outbound

This paper cites 9 and 10).

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models 9 and 10)

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:08.321787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:07.201470Z digest=sha256:6ec252fb4a94521c3c5603d100c579f3b1e250079088d9e705e90676d83e6cff

Observation 9a6cbe8e-50bc-4de3-aff2-ea4f16e4dafa · outbound

This paper cites dining table.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models dining table

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:18:08.137272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:07.338777Z digest=sha256:edddacd0062a081f7f26cce6b671387d916a7904950de78bf217f28bb6201d35

Observation e1ddd20d-6f33-40ec-93c3-69b60b169d17 · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 2014

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:14.972036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:01.204453Z digest=sha256:08be6e68b63fedf328a66099e611beee89260df5763ce3f18c06a060bc92f8e1

Observation 04c88184-dc78-4217-8d50-78eaaf56251e · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:12.451967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:03.035939Z digest=sha256:c3ff751659c7d0add4a1d5d393f3c52b7cd04b834e8031bceee877f7b526a651

Observation 5b8ed909-166b-43ac-949f-6d80c4f720eb · outbound

This paper cites an unresolved cited work.

Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:18:12.861978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:18:02.773634Z digest=sha256:1c76fb8fc93fd973512c0ce3e9877cf6ae74c7690f3175952fac1ffc414ae615

Pith citing papers

Observation 43347af8-07a4-4acb-884b-5f9e72f63f50 · inbound

Adversarial Orthogonal Disentanglement for LVLM Hallucination Mitigation cites this paper.

Adversarial Orthogonal Disentanglement for LVLM Hallucination Mitigation Causal-LLaVA: Causal Disentanglement for Mitigating Hallucination in Multimodal Large Language Models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:44:01.552618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T22:38:33.963054Z digest=sha256:4af2dddd1b17487722dfff5dcb62bf47115bbac3037f385894828635fe248cb9