Pith. sign in

Paper Citation Record · LEDGER

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG

As of 21 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 6 inbound Pith citation observations for arXiv:2506.02544.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02544 v2

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:27:42.487192Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T08:13:52.083852Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:59:46.895392Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7f73c72b-652c-42c6-a672-7a0a324433d9 · outbound

This paper cites Hallucination of Multimodal Large Language Models: A Survey.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG Hallucination of Multimodal Large Language Models: A Survey

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:41.685365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:41.685365Z digest=sha256:efb9563de9faf4e13685ef700d04324942c24a887cbaa50635fa6bedf8f6366a

Observation 3beadc01-1b76-4f3d-bbad-477785cd7ce6 · outbound

This paper cites Zhe Chen, Jiannan Wu, Wenhai Wang, Weijie Su, Guo Chen, Sen Xing, Muyan Zhong, Qinglong Zhang, Xizhou Zhu, Lewei Lu, et al.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG Zhe Chen, Jiannan Wu, Wenhai Wang, Weijie Su, Guo Chen, Sen Xing, Muyan Zhong, Qinglong Zhang, Xizhou Zhu, Lewei Lu, et al

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:27:43.089302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:27:41.799921Z digest=sha256:9daa560766e689b3f4e97345c49c00dec1783c748cb8fa4bcb12f7b31e155796

Observation d2f654ba-9ed5-4898-be63-a5944f2fc62d · outbound

This paper cites Retrieval-Augmented Generation for Large Language Models: A Survey.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG Retrieval-Augmented Generation for Large Language Models: A Survey

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:41.852787Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:41.852787Z digest=sha256:b3bba9b88c050383b08e1fc145d5d848208f2f57c516f1dc128f1c2e0464f33e

Observation ca497e44-ca88-41ad-a5bc-c696b0e9fb87 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG LLaVA-OneVision: Easy Visual Task Transfer

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:41.900497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:41.900497Z digest=sha256:5a79311416f8b413eafafb39177c7678395d669e593a61715da441451ae98c87

Observation af0be477-6e68-402b-8965-f41bcf419d97 · outbound

This paper cites RoRA-VLM: Robust Retrieval-Augmented Vision Language Models.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG RoRA-VLM: Robust Retrieval-Augmented Vision Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:42.060896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:42.060896Z digest=sha256:61a003cf4e4d07b8c8f5344a847cad84af2a692e88621a5dcc31a5b647bc4cec

Observation 46205d72-2ffe-4cef-83b9-c37970ecc8d8 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:42.125290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:42.125290Z digest=sha256:723f7c10e311faf008abe085168c25242ed4d1a56bf53174d671c6138e0aaf5c

Observation b61f0927-d9e2-4c71-894d-05bbedf80005 · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:42.212327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:42.212327Z digest=sha256:c45f75eb746c089c2edbc94f58b13aada0dc21f294f06b0fa6a478cc2e9562bf

Observation 16e1be0e-6fb7-4639-abb9-7ed1acc566bc · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:42.252278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:42.252278Z digest=sha256:2bc9c2c84df278ae3a08236dadd3a8e6bbe4d21b149f882164467ebacaf373cd

Observation 810c2194-4f26-46ea-bfed-33bb50d756c1 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:42.299624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:42.299624Z digest=sha256:87c8d74998290e815d2833ff0944d1f90b89f4d97e32f2a363053ee57b9996f5

Observation 01216d86-24ca-43d0-9595-09af80a5572e · outbound

This paper cites spring of the kid.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG spring of the kid

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:27:42.893032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:27:42.487192Z digest=sha256:97aeae1c05dce3243c1d05f8de4dc758347a9a4bdf0da4b8e40dd8500e650eaf

Observation 8dfd9098-ea8c-4823-9bbb-e6f90747cbee · outbound

This paper cites Scotland.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG Scotland

Reference 2019

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:27:42.993214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:27:42.422193Z digest=sha256:8812df13af88ba75dcc02fb37be54ea2c71cc1ec5ab7873ff237a711a74be2bd

Observation 535a8136-400b-4dc6-809c-92d4c64cfe0b · outbound

This paper cites A Survey on Hallucination in Large Vision-Language Models.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG A Survey on Hallucination in Large Vision-Language Models

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:41.996134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:41.996134Z digest=sha256:a49888e27007668b38121648360e0a4d89e56130d981c8564e263a9d797863e0

Observation 875aafb3-c194-4219-98e1-50d24b64f030 · outbound

This paper cites mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:42.371621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:42.371621Z digest=sha256:ecf6bbb7ee26196d0ad5bccb7c8e4156a94c9b8c0e79cc2bbff5bab8f4685b32

Observation 8ab4768c-4cb0-4da1-a5fe-509f928df0e3 · outbound

This paper cites GPT-4 Technical Report.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG GPT-4 Technical Report

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:41.579116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:41.579116Z digest=sha256:88d3c4aa874d999c8b39b28550fcd8110aa3a097c754d508ed39cb21f6786d8e

Observation d0a4cde3-1e51-4733-847e-47cf845e2e51 · outbound

This paper cites InProceedings of the 2024 Con- ference on Empirical Methods in Natural Language Processing, pages 16499–16513.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG InProceedings of the 2024 Con- ference on Empirical Methods in Natural Language Processing, pages 16499–16513

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:27:43.218499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-07T11:27:41.629502Z digest=sha256:77314546806facb77b036c08bc080ed7de43018e1ad216b97b6792068441f30b

Pith citing papers

Observation d1480f20-8b74-4610-b137-904eab329c4c · inbound

QKVQA: Question-Focused Filtering for Knowledge-based VQA cites this paper.

QKVQA: Question-Focused Filtering for Knowledge-based VQA CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:57:54.018414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T12:53:59.668663Z digest=sha256:097675dc91eb2d7e2d01e657f0226827221e8c2cbea17f7d1df0513c0db90f79

Observation 164cbf70-785c-49c6-b63c-109cd5b42946 · inbound

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation cites this paper.

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:47:45.523818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-16T10:46:17.411843Z digest=sha256:25113c66211f31b441bec41ecd09652401c40d56a94ed625a21853569de052a8

Observation 216cee87-7a15-40eb-bf92-d585ea4654ea · inbound

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation cites this paper.

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T08:13:52.083852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:13:52.083852Z digest=sha256:4ae9da3fbacb4efa2759d777cf78638200a20054e99dc118dd673264a884608e

Observation 3712d792-7b11-493f-af27-a0c1801b0f11 · inbound

Ground Then Rank: Revisiting Knowledge-Based VQA with Training-Free Entity Identification cites this paper.

Ground Then Rank: Revisiting Knowledge-Based VQA with Training-Free Entity Identification CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:59:46.896966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-26T08:12:14.829556Z digest=sha256:772ac2b827cbac74380bb5949668a09c30a2d65a304cfc0145fcca5e018f2229

Observation aa184d3b-8526-42fb-9614-167065862127 · inbound

Identifying and Resolving Pitfalls of Knowledge-Based VQA Benchmarks: Auditing, Repairing, and Augmenting cites this paper.

Identifying and Resolving Pitfalls of Knowledge-Based VQA Benchmarks: Auditing, Repairing, and Augmenting CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:17:17.749904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-02T19:13:44.015782Z digest=sha256:5a9cec5437b8edf691f607adce834be7fff10c6417009421692c518c5c4390c2

Observation 88f469f5-e898-4ebe-a509-99c2b00d22dc · inbound

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG cites this paper.

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-02T10:20:50.924849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:20:50.924849Z digest=sha256:1ccc3600dcee1df389b38ecc986e17f5807eca70be7ed1fa1f6cb166d7583853