Pith. sign in

Paper Citation Record · LEDGER

mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 17 inbound Pith citation observations for arXiv:2411.15041.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15041 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:41:10.213488Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T12:46:14.749364Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 11dfac29-2663-49b2-9a6a-8cff6c747a78 · inbound

OMGM: Orchestrate Multiple Granularities and Modalities for Efficient Multimodal Retrieval cites this paper.

OMGM: Orchestrate Multiple Granularities and Modalities for Efficient Multimodal Retrieval mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T22:41:10.213488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T22:41:10.213488Z digest=sha256:48d31a57e14bd990cc5c5f891cfcf0260c62c28e4ecbfe801ee6554ee2c52c44

Observation 571d8895-cb81-4a6f-9f88-fd0577c18585 · inbound

Mixture-of-Retrieval Experts for Reasoning-Guided Multimodal Knowledge Exploitation cites this paper.

Mixture-of-Retrieval Experts for Reasoning-Guided Multimodal Knowledge Exploitation mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:52:19.993606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T13:50:30.090068Z digest=sha256:16521665b5ee7f890d70d8bf206e59e51c681cfb512839825d366f290b4f1e23

Observation 875aafb3-c194-4219-98e1-50d24b64f030 · inbound

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG cites this paper.

CoRe-MMRAG: Cross-Source Knowledge Reconciliation for Multimodal RAG mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T11:27:42.371621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:27:42.371621Z digest=sha256:ecf6bbb7ee26196d0ad5bccb7c8e4156a94c9b8c0e79cc2bbff5bab8f4685b32

Observation 0028bbce-5de7-4a43-97ad-17a3629fb18e · inbound

mKG-RAG: Leveraging Multimodal Knowledge Graphs in Retrieval-Augmented Generation for Knowledge-intensive VQA cites this paper.

mKG-RAG: Leveraging Multimodal Knowledge Graphs in Retrieval-Augmented Generation for Knowledge-intensive VQA mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-19T00:06:55.031490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T00:05:08.866244Z digest=sha256:52794c31a412a9894455389e8ebbc59a2df42e74563d742a220ea13683a017b4

Observation ba192c7a-0448-4bcb-b339-ddc3dd8396d8 · inbound

Towards integrated sensors for optimized OCT with undetected photons cites this paper.

Towards integrated sensors for optimized OCT with undetected photons mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-05T23:26:28.640658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:26:28.640658Z digest=sha256:2da536fc1a3c28d5131e41d0e458b3ea616f14ef18c236fd34b8c1bc2a253e4a

Observation b2bdfe73-d43b-41af-85ab-ae2ccc4bbc60 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 111

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:53.107445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:53.107445Z digest=sha256:f9650bbf2552fef5bfba5bee41da4dca28c6ed00f3240ea368d7665e3bfa0286

Observation 1cef374f-1c2e-4445-9c40-1aae5369ac16 · inbound

Progressive Multimodal Search and Reasoning for Knowledge-Intensive Visual Question Answering cites this paper.

Progressive Multimodal Search and Reasoning for Knowledge-Intensive Visual Question Answering mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-18T20:06:49.834243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T20:04:52.852253Z digest=sha256:867f534024305c4b7b88c7b7a6798769417c997f716e6ef9d1f96a01008445a4

Observation c3f50af2-666a-4c96-9b4f-03dc62a0dc58 · inbound

Recurrence Meets Transformers for Universal Multimodal Retrieval cites this paper.

Recurrence Meets Transformers for Universal Multimodal Retrieval mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-04T20:06:09.182830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:06:09.182830Z digest=sha256:4f2c333a0b9a74a96c3664b5a7fba05de0e5ae00e044f140d5260c29142008dc

Observation 75576e0d-ca0b-4215-afda-05b752adcfb2 · inbound

QKVQA: Question-Focused Filtering for Knowledge-based VQA cites this paper.

QKVQA: Question-Focused Filtering for Knowledge-based VQA mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:57:54.028617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T12:53:59.668663Z digest=sha256:48465f8cc434a8c8581ebac256d94752ce649d9a0a77dd4d69f55974aeff66f7

Observation cb79a4e0-3510-432c-8642-b6680e020dfe · inbound

WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering cites this paper.

WikiSeeker: Rethinking the Role of Vision-Language Models in Knowledge-Based Visual Question Answering mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:20:48.065154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-10T19:59:10.657346Z digest=sha256:3c791b22b91813c581df66acaa0b15c3037bf506bd21c5cd58a7f46d7d7db351

Observation 8cb9be3d-a963-4268-b521-4e1250a7b597 · inbound

CogniVerse: Revolutionizing Multi-Modal Retrieval-Augmented Generation with Cognitive Reflection and Geometric Reasoning cites this paper.

CogniVerse: Revolutionizing Multi-Modal Retrieval-Augmented Generation with Cognitive Reflection and Geometric Reasoning mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 121

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:23:15.730102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-29T08:14:20.558527Z digest=sha256:5eb48ebd96c654850671ba23209c92223ead0286f9b0968f956af3a0d8691759

Observation c2c2eb77-cb22-404b-8831-8d3126503c4c · inbound

Ground Then Rank: Revisiting Knowledge-Based VQA with Training-Free Entity Identification cites this paper.

Ground Then Rank: Revisiting Knowledge-Based VQA with Training-Free Entity Identification mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:59:46.839046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-26T08:12:14.829556Z digest=sha256:04e5ca7b8106b493f9fea2c50db6326c557e7494a901c23792b1973e8fa4623c

Observation aa6dbd07-ebad-41b7-b8e4-c04f6228adb9 · inbound

MMAgent-R$^2$: Learning to Rerank and Reject for Agentic mRAG cites this paper.

MMAgent-R$^2$: Learning to Rerank and Reject for Agentic mRAG mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 36

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T12:46:14.750625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-09T12:37:07.906342Z digest=sha256:6b4264d416b92044536394597ea0cabecf3ea61e30f40d48e1f6032b3ef73607

Observation 11576698-4461-4e44-88f8-c8736543b8bb · inbound

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG cites this paper.

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T10:20:50.701345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:20:50.701345Z digest=sha256:efa521f0b81e8e83463f5d4085c9a364e0e6fb13257476b1c89ec07ab85ca766

Observation b8c9e70d-aa80-4e22-a27a-d2696ec8f8d1 · inbound

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering cites this paper.

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:17.751661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:17.751661Z digest=sha256:e43fd51ce4e94628512270b964e1e0e5f882f556d8bc188cbebfc315384ea740

Observation 90a75562-a2e8-494d-8a15-9392fa7bb123 · inbound

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering cites this paper.

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T04:19:53.771485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:19:53.771485Z digest=sha256:4922fe9dfa2458f2852bec6e9867db00e8b3c5ae6e633df896650ba690f7471b

Observation 5069e662-c10a-44cf-aef2-4d67d23e3bcf · inbound

M$^3$Prune: Hierarchical Collaborative Pruning for Efficient Multi-Modal Multi-Agent Retrieval-Augmented Generation cites this paper.

M$^3$Prune: Hierarchical Collaborative Pruning for Efficient Multi-Modal Multi-Agent Retrieval-Augmented Generation mR$^2$AG: Multimodal Retrieval-Reflection-Augmented Generation for Knowledge-Based VQA

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T20:15:14.692955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:15:14.692955Z digest=sha256:ba22d26638737142346da4afe0e7c02c9e34a9303f729db8641c7a76d4e7d2af