Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 41 inbound Pith citation observations for arXiv:2411.04952.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:24:49.927205Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T12:09:48.864251Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d32f2860-0616-4982-b3a8-d69e8cd7924b · inbound
VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5544092-a869-4a6b-9b09-20613d430e27 · inbound
Towards Storage-Efficient Visual Document Retrieval: An Empirical Study on Reducing Patch-Level Embeddings M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f44c4a6-e6e8-44ea-b874-4bc444b6e429 · inbound
The Next Phase of Scientific Fact-Checking: Advanced Evidence Retrieval from Complex Structured Academic Papers M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cac0801-9dc0-44f0-8672-3fb25f9d59a4 · inbound
Structured Attention Matters to Multimodal LLMs in Document Understanding M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3903e34-0903-4e2f-93c1-04ecf26cc0d3 · inbound
Docopilot: Improving Multimodal Models for Document-Level Understanding M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c37fdab9-16b7-4324-ad2a-2c8dea816da8 · inbound
MMGraphRAG: Bridging Vision and Language with Interpretable Multimodal Knowledge Graphs M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87a50d81-abf0-42b3-a340-a0dc3d45a918 · inbound
Finding Needles in Images: Can Multimodal LLMs Locate Fine Details? M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67e871a4-a0ed-4425-a552-41123f71c858 · inbound
SlideAgent: Hierarchical Agentic Framework for Multi-Page Visual Document Understanding M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e9a7a58-70a6-4b00-b6a9-7c67a6af3e3c · inbound
Attention Grounded Enhancement for Visual Document Retrieval M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0dd7a4a8-5a2d-4165-94c7-d3bfe8dd94d2 · inbound
R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7091abbc-640a-41cf-a1f8-43fb994effe0 · inbound
R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ece0a592-947f-407e-aae9-53b47dab02d0 · inbound
MoDora: Tree-Based Semi-Structured Document Analysis System M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5f01a93a-af31-4ff8-b618-a71912722062 · inbound
Towards Real-World Document Parsing via Realistic Scene Synthesis and Document-Aware Training M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb0b10d7-92e9-44ae-9928-9b6a349a1c2a · inbound
Overcoming the "Impracticality" of RAG: Proposing a Real-World Benchmark and Multi-Dimensional Diagnostic Framework M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f84d7f04-5f93-48da-9d31-11c62b0d9b6a · inbound
MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 633e4897-0f0b-48eb-8a3a-729dd8243ac2 · inbound
MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 36094f8a-e55c-40ca-8676-4de8322e971d · inbound
DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8773bdd0-5d75-45d0-bd6a-d352ee7ac2d4 · inbound
DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1bbcf13c-0eae-4fc1-bb61-f7abe1ed8b08 · inbound
PaperMind: Benchmarking Agentic Reasoning and Critique over Scientific Papers in Multimodal LLMs M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b781defd-d22b-4f41-83f0-4d1187ad4eea · inbound
DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 87c71117-c5f2-4adc-bcb2-f8428a09fa6a · inbound
LensVLM: Selective Context Expansion for Compressed Visual Representation of Text M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c0c50346-4c21-4a50-b4b1-da2cec840893 · inbound
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e4023e49-b41e-4fcd-87cc-3d352f47e59c · inbound
MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d8dc2758-9de4-4ab9-8eea-c2ddd260589f · inbound
Unveil: Unified Visual-Textual Integration and Distillation for Multi-modal Document Retrieval M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c1d3445-90a0-424e-9a15-34ecf41e108d · inbound
Overview of the EReL@MIR 2025 Multimodal Document Retrieval Challenge (Track 1) M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2616c404-a4d8-472f-92e1-5c410d18b5ec · inbound
Constrained Dominant Sets for Multimodal Document Question Answering M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b81377d0-739d-4711-9e80-6fcf20b96256 · inbound
EviProp: Seeded Relevance Diffusion on Chunk-Page Graphs for Long Multimodal Document Retrieval M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation dba1760e-c9db-4743-af65-84f72f89ff38 · inbound
The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 247aa86b-3d11-4ee9-ab1f-41c121830ab3 · inbound
Stellar: Scalable Multimodal Document Retrieval for Natural Language Queries M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7fcdb0b4-12f7-4047-b3b4-50dbcb2b6dfd · inbound
Kamera: Unified Position-Invariant Multimodal KV Cache for Training-Free Reuse M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f080281c-aff4-4c07-a7d9-95665d1e0f72 · inbound
DocArena: Turning Raw Documents into Controllable Training Environments for Document Search Agents M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b4918ab3-ca04-4dc0-b40f-06ee414649de · inbound
PIXELRAG: Web Screenshots Beat Text for Retrieval-Augmented Generation M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c66f04c0-5fd5-4ec0-8cb9-0e742f609345 · inbound
LLM-Guided Planning for Multi-hop Reasoning over Multimodal Nuclear Regulatory Documents M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c98a98fd-ce78-4469-9a8b-04c307e045cb · inbound
DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 216
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2961fbf4-4c70-4e04-b49f-5794a096b1ba · inbound
When Do Multimodal and Graph-Augmented RAG Help? A Controlled Evaluation for Document Question Answering M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c84fa5c-719b-4b6e-b92f-93ef095377da · inbound
HVM-GraphRAG: Holistic-View Multimodal Graph Retrieval-Augmented Generation on Complex Document M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cb8f18f-4f33-49d9-a1b6-0ed6457041ef · inbound
Evidence-Grounded Constraint Checking in Construction Documents M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d772b5f9-1224-4359-b01f-e0895233a708 · inbound
HierDoc: Hierarchical Page-to-Region Evidence Routing for Long-Document Visual Question Answering M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a3ed520-b8da-4ed5-b808-563486dbab8f · inbound
XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8d507f5-ccbc-47d5-b24f-74c6a588ed79 · inbound
KoVRE: Training an Efficient Embedding Model for Korean Visual Document Retrieval M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f661bfb-1bc2-4751-84fb-d6d2dcbd2e5c · inbound
DocTrace: Towards Traceable Long Document VQA via Hierarchical Evidence Graph Reasoning M3DocRAG: Multi-modal Retrieval is What You Need for Multi-page Multi-document Understanding
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.