Pith. sign in

Paper Citation Record · LEDGER

MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

As of 16 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2503.13964.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.13964 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:27:06.487511Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T03:37:35.692691Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e519a16d-56a8-4f79-9776-7c6a165a5ca3 · inbound

MMedPO: Aligning Medical Vision-Language Models with Clinical-Aware Multimodal Preference Optimization cites this paper.

MMedPO: Aligning Medical Vision-Language Models with Clinical-Aware Multimodal Preference Optimization MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T20:04:44.083780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:04:44.083780Z digest=sha256:1b1d4c5e23b6e6a2ef846998051b615e4eb4fdf22117c762df7be31181eeb709

Observation 14f94aa9-d646-4d93-b0f7-9a4f02751d4a · inbound

A Multimodal Multi-Agent Framework for Radiology Report Generation cites this paper.

A Multimodal Multi-Agent Framework for Radiology Report Generation MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T21:27:06.487511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:27:06.487511Z digest=sha256:0894b93c9b4192e38fbea3302e109703a2c074ad6b49c0120599e1f62ac9b990

Observation b1074efd-badc-440a-a15d-ab1799f567cf · inbound

From EduVisBench to EduVisAgent: A Benchmark and Multi-Agent Framework for Reasoning-Driven Pedagogical Visualization cites this paper.

From EduVisBench to EduVisAgent: A Benchmark and Multi-Agent Framework for Reasoning-Driven Pedagogical Visualization MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:57:37.280223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:57:37.280223Z digest=sha256:6bdadd8d97aa1c5a779480969ef7c3f99e65a21b129c1ecc463f1da1dc84a369

Observation 412043ba-676f-4dd1-b790-7765516916aa · inbound

Structured Attention Matters to Multimodal LLMs in Document Understanding cites this paper.

Structured Attention Matters to Multimodal LLMs in Document Understanding MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:47:37.986765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:47:37.986765Z digest=sha256:179e06e08eff3dd7d6be33d09a8c3d83c06fe593efa0f7c66124cc601d2b6379

Observation 873d134c-a90f-426f-aeeb-ed8110160504 · inbound

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends cites this paper.

A Survey on MLLM-based Visually Rich Document Understanding: Methods, Challenges, and Emerging Trends MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-19T04:42:04.383328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-05-19T04:38:49.512293Z digest=sha256:16e47a1bf6d830197767ef199d1f94c40ff202f301a0dd6b2602dc41b5a0a1c2

Observation d21a0963-2e87-4176-8f8a-e03f2160690b · inbound

MoLoRAG: Bootstrapping Document Understanding via Multi-modal Logic-aware Retrieval cites this paper.

MoLoRAG: Bootstrapping Document Understanding via Multi-modal Logic-aware Retrieval MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T16:26:48.676645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T16:26:48.676645Z digest=sha256:e0a28ca2bc871d93a1f9c6433e35780b8393d7142118899a69d91afbe0df8ba1

Observation aa78b0f3-1d76-4993-8e3f-a5cf439dfde3 · inbound

Dual Latent Memory for Visual Multi-agent System cites this paper.

Dual Latent Memory for Visual Multi-agent System MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T06:04:51.137581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:04:51.137581Z digest=sha256:05a20f55f3deebc461c13dccc06bfa8d57c648f085ee0294b67f14a8d98fc6f0

Observation 7d76c752-4654-40a7-b42e-adb102541d21 · inbound

DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning cites this paper.

DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:09.384776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-08T12:42:39.231468Z digest=sha256:468ca14791f13a405ccd11c7fdce840777e1a9d43e48dfd9712545eaebf0697f

Observation f7309a69-6b88-4852-be51-a920ee220759 · inbound

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning cites this paper.

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:19:27.908626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-14T20:16:30.070819Z digest=sha256:f36bad99a5240f3e6a8d2062e0300dcf1a9a9f316520d8ed368fab9502ba18ac

Observation 6ac94978-921b-4170-9234-04f3dc3e9f38 · inbound

EviProp: Seeded Relevance Diffusion on Chunk-Page Graphs for Long Multimodal Document Retrieval cites this paper.

EviProp: Seeded Relevance Diffusion on Chunk-Page Graphs for Long Multimodal Document Retrieval MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T03:37:35.694129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T15:04:30.757297Z digest=sha256:6477d89fdcdb518922c9550ad8c10b1d5b2869f1cd757f5c0b2451e364eb8b85

Observation 8113f3ec-f0ba-4a70-b410-772ffe9f1138 · inbound

Hybrid Retriever Evolution for Multimodal Document Reasoning Agents cites this paper.

Hybrid Retriever Evolution for Multimodal Document Reasoning Agents MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:04:21.625144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-30T06:58:56.930942Z digest=sha256:e7a345dc7c2539b6dc7334bf3f07e3a723f3cc78ff7f25bd4f987c762e94440a

Observation 8f4d180f-4297-4370-86a0-2ba6247cd0a9 · inbound

Enhancing Large Multimodal Models in Key Information Extraction via Scene-Aware Document Synthesis cites this paper.

Enhancing Large Multimodal Models in Key Information Extraction via Scene-Aware Document Synthesis MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-11T16:02:00.920066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T16:02:00.920066Z digest=sha256:7d8b73297ff247aea25673a93d6cd54f6335fa87710094d7881cc1cc784b3c11

Observation 9278ccfa-f191-4192-a14d-eda8ac272860 · inbound

FinSAgent: Corpus-Aligned Multi-Agent RAG Framework for Evidence-Grounded SEC Filing Question Answering cites this paper.

FinSAgent: Corpus-Aligned Multi-Agent RAG Framework for Evidence-Grounded SEC Filing Question Answering MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T16:10:06.364874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:10:06.364874Z digest=sha256:d0c4e2c6e57410f50078077ecf9fd109e04f452409bfbcbf4a6cde61979c3a34

Observation d2c24b7c-cd98-49a0-b79b-a9c0118b92af · inbound

HierDoc: Hierarchical Page-to-Region Evidence Routing for Long-Document Visual Question Answering cites this paper.

HierDoc: Hierarchical Page-to-Region Evidence Routing for Long-Document Visual Question Answering MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T02:57:48.542979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T02:57:48.542979Z digest=sha256:8bb893a9bc3451ce7f6c8a90388503d2db80e7d299351b8b78f931411ddac2a7

Observation c2f26311-8657-4584-a05d-a8255acd2ffa · inbound

XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding cites this paper.

XL-DocBench: Benchmarking Evidence-Grounded Extra-Long Document Understanding MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T01:39:07.462177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:39:07.462177Z digest=sha256:285b812d627e6f5569a111cec19db1a356956b5d1254cd5e4564ad650707da8e

Observation 9033eae1-b9d2-4fa2-b924-3974b6c847f8 · inbound

Locating Failure in Multi-Page Visually Rich Document Understanding: An Empirical Attribution cites this paper.

Locating Failure in Multi-Page Visually Rich Document Understanding: An Empirical Attribution MDocAgent: A Multi-Modal Multi-Agent Framework for Document Understanding

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T00:43:45.066175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:43:45.066175Z digest=sha256:9c0c18951c0f8497d81e4bc4f6cb3a000e0362a9c44282c783ad2b2a28d0b2fd