Pith. sign in

Paper Citation Record · LEDGER

MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2407.01523.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.01523 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:24:52.343132Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T16:17:09.161062Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a4ec3907-b82d-449f-a91d-604544b7d78a · inbound

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output cites this paper.

InternLM-XComposer-2.5: A Versatile Large Vision Language Model Supporting Long-Contextual Input and Output MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 104

Resolution
verified exact
arxiv_id, observed 2026-05-17T10:46:28.895841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T10:46:28.447347Z digest=sha256:2d5b94cebe5db6942a7691561c49594522c1673e183da4f673a3a8906a4b7ebd

Observation f474b427-eae1-430a-ba6e-6291ebc0328f · inbound

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning cites this paper.

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:33:26.769132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:33:26.613927Z digest=sha256:94e12f3d5a93c7508b7c0cd2c60e65d57ba88e718ca8727f2da9166d6f1d1e71

Observation c0ed3f2c-aa1f-435a-90bb-bee0c89944d7 · inbound

VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning cites this paper.

VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:24:52.343132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:24:52.343132Z digest=sha256:f9231d124cb5d2763ec709b865da54d67e0a5b55e1a42a7d8483515b36f6ca95

Observation aee3b0cc-1206-4230-a44e-afc8d3ed86ee · inbound

WikiMixQA: A Multimodal Benchmark for Question Answering over Tables and Charts cites this paper.

WikiMixQA: A Multimodal Benchmark for Question Answering over Tables and Charts MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T23:58:26.366936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:58:26.366936Z digest=sha256:62c0d7b7b294a6959eed76550177a3caa971fa8a1941f7209e6c52a60f46540a

Observation 3b933033-63d9-44ac-aff3-402f563405c4 · inbound

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU cites this paper.

Breaking the Boundaries of Long-Context LLM Inference: Adaptive KV Management on a Single Commodity GPU MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T23:00:11.143406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:00:11.143406Z digest=sha256:80c7ad1aeeb3b780f31f96f34cfeb7298eb9559f9f22a21c17bb47018e2b9116

Observation 1260f822-738f-40b9-996b-4ff5567b5c5f · inbound

The Next Phase of Scientific Fact-Checking: Advanced Evidence Retrieval from Complex Structured Academic Papers cites this paper.

The Next Phase of Scientific Fact-Checking: Advanced Evidence Retrieval from Complex Structured Academic Papers MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T22:45:54.555828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:45:54.555828Z digest=sha256:223822b8883ceded3a998aedc276215181af31f6c7349c77781e767fc76d4d9d

Observation 79f003c5-613d-490b-919f-c1a4b29a166a · inbound

Structured Attention Matters to Multimodal LLMs in Document Understanding cites this paper.

Structured Attention Matters to Multimodal LLMs in Document Understanding MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T23:47:38.698127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:47:38.698127Z digest=sha256:83ba0ae07ecdfabcdf01113fdd3d9843938cc27c2b9a31d1404bb4a10d1c40dc

Observation ec58836f-28a4-4deb-82ba-8570b98fcec4 · inbound

VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents cites this paper.

VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:10:15.201235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T14:10:14.929207Z digest=sha256:3b39f7630e23ae5cf3966cf42f97998ceb618848a1266395006cdb06d251a8af

Observation d4acb8a5-5d1b-43b1-94a9-f204c153c49a · inbound

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark cites this paper.

Document Haystack: A Long Context Multimodal Image/Document Understanding Vision LLM Benchmark MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T16:12:02.996244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:12:02.996244Z digest=sha256:cf58b145c4425a856eab6b4e74ee7cd17287d99db21eb5fbef4f24e9cb119841

Observation b6602d08-cec9-4ae5-8d63-dd8c648ee984 · inbound

Multi-Agent Interactive Question Generation Framework for Long Document Understanding cites this paper.

Multi-Agent Interactive Question Generation Framework for Long Document Understanding MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T13:45:18.451030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:45:18.451030Z digest=sha256:665579f1bc3e463b1472f50fdee161783c30da6e07742234a49bb6dda3e99cd6

Observation 539bd716-b703-4f35-bd20-4a0174b3bf81 · inbound

VisR-Bench: An Empirical Study on Visual Retrieval-Augmented Generation for Multilingual Long Document Understanding cites this paper.

VisR-Bench: An Empirical Study on Visual Retrieval-Augmented Generation for Multilingual Long Document Understanding MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T22:10:08.166980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:10:08.166980Z digest=sha256:d8c0f9557eedffbad9502a6f07ca6ca8a3b02cfeeed50f005d99ef1b9d3dfb7c

Observation c97f5a42-b1d1-455e-a20e-a55081a197d1 · inbound

A-SEA3L-QA: A Fully Automated Self-Evolving, Adversarial Workflow for Arabic Long-Context Question-Answer Generation cites this paper.

A-SEA3L-QA: A Fully Automated Self-Evolving, Adversarial Workflow for Arabic Long-Context Question-Answer Generation MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T11:23:43.927169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:23:43.927169Z digest=sha256:f5144229d7803a8af32410a3fdaafd584304b5271d0f90fb1b928237e66e99e1

Observation 8b59d568-3960-45b1-a816-be2561950209 · inbound

MTraining: Distributed Dynamic Sparse Attention for Efficient Ultra-Long Context Training cites this paper.

MTraining: Distributed Dynamic Sparse Attention for Efficient Ultra-Long Context Training MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T19:44:19.671447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T19:44:04.833504Z digest=sha256:d1e30b0a4d33c1faf0278f67e692b0eec3a5b7822ee2b9571510e66734defc1b

Observation 46761a21-751d-45be-8914-1f098d401b18 · inbound

Internalized Reasoning for Long-Context Visual Document Understanding cites this paper.

Internalized Reasoning for Long-Context Visual Document Understanding MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:53:28.106451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T23:53:19.148407Z digest=sha256:1ba08457f0d9c1081f31a6e77892a4e9cbf49d927312f7bc75696da78b4e10a3

Observation 31ed8674-e872-47a7-bcc2-0d1907688e01 · inbound

Internalized Reasoning for Long-Context Visual Document Understanding cites this paper.

Internalized Reasoning for Long-Context Visual Document Understanding MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-13T15:50:49.083652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T15:50:49.083652Z digest=sha256:8b2ad61fe26995ee22745f59a37ca31f01106fb3e564202645459e9e34dd6f6f

Observation d454812c-67b1-4dda-a968-0520fdc06ff7 · inbound

FileGram: Grounding Agent Personalization in File-System Behavioral Traces cites this paper.

FileGram: Grounding Agent Personalization in File-System Behavioral Traces MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:45:51.251477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T19:33:57.643549Z digest=sha256:72881b32f5f05ebe3f0f246bd438f4e627d9ecc02438af14aff94ff50b59f604

Observation 76cc8f0e-d4c9-4a96-b6e1-ab35d84a2c13 · inbound

Unifying Sparse Attention with Hierarchical Memory for Scalable Long-Context LLM Serving cites this paper.

Unifying Sparse Attention with Hierarchical Memory for Scalable Long-Context LLM Serving MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:01:25.920832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T13:15:21.201950Z digest=sha256:41ca11cab7c9966122543da1ee20505116c764458638408dc38e0f747e40538c

Observation 164a2a23-3a53-4da0-9a95-cc889f31716a · inbound

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models cites this paper.

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-30T21:05:04.225422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T21:00:25.664841Z digest=sha256:197e5d4246acef839a2ca623ecb1c470464b7e129a7a88d70339a302ef84a59c

Observation 4d4608b8-b9b3-4689-bd9c-51a078a89607 · inbound

DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark cites this paper.

DocRetriever: A Plug-and-Play Framework for Multimodal Document Retrieval with Comprehensive Benchmark MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:15.101581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T08:09:41.068000Z digest=sha256:7786e322f6a4da1d724398790ab9a703f21cceb497b31fd02a42c291f0be2240

Observation 00d1f3ac-5416-44a1-a76b-60d1f6ccd64d · inbound

How Much Dense Attention is Necessary? Oracle-Guided Sparse Prefill for Full/GQA Layers in Hybrid Long-Context Models cites this paper.

How Much Dense Attention is Necessary? Oracle-Guided Sparse Prefill for Full/GQA Layers in Hybrid Long-Context Models MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:17:09.162985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T22:52:13.337419Z digest=sha256:3f8ef1718aaa60deb3e6641435ab570d4887b848d1f7a6435e7bedf85eecb15f

Observation fe90420f-cffe-47ca-87c1-41f41a1f507f · inbound

LLM-Guided Planning for Multi-hop Reasoning over Multimodal Nuclear Regulatory Documents cites this paper.

LLM-Guided Planning for Multi-hop Reasoning over Multimodal Nuclear Regulatory Documents MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T07:24:22.472263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T07:15:01.669896Z digest=sha256:d79e3c75ef3600a8eb18492c611319cc1a71ccde3896590d60f1777b574c2145

Observation 9f59fb12-2cf2-4986-8f1f-b974419dd301 · inbound

Hybrid Retriever Evolution for Multimodal Document Reasoning Agents cites this paper.

Hybrid Retriever Evolution for Multimodal Document Reasoning Agents MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:04:21.631252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-30T06:58:56.930942Z digest=sha256:d9bc9d58291b8fe1e2849d759b88e40c55ec13b9cac6dd1325ec026c9e7cb843

Observation be5c54ce-6652-4743-a806-d350eec7c2b1 · inbound

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition cites this paper.

CLBench-V: Evaluating Multimodal Context Learning from Grounding to Knowledge Acquisition MMLongBench-Doc: Benchmarking Long-context Document Understanding with Visualizations

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T02:55:23.398613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:55:23.398613Z digest=sha256:93b86d80ba0fff55faeef402d886d2c88f75802d893b8d03ccbc846b8f584041