Pith. sign in

Paper Citation Record · LEDGER

CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2401.11944.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.11944 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:19:05.167967Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T04:32:32.672779Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 623d5a10-821a-485a-a79f-434c82c4544c · inbound

DeepSeek-VL: Towards Real-World Vision-Language Understanding cites this paper.

DeepSeek-VL: Towards Real-World Vision-Language Understanding CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:58:54.760355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T17:58:54.177359Z digest=sha256:def3622b40a942e2a3a4476ed7bbf460ed789449d604015f86b1e2b52e40ccf5

Observation e18477da-9ef8-4c40-889f-ce3c5099ed29 · inbound

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning cites this paper.

OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 131

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:33:26.885629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T20:33:26.613927Z digest=sha256:68e1b0c44c5587d49e28a8e88a5906f1e827aaba44235a31c326320ba786fb7e

Observation 68fc532c-d9aa-4e30-9e14-86a8f87eb304 · inbound

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning cites this paper.

Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 247

Resolution
verified exact
arxiv_id, observed 2026-05-23T04:32:32.676271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-23T04:30:38.804702Z digest=sha256:0cc25c86c4087ba7b6d1cf85c54dc4015b9b8011b5570dd438ed49a7450101f9

Observation 8dd42996-6022-4616-9e13-33f11b1ae93e · inbound

FlagEvalMM: A Flexible Framework for Comprehensive Multimodal Model Evaluation cites this paper.

FlagEvalMM: A Flexible Framework for Comprehensive Multimodal Model Evaluation CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T05:14:44.959144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:14:44.959144Z digest=sha256:d6edd5f3f4b1de2359fd6d6c552496d030a498ebe8a5a7d9080f43007fc908cf

Observation 5aff777f-413d-4142-9836-e8741bc8053b · inbound

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos cites this paper.

VRBench: A Benchmark for Multi-Step Reasoning in Long Narrative Videos CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T04:22:56.098198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:22:56.098198Z digest=sha256:993a3a87d243137cfe9f65260cd5cae8f0ad504f4ed5c61082d900bbdbf98329

Observation b1b73b5f-c11d-462a-87d9-feb48b3446d9 · inbound

Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? cites this paper.

Argus Inspection: Do Multimodal Large Language Models Possess the Eye of Panoptes? CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T11:19:05.167967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:19:05.167967Z digest=sha256:7e25caa2afc39a9bcbba635d927719b6d8dd493818e27ee747006a8e6ec43800

Observation ed9c1492-3022-4bba-a86d-b40b58a226c2 · inbound

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique cites this paper.

EduFlow: Advancing MLLMs' Problem-Solving Proficiency through Multi-Stage, Multi-Perspective Critique CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:04:55.471250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:04:55.471250Z digest=sha256:83f7398d9a69597ff29c515fcdb3472c969e6b6a9d3297396c3dc714021658b8

Observation 4666b248-d71e-4a20-91fa-9d307d542e3f · inbound

From UAV Imagery to Agronomic Reasoning: A Multimodal LLM Benchmark for Plant Phenotyping cites this paper.

From UAV Imagery to Agronomic Reasoning: A Multimodal LLM Benchmark for Plant Phenotyping CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:25:59.778943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:11:50.265856Z digest=sha256:9e24ca3bd96bbe3ff669090f7209288cf605fe70c6226651bb307cf46eb28398

Observation 6b981cda-3a3a-4f1d-aa23-b32c0d90ebe5 · inbound

OProver: A Unified Framework for Agentic Formal Theorem Proving cites this paper.

OProver: A Unified Framework for Agentic Formal Theorem Proving CMMMU: A Chinese Massive Multi-discipline Multimodal Understanding Benchmark

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T14:48:23.548793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T14:43:46.517807Z digest=sha256:d8476bfdbb7805170b95fb7e920b96108d43571813cbb01ae6806dadc2afd4cf