Pith. sign in

Paper Citation Record · LEDGER

GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2408.03361.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2408.03361 v7

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:01:39.697849Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T12:56:57.641234Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9992be07-b41d-4250-8c0d-7bb53dc2a890 · inbound

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI cites this paper.

GMAI-VL & GMAI-VL-5.5M: A Large Vision-Language Model and A Comprehensive Multimodal Dataset Towards General Medical AI GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T15:16:27.659001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:16:27.659001Z digest=sha256:77291e11f445017e364330c126a9c8550167f57dfb043909affa1f54fa23a31f

Observation 653d52f7-f8b6-4a0d-9ffd-b65b81a91141 · inbound

MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs cites this paper.

MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 130

Resolution
unresolved
no resolver link, observed 2026-08-12T14:31:37.234152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:31:37.234152Z digest=sha256:721660344623a975aea1d6b38b3c398beb997340b59b7f33d98dcc851a04ef51

Observation c0c409ec-f4cd-4254-b5ac-a23523ad3eec · inbound

Enhanced Multimodal RAG-LLM for Accurate Visual Question Answering cites this paper.

Enhanced Multimodal RAG-LLM for Accurate Visual Question Answering GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T23:10:33.010777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T23:10:33.010777Z digest=sha256:10ce890befd127f5fe0aad0cde782c3bcd1f8798d1cbd77d9ffa251be9278dd4

Observation 05c95cc0-e522-4712-a287-87ecab6a8ef2 · inbound

EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery cites this paper.

EndoChat: Grounded Multimodal Large Language Model for Endoscopic Surgery GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T18:24:41.806356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:24:41.806356Z digest=sha256:983331a7ede920a411ba04a35e27c149ee6fed8d652372321262c78bcf7202f5

Observation 1536fbc9-d5a1-4d3d-bd9a-969cd4e1f9d7 · inbound

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models cites this paper.

ZeroBench: An Impossible Visual Benchmark for Contemporary Large Multimodal Models GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T20:55:05.830828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:55:05.830828Z digest=sha256:545cad84f41b6ded48aa802f2f1610ff8dcdfb668b694e0b69a8f783c5022867

Observation ba242164-361b-49a3-96e9-e71874d61941 · inbound

Unveiling the Lack of LVLM Robustness to Fundamental Visual Variations: Why and Path Forward cites this paper.

Unveiling the Lack of LVLM Robustness to Fundamental Visual Variations: Why and Path Forward GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-16T11:01:39.697849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:01:39.697849Z digest=sha256:76e7a3edfdacd74f98f711b96bc0cca0e0fe693d74c53c5dd695c109da618bd8

Observation 6b9d6fb6-13d8-487f-bef0-ce7d69091723 · inbound

Advancing Conversational Diagnostic AI with Multimodal Reasoning cites this paper.

Advancing Conversational Diagnostic AI with Multimodal Reasoning GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T23:45:13.590280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:45:13.590280Z digest=sha256:a77c48c4e07b72bff9c2d300c9e7aac6538abb17b0b1beb9cf19e76c5e7e409a

Observation 3b84247e-1e81-463f-b92d-a5e0a6498418 · inbound

SeedBench: A Multi-task Benchmark for Evaluating Large Language Models in Seed Science cites this paper.

SeedBench: A Multi-task Benchmark for Evaluating Large Language Models in Seed Science GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:21:14.656584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:21:14.656584Z digest=sha256:21f4a04a6829c9aa4cbe48a20e15a9ccfe0350ec7b05590e6ef08a3c1f4e07b0

Observation 1119b38b-4acd-4913-b05a-11710da19228 · inbound

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book cites this paper.

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:02.897723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:02.897723Z digest=sha256:4e1df80a5064b66665d7cf9f3a082f9cc4701a467a867c5ef3d009ecc9502bb2

Observation faf179a8-d2c9-4552-931e-acdd0bd8466a · inbound

MANBench: Is Your Multimodal Model Smarter than Human? cites this paper.

MANBench: Is Your Multimodal Model Smarter than Human? GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:33.269754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:04:33.269754Z digest=sha256:5e5b250324055319a21e74ed9ca8f5d55d2c98347382a214ff1cf481fe36e24b

Observation 8cae3ded-b0c8-4b9c-94d9-bffe5b453791 · inbound

A Comprehensive Survey of Deep Research: Systems, Methodologies, and Applications cites this paper.

A Comprehensive Survey of Deep Research: Systems, Methodologies, and Applications GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T00:48:12.685232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:48:12.685232Z digest=sha256:567135eac7a9b6cbcf61453f02bbda2a3d92ccb8ad922440a0264a7b1f46ab1d

Observation dbda4ea0-345f-4754-8794-444549350d65 · inbound

Constructing Ophthalmic MLLM for Positioning-diagnosis Collaboration Through Clinical Cognitive Chain Reasoning cites this paper.

Constructing Ophthalmic MLLM for Positioning-diagnosis Collaboration Through Clinical Cognitive Chain Reasoning GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T14:51:09.584692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:51:09.584692Z digest=sha256:ce7d1bc1733812de24abda96ad656ef0a015d992c76c16e5b3a2b27f1e441f29

Observation 301888e4-a610-4cda-8fa9-3cbee294a3ef · inbound

AMVICC: A Novel Benchmark for Cross-Modal Failure Mode Profiling for VLMs and IGMs cites this paper.

AMVICC: A Novel Benchmark for Cross-Modal Failure Mode Profiling for VLMs and IGMs GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T09:35:26.655659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T09:35:26.655659Z digest=sha256:19437f2024f40fb77c8d785866d6b32c9054263809eedafabd120b9ce0d73a45

Observation 4048306c-be35-42aa-bd1a-50e78abf1bcc · inbound

CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs cites this paper.

CXR-ContraBench: Benchmarking Negated-Option Attraction in Medical VLMs GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:09.736536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T14:49:53.357083Z digest=sha256:e9e1fa54ba375be6a5bdd2dc29d6fda2c51725f12d4f7611785d1b4f84c58fe5

Observation 361d4519-3a74-4f49-a9bf-5a7e8b7e51c7 · inbound

MMBU: A Massive Multi-modal Biomedical Understanding Benchmark to Probe the Perception Capabilities of Vision-Language Models cites this paper.

MMBU: A Massive Multi-modal Biomedical Understanding Benchmark to Probe the Perception Capabilities of Vision-Language Models GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:56:57.642951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-28T01:39:30.851276Z digest=sha256:ca0eb81e98f796c15c451a23b84f525f6263a155352521f817800749c57b22b3

Observation e0052204-d841-44ec-97ea-fc217cd4c0cc · inbound

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation cites this paper.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-13T05:07:42.040673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:07:42.040673Z digest=sha256:252f70f2220fb7015297652930a6eac87f061834a26f78e2dcd0b877003a0e8f

Observation 3021863a-813b-46d5-95f9-a8515395b7e7 · inbound

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation cites this paper.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:30.865915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:30.865915Z digest=sha256:470c2fddddc9efb593fd4b341136d05b8baceb20f79fba75ad1f3568301c1801

Observation 0c79f66d-e09c-4af3-a27e-78e5f01b2b7f · inbound

Marking the Wrong Symptoms: Evaluating LLM Watermarks in Medical Texts cites this paper.

Marking the Wrong Symptoms: Evaluating LLM Watermarks in Medical Texts GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-02T13:52:07.866376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:52:07.866376Z digest=sha256:9ce73bcebf532eba4464e034832ae4634a9611921e034fb24bbd31911424b36f

Observation 2246ef7e-d5c6-45be-bcc4-6ebde0bb7290 · inbound

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA cites this paper.

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T05:36:45.204851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:36:45.204851Z digest=sha256:9edce2306c3d344fa8869b15a49168fba045d8b1ff089909bdb85d579cc41fc7

Observation 90755ed4-8b29-4e2b-bfcf-421995badfa9 · inbound

How Good are Foundation Models in Longitudinal MRI Disease Progression Reasoning? cites this paper.

How Good are Foundation Models in Longitudinal MRI Disease Progression Reasoning? GMAI-MMBench: A Comprehensive Multimodal Evaluation Benchmark Towards General Medical AI

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T14:04:08.989376Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T14:04:08.989376Z digest=sha256:28faad32dd8f5e2a8801d92f6861dfdab4d5e22dd9440e2d915a8e075769f7a6