Pith. sign in

Paper Citation Record · LEDGER

MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2407.10990.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.10990 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:47.005013Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T04:13:53.226694Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fe3c8a0f-4023-446d-9cd5-b0cd3398d2e1 · inbound

Data-Centric Foundation Models in Computational Healthcare: A Survey cites this paper.

Data-Centric Foundation Models in Computational Healthcare: A Survey MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-05-24T04:13:53.229440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-24T04:13:05.328492Z digest=sha256:03606e832bb4ac5b4b75328cb67258dac79efe52924d2d89bb242c0b4998332d

Observation ce415c53-29d1-43eb-9314-0b3d91eec600 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 215

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:47.005013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:47.005013Z digest=sha256:3efd36a4d15c4988556ff40e84db8f203f8c3da0a3bb28c1de32f5ff230c9116

Observation 1c088b28-fcb0-4393-bba6-5d3efbb4fcf1 · inbound

A Novel Evaluation Benchmark for Medical LLMs: Illuminating Safety and Effectiveness in Clinical Domains cites this paper.

A Novel Evaluation Benchmark for Medical LLMs: Illuminating Safety and Effectiveness in Clinical Domains MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:46:10.452156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:46:10.452156Z digest=sha256:f205c46be7ae7d3a5e7eba22026652c8a13ed10357dc3f13c5655e78441c7b8e

Observation 27d9ffa7-dd00-46aa-baa6-5383c9bec9e3 · inbound

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation cites this paper.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T05:07:42.040673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:07:42.040673Z digest=sha256:f9334b09f66d3d4730f2a97348e4feb43f9f78ddb3c575e4bcd310625e8de979

Observation cde0045c-0945-4411-bed9-9f84f6eb9ca6 · inbound

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation cites this paper.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.951292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.951292Z digest=sha256:7db07c5c60153ffe5d925667601ed8b09a2bf518f3b7863c9cac6969aef16565