Pith. sign in

Paper Citation Record · LEDGER

MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2407.10990.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.10990 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:47.005013Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T04:13:53.226694Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fe3c8a0f-4023-446d-9cd5-b0cd3398d2e1 · inbound

Data-Centric Foundation Models in Computational Healthcare: A Survey cites this paper.

Data-Centric Foundation Models in Computational Healthcare: A Survey MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-05-24T04:13:53.229440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-24T04:13:05.328492Z digest=sha256:e5fea338e93dec7665d1d73e9223ae735d9894bbf23de251dc1ae4b4f05ce5cb

Observation ce415c53-29d1-43eb-9314-0b3d91eec600 · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 215

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:47.005013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:47.005013Z digest=sha256:1914cc482dc8f00ac9796ca5a26978e22a7ad2d5b6886f143dd319283330da47

Observation 1c088b28-fcb0-4393-bba6-5d3efbb4fcf1 · inbound

A Novel Evaluation Benchmark for Medical LLMs: Illuminating Safety and Effectiveness in Clinical Domains cites this paper.

A Novel Evaluation Benchmark for Medical LLMs: Illuminating Safety and Effectiveness in Clinical Domains MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T10:46:10.452156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:46:10.452156Z digest=sha256:afa26932ac9e27d6f9e9bbf7f5d2d3321a21988ad57beeeb229982aefe93bbcb

Observation 27d9ffa7-dd00-46aa-baa6-5383c9bec9e3 · inbound

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation cites this paper.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-13T05:07:42.040673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T05:07:42.040673Z digest=sha256:a352a5e500c324cb7345f7dc1f2eb44dc975c96f7c03887c85a3f28b6d9d46aa

Observation cde0045c-0945-4411-bed9-9f84f6eb9ca6 · inbound

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation cites this paper.

MedRealMM: A Real-World Multimodal Benchmark for Chinese Online Medical Consultation MedBench: A Comprehensive, Standardized, and Reliable Benchmarking System for Evaluating Chinese Medical Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T07:43:27.951292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:43:27.951292Z digest=sha256:962e0618f0ebfc51badb4911f0d4d2ad7373ff3e1b4be54cdb307459b36fa69c