Pith. sign in

Paper Citation Record · LEDGER

MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2412.18947.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.18947 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:19:11.288187Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:38:28.969559Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dd44847f-7dc3-4363-a261-9311cbbb39d0 · inbound

Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models cites this paper.

Loki's Dance of Illusions: A Comprehensive Survey of Hallucination in Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 200

Resolution
unresolved
no resolver link, observed 2026-08-07T10:19:11.288187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:19:11.288187Z digest=sha256:f50e36f6801f1cf4f786e7e70ff33fb31a13280a9faa1b46ec55135002eeb3a7

Observation 04ce2e7e-d737-4d09-b44d-e3a85a5e8824 · inbound

A comprehensive taxonomy of hallucinations in Large Language Models cites this paper.

A comprehensive taxonomy of hallucinations in Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-06T05:29:21.056789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:29:21.056789Z digest=sha256:9abc8bb372adf9a523ed74aea065dcb6541677b435751e2c4ff205345a7a0e9f

Observation 038607a1-13e2-41aa-a690-ab5f2d7550b7 · inbound

TerraMAE: Learning Spatial-Spectral Representations from Hyperspectral Earth Observation Data via Adaptive Masked Autoencoders cites this paper.

TerraMAE: Learning Spatial-Spectral Representations from Hyperspectral Earth Observation Data via Adaptive Masked Autoencoders MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T22:26:40.829467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:26:40.829467Z digest=sha256:6db0dfb13a8c25fec8461eb3160273030cd54f2010121d0a88b8f3c26a327b4b

Observation baa9a7b1-57c2-4848-91ed-e9d0ce1365f3 · inbound

Trustworthy Medical Imaging with Large Language Models: A Study of Hallucinations Across Modalities cites this paper.

Trustworthy Medical Imaging with Large Language Models: A Study of Hallucinations Across Modalities MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T22:26:06.534145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:26:06.534145Z digest=sha256:f28bba1f63e67f8d812e297e68051e1a413919405bb9c3a6e1edce43b8b831d5

Observation 456e0b79-1c9f-4d9e-b161-5bfb1f2679aa · inbound

A Multi-Task Evaluation of LLMs' Processing of Academic Text Input cites this paper.

A Multi-Task Evaluation of LLMs' Processing of Academic Text Input MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T19:49:52.789854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:49:52.789854Z digest=sha256:5d8889b5f6249528bd61bc890e0bff645b72cc46c9609cde170662c73ef82366

Observation f75f4bb8-a881-4007-bea9-ad58fdb4dc7d · inbound

PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts cites this paper.

PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:12:44.896323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-19T20:09:47.750043Z digest=sha256:374cc08bb29c3b1cc8dc839cfc8a330a084a8f6d830915d1dfc0dc9485a7e136

Observation 1e92785f-c4af-4572-bda8-1688b7c71bfa · inbound

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models cites this paper.

ERRORQUAKE: Heavy-Tailed Error Severity Distributions in Open-Weight Large Language Models MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T20:40:22.043187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T20:40:22.043187Z digest=sha256:0e99e9f94d142bc4dd90acf90cf0bb3f0451e0e93322897dc588f0c4a736fa20

Observation be9bd705-9c52-4a50-8ad7-ff9c51334350 · inbound

Hallucination in Medical Imaging AI: A Cross-Modality Analytical Framework for Taxonomy, Detection, and Mitigation under Regulatory Constraints cites this paper.

Hallucination in Medical Imaging AI: A Cross-Modality Analytical Framework for Taxonomy, Detection, and Mitigation under Regulatory Constraints MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:38:28.970985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T07:00:54.990637Z digest=sha256:5a275ef967f482073012c0f216d78192e8d59754ed1062dced00d8913f61605e

Observation eff9e784-fec0-4680-a789-70fef4a19c29 · inbound

KnowHal: A Knowledge-Driven Benchmark for Comprehensive Multimodal Hallucination Evaluation cites this paper.

KnowHal: A Knowledge-Driven Benchmark for Comprehensive Multimodal Hallucination Evaluation MedHallBench: A New Benchmark for Assessing Hallucination in Medical Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T12:40:15.241605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:40:15.241605Z digest=sha256:1f7cbf68dfb0a097809f6746c266411a0482a32fd9db0a3bda00797a29f326ef