Pith. sign in

Paper Citation Record · LEDGER

MedCalc-Bench: Evaluating Large Language Models for Medical Calculations

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2406.12036.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.12036 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:56:15.630445Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:37:30.304423Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b76d961a-4776-4dd0-b14d-602468740e1c · inbound

OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology cites this paper.

OphthBench: A Comprehensive Benchmark for Evaluating Large Language Models in Chinese Ophthalmology MedCalc-Bench: Evaluating Large Language Models for Medical Calculations

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T16:10:24.741545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T16:10:24.741545Z digest=sha256:cd0506e868a0a549340b12c2b4db21c2234cf5f87cb0bfd156c078cd8edfad03

Observation b6141c47-7b77-4238-9954-ddf8308caf09 · inbound

MedGUIDE: Benchmarking Clinical Decision-Making in Large Language Models cites this paper.

MedGUIDE: Benchmarking Clinical Decision-Making in Large Language Models MedCalc-Bench: Evaluating Large Language Models for Medical Calculations

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T20:56:15.630445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:56:15.630445Z digest=sha256:542bc61755aee54ae9a66cbe5b507ebb404c497434b686700123e2a6e6f9ebcf

Observation 5e2eb2f6-8d0e-47b9-ac02-615e2d83ecc7 · inbound

Loong: Synthesize Long Chain-of-Thoughts at Scale through Verifiers cites this paper.

Loong: Synthesize Long Chain-of-Thoughts at Scale through Verifiers MedCalc-Bench: Evaluating Large Language Models for Medical Calculations

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T16:37:52.838834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:37:52.838834Z digest=sha256:70b3a9090a85da5efe187a98f747960b602b2a56f583c21b46525f592a775e2b

Observation a60b1f3f-f7b6-4e8a-8ef6-c75f54fdd677 · inbound

CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows? cites this paper.

CHI-Bench: Can AI Agents Automate End-to-End, Long-Horizon, Policy-Rich Healthcare Workflows? MedCalc-Bench: Evaluating Large Language Models for Medical Calculations

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-20T17:48:48.850449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T17:45:02.896703Z digest=sha256:0026195a74a0925616903b524aa4e733fc71f83448130f69a62cd1a4527a21ee

Observation d192d750-aa67-4c78-b149-3831ff2417b7 · inbound

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches cites this paper.

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches MedCalc-Bench: Evaluating Large Language Models for Medical Calculations

Reference 136

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:56:14.342484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T17:35:01.285534Z digest=sha256:108710963d63788e54e49f3c4659080925b8a6cfb2747510c75212c33445d2fb

Observation 69538a6b-6d28-4190-a67d-76ba7dea4148 · inbound

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches cites this paper.

Reasoning4Sciences: Bridging Reasoning Language Models to All Scientific Branches MedCalc-Bench: Evaluating Large Language Models for Medical Calculations

Reference 148

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:35:34.030925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-01T07:22:25.349398Z digest=sha256:12abf33a2fab85c5ae5feec4362b7c7692579a765b5fb94fdb076473e8d6b732

Observation 46256d06-891e-4594-b086-fe623703490c · inbound

Alignment Defends LLMs from Property Inference Attacks cites this paper.

Alignment Defends LLMs from Property Inference Attacks MedCalc-Bench: Evaluating Large Language Models for Medical Calculations

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:37:30.305978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T17:05:19.853450Z digest=sha256:0866248633071ebbb4628454df741a13bce9785c68fb8d7f35f048894babf810