Pith. sign in

Paper Citation Record · LEDGER

Teaching large language models to reason like expert diagnosticians

As of 21 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2509.12194.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.12194 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T16:43:19.603410Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d49e60be-48dd-4245-848f-02bf6179c7b7 · outbound

This paper cites Case 9431.

Teaching large language models to reason like expert diagnosticians Case 9431

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:16.656179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:16.656179Z digest=sha256:313798d726d74dfa4fbd6da552650971a35faec0dc4c9d9c08cf210a7f9227bb

Observation c84470d1-76b5-44b7-8396-72da7136470d · outbound

This paper cites Building a community of medical learning - A century of case records of the Massachusetts general hospital in the journal.

Teaching large language models to reason like expert diagnosticians Building a community of medical learning - A century of case records of the Massachusetts general hospital in the journal

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:16.772075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:16.772075Z digest=sha256:d2b99b5df1862df61a48946d780537bfab527c5fb66008039439609a157fc40e

Observation 4d48cadf-50bd-44f7-9277-ff238cff7e11 · outbound

This paper cites The clinicopathological conferences (CPCs).

Teaching large language models to reason like expert diagnosticians The clinicopathological conferences (CPCs)

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:16.880999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:16.880999Z digest=sha256:75ae22000795daef73aa71d04dd93c3081a051da741933fe51d522fd12b1dee1

Observation b9f69243-cfda-4b6b-90e5-67093eb29781 · outbound

This paper cites Reasoning foundations of medical diagnosis; symbolic logic, probability, and value theory aid our understanding of how physicians reason.

Teaching large language models to reason like expert diagnosticians Reasoning foundations of medical diagnosis; symbolic logic, probability, and value theory aid our understanding of how physicians reason

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:17.035089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:17.035089Z digest=sha256:374b7e9cc77eb73f7f8f2984e5769397eddf4730e3e56ee5dcf7e27167d0cc6a

Observation f432e112-4a3c-4555-b50c-ba5da1ecd991 · outbound

This paper cites Digital computers and medical logic.

Teaching large language models to reason like expert diagnosticians Digital computers and medical logic

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:17.252755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:17.252755Z digest=sha256:088993434d84ecaa711cfdfe050e8aa30c68420259177dd003504e60bf2c5638

Observation da73b3fc-ed84-486a-b7ce-787eb413aca9 · outbound

This paper cites Digitizing diagnosis: Medicine, minds, and machines in twentieth-century America.

Teaching large language models to reason like expert diagnosticians Digitizing diagnosis: Medicine, minds, and machines in twentieth-century America

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:17.363151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:17.363151Z digest=sha256:2cc9e5ff5aa7042d9cab538a24acf489b35d1dd41f58f9cd3717f5b3f838b284

Observation 6ff66c39-36f4-48c8-84cf-bc1489342f08 · outbound

This paper cites Internist-1, an experimental computer-based diagnostic consultant for general internal medicine.

Teaching large language models to reason like expert diagnosticians Internist-1, an experimental computer-based diagnostic consultant for general internal medicine

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:17.481089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:17.481089Z digest=sha256:35bf215ababaa9ffee6790815015d75841a5605fc13654cc95576b5755086ffa

Observation 07dce628-b8e8-42bc-8c1d-c6ee589d4cec · outbound

This paper cites Differential diagnosis generators: an evaluation of currently available computer programs.

Teaching large language models to reason like expert diagnosticians Differential diagnosis generators: an evaluation of currently available computer programs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:17.603719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:17.603719Z digest=sha256:8f8e608380d3304129c3143c384a90b05d5b10c82c1b3c2c3a62a268877beb2f

Observation 484d4e69-3da5-4dba-bd47-2c6d1bdc9c89 · outbound

This paper cites Evaluation of medical decision support systems (DDX generators) using real medical cases of varying complexity and origin.

Teaching large language models to reason like expert diagnosticians Evaluation of medical decision support systems (DDX generators) using real medical cases of varying complexity and origin

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:17.648517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:17.648517Z digest=sha256:642475eb4d414a6cd37845756185e1bef29250f52c36ebdf6ea8ccf7fb5ed52c

Observation 2522cf0b-a89c-4364-be94-241854f98b0c · outbound

This paper cites Accuracy of a generative artificial intelligence model in a complex diagnostic challenge.

Teaching large language models to reason like expert diagnosticians Accuracy of a generative artificial intelligence model in a complex diagnostic challenge

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:17.698986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:17.698986Z digest=sha256:c5d685d6e34e5620e63411049726ddc3dee5e3532cd412028b459ce9c6727cc8

Observation 5389475e-2c43-4993-b197-d0d1091394a3 · outbound

This paper cites Comparison of frontier open-source and proprietary large language models for complex diagnoses.

Teaching large language models to reason like expert diagnosticians Comparison of frontier open-source and proprietary large language models for complex diagnoses

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:17.747323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:17.747323Z digest=sha256:f3d88a8564ea33a5f3858b11e0c02013c36d64e9bc91641dfa7c4bcb85d5cf4a

Observation 6775379a-521e-446b-8223-67bf6d4573f4 · outbound

This paper cites Towards accurate differential diagnosis with large language models.

Teaching large language models to reason like expert diagnosticians Towards accurate differential diagnosis with large language models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:17.868360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:17.868360Z digest=sha256:eba54ec9b9f3c412d9edcc8aca387e20ec070febb79efb0a43e30b2bbedb8f19

Observation 92ad0103-21d3-499f-bb4c-8ddcfc67bcab · outbound

This paper cites Superhuman performance of a large language model on the reasoning tasks of a physician.

Teaching large language models to reason like expert diagnosticians Superhuman performance of a large language model on the reasoning tasks of a physician

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:18.002881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:18.002881Z digest=sha256:307834db43d37457bd6cc65a5f5312cf48621ed4308b75ba9eb3f94592ed2d6a

Observation c9b044ff-9e91-424f-b685-b6d12ece3f6c · outbound

This paper cites Sequential Diagnosis with Language Models.

Teaching large language models to reason like expert diagnosticians Sequential Diagnosis with Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:18.083351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:18.083351Z digest=sha256:092af6d895ab1c8eee1ea1963608348d6af6ab840cf86a9f61e311ac85234dc5

Observation 26b49156-b3b9-4a69-9910-6c65d4d16fe4 · outbound

This paper cites It’s time to bench the medical exam benchmark.

Teaching large language models to reason like expert diagnosticians It’s time to bench the medical exam benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:18.268216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:18.268216Z digest=sha256:3ebb886da287e264c7ac97bfc87b72e71e8737db7396bb9039d645d4aed24f89

Observation d9a03f10-22e2-4e12-beb6-ea6fc0b98a45 · outbound

This paper cites When it comes to benchmarks, humans are the only way.

Teaching large language models to reason like expert diagnosticians When it comes to benchmarks, humans are the only way

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:18.373831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:18.373831Z digest=sha256:5bc5d2dcefc50156346fe6bad5ef1f291e283c21bed9d9e15f13e35e5119228c

Observation 8efba316-5ffa-4bef-8522-c037225f4c7e · outbound

This paper cites Representation and misdiagnosis of dark skin in a large-scale visual diagnostic challenge.

Teaching large language models to reason like expert diagnosticians Representation and misdiagnosis of dark skin in a large-scale visual diagnostic challenge

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:18.524477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:18.524477Z digest=sha256:10633d64b0a6079c0ad19b98079163a6434fff1e1c819b8fb100ca80acb74c78

Observation 4477b600-045a-468a-8122-68b3649cbfe2 · outbound

This paper cites OpenAlex: A fully-open index of scholarly works, authors, venues, institutions, and concepts.

Teaching large language models to reason like expert diagnosticians OpenAlex: A fully-open index of scholarly works, authors, venues, institutions, and concepts

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:18.659456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:18.659456Z digest=sha256:54271fa60c2bedbc39d8f4b51556113af88a11160c9e0a22f6df9d9cb5c52a26

Observation 42f775e5-14d9-43ce-b516-7e1288372353 · outbound

This paper cites Hippocratic corpus.

Teaching large language models to reason like expert diagnosticians Hippocratic corpus

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:18.758720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:18.758720Z digest=sha256:44d5d33141733521afd6f514b9bfe7ee10892a9d54f5601011c82aaea0152919

Observation c409f1b7-a516-478b-8f56-419faac26ac9 · outbound

This paper cites Hippocratic.

Teaching large language models to reason like expert diagnosticians Hippocratic

Reference 20

Resolution
verified exact
doi, observed 2026-08-04T16:48:34.842283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-04T16:43:18.950423Z digest=sha256:4613ff612f54dbc04ff76ff7082217cc800e561514f486dad4cfab98e51dc573

Observation 54159444-7804-4a23-8072-f4ddfe36fa9c · outbound

This paper cites MEDITRON-70B: Scaling Medical Pretraining for Large Language Models.

Teaching large language models to reason like expert diagnosticians MEDITRON-70B: Scaling Medical Pretraining for Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:19.005255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:19.005255Z digest=sha256:12a90f4275e1a8ababa9bc573a8a33481432182519ce3cf4c655c3a46a0a621d

Observation 0ca4b183-dea0-4650-b829-082df593e64e · outbound

This paper cites MedGemma Technical Report.

Teaching large language models to reason like expert diagnosticians MedGemma Technical Report

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:19.113616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:19.113616Z digest=sha256:35a8a45be78e627ff3d914d0b6331961d95f4a5be50622459d0ff59481bca258

Observation f758eee5-9639-439a-8804-f8cb20507442 · outbound

This paper cites Limitations of learning new and updated medical knowledge with commercial fine-tuning large language models.

Teaching large language models to reason like expert diagnosticians Limitations of learning new and updated medical knowledge with commercial fine-tuning large language models

Reference 23

Resolution
verified exact
doi, observed 2026-08-04T16:48:34.579892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-04T16:43:19.193492Z digest=sha256:04d440e4acb1115c9931956fbe21064448fb78d1f098f4d004f1e40820c58686

Observation a68c6e94-52e0-42b4-b3cd-851a038d854c · outbound

This paper cites Evaluating the effectiveness of biomedical fine-tuning for large language models on clinical tasks.

Teaching large language models to reason like expert diagnosticians Evaluating the effectiveness of biomedical fine-tuning for large language models on clinical tasks

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:19.264742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:19.264742Z digest=sha256:b47007b30bc79e1d385b04bedee550650538db21a8f98a229bb2f95c4c91d81d

Observation 027679e9-950c-49ad-9b4a-30379621858b · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Teaching large language models to reason like expert diagnosticians Chain-of-thought prompting elicits reasoning in large language models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:19.326115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:19.326115Z digest=sha256:63a3463c71d7eda3acd6a08113f717d63c71467b5e83c1c99e84e9a7240d40d0

Observation 59938cd2-eb3b-481c-9f91-ce3c1fd65b46 · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

Teaching large language models to reason like expert diagnosticians Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:19.371532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:19.371532Z digest=sha256:14fc9f42641ee83f1225fee6447e5a11a68f3570cc02bbbae038665341748292

Observation 13fa9db0-d181-4d40-b6b8-b06da1206676 · outbound

This paper cites From Medprompt to o1: Exploration of Run-Time Strategies for Medical Challenge Problems and Beyond.

Teaching large language models to reason like expert diagnosticians From Medprompt to o1: Exploration of Run-Time Strategies for Medical Challenge Problems and Beyond

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:19.424876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:19.424876Z digest=sha256:c332b75bf75c9c2460e1b18cbf2fb1c2ed4278f54c272a5e90ec0a62fc2734dd

Observation 2dd01417-12a7-4027-be51-2950ac69ed3c · outbound

This paper cites The Bitter Lesson [Internet].

Teaching large language models to reason like expert diagnosticians The Bitter Lesson [Internet]

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:19.458759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:19.458759Z digest=sha256:670b4513527eeff84b79110bdce723217e851b3c3b099956d4ba6356a90d14c8

Observation 526563e2-3663-447a-a733-8f4dc7d40bdc · outbound

This paper cites LongHealth: A question answering benchmark with long clinical documents.

Teaching large language models to reason like expert diagnosticians LongHealth: A question answering benchmark with long clinical documents

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:19.545635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:19.545635Z digest=sha256:c687acbb8c5dcf1ab3f7b2a95d8a9764658d605010fa46fb33fb23c2b4928527

Observation cb5bc586-7c22-45fa-bb24-6b26abd0cdbf · outbound

This paper cites Health system-scale language models are all-purpose prediction engines.

Teaching large language models to reason like expert diagnosticians Health system-scale language models are all-purpose prediction engines

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T16:43:19.603410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:43:19.603410Z digest=sha256:bd5082bd3c5454e5a2831f534c5a23c0a0bd370fe4c425d80989a2888ddf91e8

Pith citing papers

No inbound Pith citation observations are available.