Pith. sign in

Paper Citation Record · LEDGER

Information-seeking failures of large language models in agentic clinical reasoning

As of 21 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2607.10275.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.10275 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-14T12:59:50.759055Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact5
  • verified fuzzy0
  • unresolved44
  • parse uncertain0
  • malformed identifier4
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 01ddeb8e-6ddc-4f45-ab24-21fd41bc2d73 · outbound

This paper cites Application of precision medicine in clinical routine in haematology-challenges and opportunities.J.

Information-seeking failures of large language models in agentic clinical reasoning Application of precision medicine in clinical routine in haematology-challenges and opportunities.J

Reference 1

Resolution
verified exact
doi, observed 2026-07-14T13:00:29.213431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:8b2367102bc91ff59a2e34bcfeba37fd5fb743dd15bc1758d93d0b5875713599

Observation 08d4193e-09f0-41fb-9388-80965443c67b · outbound

This paper cites Large language models encode clinical knowledge.Nature, 620(7972):172–180, August 2023.

Information-seeking failures of large language models in agentic clinical reasoning Large language models encode clinical knowledge.Nature, 620(7972):172–180, August 2023

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:9df058d6522e57ea452ff616b753035d6cf2dbf184237725603e053788f93284

Observation 249ab167-8c1b-45a4-854d-47face9fd0fa · outbound

This paper cites Capabilities of GPT-4 on Medical Challenge Problems.

Information-seeking failures of large language models in agentic clinical reasoning Capabilities of GPT-4 on Medical Challenge Problems

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:47722d5c97701bcc9ea45b03a392fc72414577386021b5831b8e363fe3bb48eb

Observation c63f30bf-d55d-407e-bb53-87e89023f522 · outbound

This paper cites How does ChatGPT perform on the united states medical licensing examination (USMLE)? theimplicationsoflargelanguagemodelsformedicaleducationandknowledgeassessment.JMIR Med.

Information-seeking failures of large language models in agentic clinical reasoning How does ChatGPT perform on the united states medical licensing examination (USMLE)? theimplicationsoflargelanguagemodelsformedicaleducationandknowledgeassessment.JMIR Med

Reference 4

Resolution
malformed identifier
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:83cb6a6b75de7d335dab58aa66ca417c9364a14abb8e8081bc142305f2bb71b9

Observation ab63eb7d-17c0-4e8e-855a-17e77ad82402 · outbound

This paper cites GPT-4 for information retrieval and comparison of medical oncology guidelines.NEJM AI, 1(6), May 2024.

Information-seeking failures of large language models in agentic clinical reasoning GPT-4 for information retrieval and comparison of medical oncology guidelines.NEJM AI, 1(6), May 2024

Reference 5

Resolution
verified exact
doi, observed 2026-07-14T13:00:29.246671Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:c0c0626c7d0aa7f193c282c2ed817285d5f2e0235a9ad01694c7a9d338312f82

Observation b7589acf-32a7-4a82-b347-df8688d47f78 · outbound

This paper cites Large language models should be used as scientific reasoning engines, not knowledge databases.Nat.

Information-seeking failures of large language models in agentic clinical reasoning Large language models should be used as scientific reasoning engines, not knowledge databases.Nat

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:0c2019da8aa1411a9bf9b88af98085ad3ec28e9b735dfc18521f00a1854c40fb

Observation 245d5314-64b8-4c19-afd6-eadaaa6c933a · outbound

This paper cites Large language models in medicine.Nat.

Information-seeking failures of large language models in agentic clinical reasoning Large language models in medicine.Nat

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:f23b6c692acfd323f9741a5a35a2b2897cecdeb419aa54c059ffa0c0c2f1b32d

Observation e66257bb-48c3-4fd0-b562-abfd5ac36ea8 · outbound

This paper cites Multimodal biomedical AI.Nat.

Information-seeking failures of large language models in agentic clinical reasoning Multimodal biomedical AI.Nat

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:ea4d5199277cd08692c986e81a12f3b846124076a5959a63af19767a11136cad

Observation 016d41ac-18a6-43a3-9bcc-42692f6935b6 · outbound

This paper cites Artificialintelligenceformultimodaldataintegrationinoncology.Cancer Cell,40(10):1095–1110, October 2022.

Information-seeking failures of large language models in agentic clinical reasoning Artificialintelligenceformultimodaldataintegrationinoncology.Cancer Cell,40(10):1095–1110, October 2022

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:2d0a5f659e849a9c127f857517d12ed536c47a7f81fe40e6f74cbe075aa63a74

Observation 7abf7370-6da5-4886-92e6-ec165f2382be · outbound

This paper cites Foundation models for generalist medical artificial intelligence.Nature, 616 (7956):259–265, April 2023.

Information-seeking failures of large language models in agentic clinical reasoning Foundation models for generalist medical artificial intelligence.Nature, 616 (7956):259–265, April 2023

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:25d70d14a5f14f4003800364c7ab4403b6a569159bbb84293956a88ec458e006

Observation 4795a23b-8fc5-42be-9435-c527b67652bf · outbound

This paper cites Towards generalist biomedical AI.NEJM AI, 1(3), February 2024.

Information-seeking failures of large language models in agentic clinical reasoning Towards generalist biomedical AI.NEJM AI, 1(3), February 2024

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:55ebee67450063ba3081079c15c2e83dd1e42e72b0ad77a9d808f2aff4c6425f

Observation 6c736a9d-d193-4f6e-a450-ceae401273b2 · outbound

This paper cites Development and validation of an autonomous artificial intelligence agent for clinical decision-making in oncology.Nat.

Information-seeking failures of large language models in agentic clinical reasoning Development and validation of an autonomous artificial intelligence agent for clinical decision-making in oncology.Nat

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:a8e12246d03aeff539dc494844a4c96a2241d546fd2ab648c9288b571192a341

Observation 12d233ed-84f3-4c2a-add5-e20419c2c7d8 · outbound

This paper cites an unresolved cited work.

Information-seeking failures of large language models in agentic clinical reasoning Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:e3376763f7def2638b2c5df7e7850c522a958fc1f9b0068cedf9a312a9c3d259

Observation beda8c3b-7589-48d1-8b18-db40e55119c1 · outbound

This paper cites HowAIagentswillchange cancer research and oncology.Nat.

Information-seeking failures of large language models in agentic clinical reasoning HowAIagentswillchange cancer research and oncology.Nat

Reference 14

Resolution
malformed identifier
doi_truncated, observed 2026-07-14T13:00:29.195322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:cc4bfd15124ba5ebd2c5a81b5faf529ffc3c1fe07d791d264f78a21de1078620

Observation 15b0a9de-2063-4e96-b0c6-a1d4672050b9 · outbound

This paper cites Holistic evaluation of large language models for medical tasks with MedHELM.Nat.

Information-seeking failures of large language models in agentic clinical reasoning Holistic evaluation of large language models for medical tasks with MedHELM.Nat

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:ddac44c91f11d2694404ffac6da0150a70fe398c3c4a83073d3313ba77900c40

Observation 94bab2c2-8638-4538-a475-cc4534c79730 · outbound

This paper cites Comparative benchmarking of the DeepSeek large language model on medical tasks and clinical reasoning.Nat.

Information-seeking failures of large language models in agentic clinical reasoning Comparative benchmarking of the DeepSeek large language model on medical tasks and clinical reasoning.Nat

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:dc5ed136a0fbd526684f6339089025528c1f58a0fc9086b306fc2e327116f0e9

Observation d322678b-79a5-4b1a-800a-989cc00a8df3 · outbound

This paper cites Toward expert-level medical question answering with large language models.Nat.

Information-seeking failures of large language models in agentic clinical reasoning Toward expert-level medical question answering with large language models.Nat

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:c83bcf503dd78c8df5869366c9a98b6df001a3d41accb8a7667ab0af004f6fc7

Observation d9e46336-7f5b-464e-a1eb-b1c859f97ae5 · outbound

This paper cites On the Planning Abilities of Large Language Models : A Critical Investigation.

Information-seeking failures of large language models in agentic clinical reasoning On the Planning Abilities of Large Language Models : A Critical Investigation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:0865984310116fe3bdfb5b26ecd3a6e342ee97fa480c2ef1e4d2ff4df9bbf64e

Observation 8d3b0e49-3111-4c1c-bbee-1b090868200d · outbound

This paper cites Large Language Models for Planning: A Comprehensive and Systematic Survey.

Information-seeking failures of large language models in agentic clinical reasoning Large Language Models for Planning: A Comprehensive and Systematic Survey

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:7cd630ba9d9e84af5af7fc915e99bcfcc89098a24489266743d1755da80f6a8b

Observation 3fd4a10b-2c63-41a6-83e1-ca355f723927 · outbound

This paper cites A Survey on LLM-as-a-Judge.

Information-seeking failures of large language models in agentic clinical reasoning A Survey on LLM-as-a-Judge

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:6aab016193bac3cc366f9fd1ea79086d3d33a0dbf4d2e16dbc425efd56b6a10b

Observation 0045cc91-7f33-4fd0-bc52-ba3a3a0eb115 · outbound

This paper cites Development of a clinical reasoning documentation assessment tool for resident and fellow admission notes: A shared mental model for feedback.J.

Information-seeking failures of large language models in agentic clinical reasoning Development of a clinical reasoning documentation assessment tool for resident and fellow admission notes: A shared mental model for feedback.J

Reference 21

Resolution
verified exact
doi, observed 2026-07-14T13:00:29.242151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:f0ae829660c126cc197940d647d0df3a077f68a5737fd82536f800693bf43367

Observation d7e73094-f550-4595-a78f-ae31c0cb4a6f · outbound

This paper cites A universal model of diagnostic reasoning.Acad.

Information-seeking failures of large language models in agentic clinical reasoning A universal model of diagnostic reasoning.Acad

Reference 22

Resolution
verified exact
doi, observed 2026-07-14T13:00:29.200573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:91d6deea8058697258e5afeeef379f9cf0c176f2c77f210452692b86aee32d7e

Observation f90c4402-4042-4b57-a828-2409f91f8b9e · outbound

This paper cites AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments.

Information-seeking failures of large language models in agentic clinical reasoning AgentClinic: a multimodal agent benchmark to evaluate AI in simulated clinical environments

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:f44d2cdfc30c1c6cda42052dd1f5d62c655bfc97e059563aee5244621ae8ce9f

Observation 489483b9-e808-443e-99d2-e3706d43130f · outbound

This paper cites MedDialogRubrics: A comprehensive benchmark and evaluation framework for multi-turn medical consultations in large language models.

Information-seeking failures of large language models in agentic clinical reasoning MedDialogRubrics: A comprehensive benchmark and evaluation framework for multi-turn medical consultations in large language models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:8e9e8d07ad5c9d3c8aca268f07e69585e4d9706d869d3808385f2e550931d18d

Observation a8fa9123-133f-4884-a662-eb59d587dd37 · outbound

This paper cites Self-evolving multi-agent simulations for realistic clinical interactions.

Information-seeking failures of large language models in agentic clinical reasoning Self-evolving multi-agent simulations for realistic clinical interactions

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:39ded01930dd4bf3f872a4b3348721f8d90c89f43b91dbdd9425dd5230d67215

Observation 0f8a1ec1-1998-4222-9a01-739385786efb · outbound

This paper cites Lost in the Middle: How Language Models Use Long Contexts.

Information-seeking failures of large language models in agentic clinical reasoning Lost in the Middle: How Language Models Use Long Contexts

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:ee50c5a5579f96c34e90f4369e117dc7b04efe91b67215743528adbc766cc754

Observation d772dedf-a095-4830-98ae-fcf4aedc8f34 · outbound

This paper cites Systematic Evaluation of Long-Context LLMs on Financial Concepts.

Information-seeking failures of large language models in agentic clinical reasoning Systematic Evaluation of Long-Context LLMs on Financial Concepts

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:f660b1573b32c80fcae1432e2dd15e5fb21bd6b38343fde4d321816491ade3c1

Observation 48e19e03-25cd-487a-8de0-5b1c0f143e6f · outbound

This paper cites LLMs Get Lost In Multi-Turn Conversation.

Information-seeking failures of large language models in agentic clinical reasoning LLMs Get Lost In Multi-Turn Conversation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:64c1d84b8bc666ac7bfe1aec1c658413c6d3fcbacbfae9dc9d897c155ca17982

Observation 97cd46cd-c140-45ce-acbb-7de9a2332c51 · outbound

This paper cites Medical reasoning in LLMs: an in-depth analysis of DeepSeek R1.Front.

Information-seeking failures of large language models in agentic clinical reasoning Medical reasoning in LLMs: an in-depth analysis of DeepSeek R1.Front

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:eed69461b76f655c1b89f1328f41dc1277969ee121cc07bec622ac760f505320

Observation bb547611-d6f5-48cc-a3b2-f5aa674dd0f7 · outbound

This paper cites Addressing cognitive bias in medical language models.

Information-seeking failures of large language models in agentic clinical reasoning Addressing cognitive bias in medical language models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:2dcdb5e2d7a64737670ee971049efa5b39b869d862177515e84c370082636242

Observation 6b7a4404-a093-4080-918d-c669f031c4e6 · outbound

This paper cites Mitigating cognitive biases in clinical decision-making through multi-agent conversations using large language models: Simulation study.J.

Information-seeking failures of large language models in agentic clinical reasoning Mitigating cognitive biases in clinical decision-making through multi-agent conversations using large language models: Simulation study.J

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:f96101c4e9ce40c47fef80c6561f264c813eb925486e49b4365e5702b447198d

Observation 347341cf-d326-4334-b7b8-4df292f772ca · outbound

This paper cites Medical Reasoning in the Era of LLMs: A Systematic Review of Enhancement Techniques and Applications.

Information-seeking failures of large language models in agentic clinical reasoning Medical Reasoning in the Era of LLMs: A Systematic Review of Enhancement Techniques and Applications

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:873fbd23d2e3887616729c2bbb4dc957e826b5bcfc02ad095509897f68903eef

Observation 7aef1882-e036-49d3-9afd-0bd586b0d8fb · outbound

This paper cites Automating expert-level medical reasoning evaluation of large language models.NPJ Digit.

Information-seeking failures of large language models in agentic clinical reasoning Automating expert-level medical reasoning evaluation of large language models.NPJ Digit

Reference 34

Resolution
verified exact
doi, observed 2026-07-14T13:00:29.200799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:1d691a5d3fe627269ae707cd00ebdbd0ebed05ab26bd2b3df4713d1f002bcfbf

Observation 35f443b9-0625-4e54-9c38-d5b5c6369aaf · outbound

This paper cites morphologic remission.

Information-seeking failures of large language models in agentic clinical reasoning morphologic remission

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:c6ae1ecb704b4ced16893dcf4571af3d67224dbe88f948637e8633d8fe4fe1ca

Observation 0adc449b-a5ac-4ab9-befc-8fcadb2537eb · outbound

This paper cites an unresolved cited work.

Information-seeking failures of large language models in agentic clinical reasoning Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:db8c121dde80c5f27e8b4d46ea8b5a0a5baaf8fe56589efe0306a63c8d4108d8

Observation c180bc20-dea2-4279-99ac-099936e29e1f · outbound

This paper cites an unresolved cited work.

Information-seeking failures of large language models in agentic clinical reasoning Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:c1f617896e047abc7920b658661eb2569a57d024d3191f57f62d4dd7a913e231

Observation 7c7efa82-2739-432c-997f-46f4a315c0df · outbound

This paper cites Solve: The patient is in septic shock secondary to presumed community-acquired pneumonia. Immediate management includes.

Information-seeking failures of large language models in agentic clinical reasoning Solve: The patient is in septic shock secondary to presumed community-acquired pneumonia. Immediate management includes

Reference 38

Resolution
malformed identifier
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:e1afd75bc68a17df537aa55c064e7eda26c9ab70cade79afb633ad6d6b4d3a26

Observation c2b0958b-921f-47c2-8c10-beb8cce3ff32 · outbound

This paper cites an unresolved cited work.

Information-seeking failures of large language models in agentic clinical reasoning Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:849da4f8155870566a8a264597ceb03e4f92fb224685e54ffae49ddca0031b50

Observation 25b3f52a-b648-4856-b88a-8cd79dbb001d · outbound

This paper cites an unresolved cited work.

Information-seeking failures of large language models in agentic clinical reasoning Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:64e90b88b2b198bba4fb26197cfc10bf016aaafd2a1fac494ad74f5938d4d03f

Observation f58273a9-b039-4fd7-9b45-22c53cf33c46 · outbound

This paper cites inactive post-treatment myeloma lesions.

Information-seeking failures of large language models in agentic clinical reasoning inactive post-treatment myeloma lesions

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:20c81ad43786cfd6e30aac17d509df08ab06825f6e74f7fad0be9484b288293d

Observation 34e07c82-afae-4173-972e-9d713f8008e0 · outbound

This paper cites next treatment options.

Information-seeking failures of large language models in agentic clinical reasoning next treatment options

Reference 42

Resolution
malformed identifier
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:073639e24ad354cd3a86999e41e7532e16f3efc2a3d3d051bf3f8be53ea8b567

Observation 999835f6-7801-431d-b376-eee328b81e4d · outbound

This paper cites an unresolved cited work.

Information-seeking failures of large language models in agentic clinical reasoning Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:3acabc058de4dc4fd65e30b47845d3bd9bdd1ee20b97c3f1ce87e92f3739716b

Observation 38d2ae96-4a5c-4dee-ae89-a02e0e0599e0 · outbound

This paper cites an unresolved cited work.

Information-seeking failures of large language models in agentic clinical reasoning Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:6bf029f6bdeea1b6e5774f6db18903f0d55cc149e1051e8b9b06d4f1721f255f

Observation 8d6e751a-8bb5-40c8-8805-7df80bcee946 · outbound

This paper cites This might fit.

Information-seeking failures of large language models in agentic clinical reasoning This might fit

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:62adb5bfef59a15ba20ab23439ab86aa7d069022198fe35f1cae1866ff467edb

Observation 24e36c0f-9632-4a00-8fe0-8fea358c42dd · outbound

This paper cites Coordinate anticoagulation management for procedural hold.

Information-seeking failures of large language models in agentic clinical reasoning Coordinate anticoagulation management for procedural hold

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:85b61d785eb908174d1fb2befa1328f81b3b9e3d4676d69919f3bd9c9e31ba2a

Observation 3396f878-63b0-494e-81b0-215dc02e5dc1 · outbound

This paper cites an unresolved cited work.

Information-seeking failures of large language models in agentic clinical reasoning Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:1ef3d7c943039cb6c46e9038e6ee60af369db37673500f339ef86bc4bd031d09

Observation e3c00399-7fbd-4ec9-bcaa-aa36b0a70e66 · outbound

This paper cites pneumoniae and VRE E.

Information-seeking failures of large language models in agentic clinical reasoning pneumoniae and VRE E

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:27b466ff690a9a508d0170571cdae5487e3c43219ef6c29afde2d98bc8340822

Observation b6f48dcc-91a0-4d36-b5e3-76b5a6a8fc1d · outbound

This paper cites partially_correct.

Information-seeking failures of large language models in agentic clinical reasoning partially_correct

Reference 49

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:121d026503aebdeb47c259b3780ce9a09698789da8f7fa32a4c7df760a7caf62

Observation 634829b0-6982-48c5-95b8-ca0f389339ae · outbound

This paper cites an unresolved cited work.

Information-seeking failures of large language models in agentic clinical reasoning Unresolved cited work

Reference 50

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:6789b96c8d2661120e5d78fa1b387264b598a421924787a5ccdbf76c632dc553

Observation aa82e435-f4f0-4446-8ee5-e018f8ea3c2f · outbound

This paper cites After reviewing treatment history (5 lines including ASCT x2, CAR-T, multiple PI/IMiD/anti-CD38 combinations), correctly identifies penta-refractory status.

Information-seeking failures of large language models in agentic clinical reasoning After reviewing treatment history (5 lines including ASCT x2, CAR-T, multiple PI/IMiD/anti-CD38 combinations), correctly identifies penta-refractory status

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:4c2d50a7654ba9fb43f97f086939b44cf2edb6585e31ca79a8744700e5f7e4d0

Observation 0574f479-a91d-4a60-927e-006ba334dda3 · outbound

This paper cites •Subtype: AML with myelomonocytic differentiation.

Information-seeking failures of large language models in agentic clinical reasoning •Subtype: AML with myelomonocytic differentiation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:de2c46dd5c002b4cadd3e669b545d6c2380fc99c8bc223c726dd7745203c9428

Observation cdfa6f18-200a-4dfb-9066-baabccc75077 · outbound

This paper cites incorrect.

Information-seeking failures of large language models in agentic clinical reasoning incorrect

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:4cfd08e20cbeae1878d95168e061e7fd878f07e0f0614ca7858add4ca23a9b4b

Observation f28e267b-75b4-4571-a870-d87d96a98aa0 · outbound

This paper cites Solve” Request: •“Final round - provide definitive diagnosis and treatment plan.

Information-seeking failures of large language models in agentic clinical reasoning Solve” Request: •“Final round - provide definitive diagnosis and treatment plan

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-14T12:59:50.759055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:59:50.759055Z digest=sha256:e966f665f88398dd8f1922f835df83287ed5549763786bf26844f6ab384ac354

Pith citing papers

No inbound Pith citation observations are available.