Pith. sign in

Paper Citation Record · LEDGER

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting

As of 16 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 2 inbound Pith citation observations for arXiv:2507.22902.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22902 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:08:30.490781Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-12T02:21:16.825681Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T07:41:42.559848Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation f1653e6f-6e4e-43fc-a7bc-d8dfdb901bff · outbound

This paper cites Lower Back Pain.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Lower Back Pain

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:34.156489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:28.761844Z digest=sha256:2086836c2777d55b746c0f9c615766f10085cee6ecfe5ca24c0c8ed73478f72a

Observation 19090acd-74fb-4ed8-92ef-ec0e82e5544f · outbound

This paper cites Washington, DC: Association of American Medical Colleges, 2021.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Washington, DC: Association of American Medical Colleges, 2021

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:35.590065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:27.200231Z digest=sha256:e07bf2788fbc9353071cc046e1fa73ceb32f5b2aae8ce09c6cb709a78840bae8

Observation ab73543e-c008-4c64-83f1-d66d139d7836 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:35.434081Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:27.268965Z digest=sha256:09e9324fe33931df5e22c7b8e2a878d54d62cd9f04f685b1d744e0f1537f5057

Observation 7c53d8af-22c7-45c9-9552-887fdef9a442 · outbound

This paper cites Arch Intern Med 18: 1377-1385, 2012.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Arch Intern Med 18: 1377-1385, 2012

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:35.319229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:27.384233Z digest=sha256:b6bc56147da10f3c285379ee31f75ae1cb08e601ae8fc06a238a85e7acb52bf0

Observation 02de3724-d4cf-492e-9ad4-8d2d2ce57be2 · outbound

This paper cites New York: Basic Books, 2019.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting New York: Basic Books, 2019

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:35.050725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:27.533897Z digest=sha256:abdbc36ccd6bbf83ee2d1b95fb7ca140d5e6b3bb58d5e8d0ec6aa2b9a37a9348

Observation edce1057-7f50-4b21-8f7e-cd67909ceb1f · outbound

This paper cites NPJ Digit Med 4: 93, 2021.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting NPJ Digit Med 4: 93, 2021

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:34.833681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:27.663240Z digest=sha256:4d125a6957da89d1f1fe7edbf60cc2d7de366734c6b11daf6e24d93c28975346

Observation 59068972-6822-4a6c-8934-d1dc6201d54e · outbound

This paper cites Ambient AI Scribing Support: Comparing the Performance of Specialized AI Agentic Architecture to Leading Foundational Models.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Ambient AI Scribing Support: Comparing the Performance of Specialized AI Agentic Architecture to Leading Foundational Models

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T22:08:30.804816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:27.815801Z digest=sha256:227b6ab3eecd86839d3b8e25c914d1fbebe9480355a2add2dace05b39f9b8021

Observation 24522d0a-1861-4ffb-9a96-a8e34b8b0448 · outbound

This paper cites AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:27.981124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:27.981124Z digest=sha256:d0c0e1be28c85e0abcc1ca59bcfc970f48748198a93b71a0e93d64332de860c8

Observation 01a3414b-a0c5-43dc-b4f3-82baca1b047b · outbound

This paper cites Improving Clinical Note Generation from Complex Doctor-Patient Conversation.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Improving Clinical Note Generation from Complex Doctor-Patient Conversation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:28.055295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:28.055295Z digest=sha256:3407cf45e9886696b46edca92bef11cc04dd6b143f3a5807f232e9f78f854b57

Observation b1509b49-a100-40c3-8b12-eb6e16c65fcc · outbound

This paper cites Towards accurate differential diagnosis with large language models.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Towards accurate differential diagnosis with large language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:28.172326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:28.172326Z digest=sha256:b135a218f17477c86f0d4c00afe8f4a12d139b7c842a343ea321b7679d2e061d

Observation 7c4ee0ab-67be-499a-b1ad-bfff77976ffe · outbound

This paper cites Ann Intern Med 178:498-506, 2025.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Ann Intern Med 178:498-506, 2025

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:34.615507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:28.280044Z digest=sha256:7323d50a8e59d75d6d141e3eb93896c62bf469050cc1379f96e1f6a99ddc6cfb

Observation 4b8687d1-1c88-470a-9ae5-eb25ded516ad · outbound

This paper cites LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting LLMs-as-Judges: A Comprehensive Survey on LLM-based Evaluation Methods

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:28.413051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:28.413051Z digest=sha256:81196edd265b94f4ab72b576d7dca0c7d822eb7f9e273e0d6759e1dcb35efb3f

Observation d7b71641-b679-4d82-b8e9-deb1809ba787 · outbound

This paper cites Assessing the Quality of AI-Generated Clinical Notes: A Validated Evaluation of a Large Language Model Scribe.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Assessing the Quality of AI-Generated Clinical Notes: A Validated Evaluation of a Large Language Model Scribe

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:28.562592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:28.562592Z digest=sha256:cf2c192c89eec8ab6663f6c8f72af93f7d0dc5a17f4895966de2cabe7deea091

Observation a860dda0-c6f7-4706-adbc-2c7c7d28de7b · outbound

This paper cites clinically consistent.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting clinically consistent

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:34.380831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:28.686148Z digest=sha256:76e15529c4981e8267b50278a523b6a3f4136165493958c0d77c2fc049da3884

Observation 5951218f-c975-4f04-9405-8579d66a52fe · outbound

This paper cites Eczema” in one SOAP note would be clinically consistent with a diagnosis of “Atopic Dermatitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Eczema” in one SOAP note would be clinically consistent with a diagnosis of “Atopic Dermatitis

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.925016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:28.887270Z digest=sha256:3fd15fffc4df37371665a542c1cf69d34fcfa95c766461b78e3720de4dbe68c7

Observation d0a10454-1b36-4b30-a557-c104d2d6c523 · outbound

This paper cites Sinusitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Sinusitis

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.750763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:29.050632Z digest=sha256:cf079c2c17bcc8214c3026d7cb645efe84c080881c382ca1dc3bc8e93062aa7f

Observation 1cd01b5e-d58a-4ad5-8c12-07a9c074d8d2 · outbound

This paper cites gallstones.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting gallstones

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.510854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:29.181371Z digest=sha256:29dadd32c753a8f27b50deed60da7e0ab13a37715c973a9438fa77e3bd13caa2

Observation 75cb798b-31de-493f-888d-8a467ce7fc29 · outbound

This paper cites clinically consistent.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting clinically consistent

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.320166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:29.309075Z digest=sha256:0c844c50111b7c00fa87a6bfbbd9340d572aab71a5099df377608799993cf269

Observation 84fdff10-6ae7-499a-8741-f70684d124a0 · outbound

This paper cites Lower Back Pain.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Lower Back Pain

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:33.119545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:29.424324Z digest=sha256:8eb7035515fc29d4c7ec29566d84635f30341e779d1fce719c4122ecfd3e67d5

Observation f14d327c-6303-4522-90ac-6411cbe2fc08 · outbound

This paper cites Eczema” in one SOAP note would be clinically consistent with a diagnosis of “Atopic Dermatitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Eczema” in one SOAP note would be clinically consistent with a diagnosis of “Atopic Dermatitis

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:32.866796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:29.560209Z digest=sha256:077367b28f05fc307ad47d206efb0afac31bbef51e7b72d1fd765adb0ac238ec

Observation a25741eb-a066-4976-b93e-a0a36d374796 · outbound

This paper cites Sinusitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Sinusitis

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:32.644204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:29.668139Z digest=sha256:77dab6e5a6483792f802765db63238193e063022f40aa1366452aeed3a1197c1

Observation 46f96c63-ad5c-42d5-bfbd-d011cb1daf25 · outbound

This paper cites gallstones.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting gallstones

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:32.410739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:29.791790Z digest=sha256:61a4a5357cf59195146bea6a5ebac7e29410de9947407d449e4e691639870bbe

Observation 5ff83ca4-6cea-4236-b5c6-4d109b2853c8 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:32.192580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:29.948546Z digest=sha256:f361ad28069c2c5c5b9734048b2aee27453e163c9164d76e7d3d8789b4719231

Observation c597a882-8a9f-44b2-9b21-8291b9485525 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:31.878329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:30.075323Z digest=sha256:a13481bf644b770df69c6ec7fa22b9a41a0fede60e1d699637b0812e57cfa2ca

Observation 01a8ec30-1e17-4716-8096-1f4c1e676af8 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:31.634192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:30.184215Z digest=sha256:a6b0443ad62e82649b16e6f42eb23f8a5c6c6f1a17ee48b4f977d7c77342729d

Observation f7554707-485d-4c9e-80a0-9cbc5eb39e39 · outbound

This paper cites Ibuprofen 600 mg three times a day.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Ibuprofen 600 mg three times a day

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:31.394529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:30.242958Z digest=sha256:09feab4017cfafed4b140223265fcd20a278d54837144ad7b89609e11e8bbed9

Observation 8643c278-7af9-4a5c-b77e-e97d3afd7ce3 · outbound

This paper cites an unresolved cited work.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-06T22:08:31.222079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:30.375897Z digest=sha256:6fc122d19d1fbf04806f2fae3056ce26f0a56b2d74a3986bcd93c143ffbc20bd

Observation 12e8e079-c39b-441c-80f9-e4d537dba161 · outbound

This paper cites Acute Sinusitis.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting Acute Sinusitis

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:08:31.054751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:08:30.490781Z digest=sha256:b630a5d461bb67ba5b0a5559708979a138f2bd31c90b8c80e57ba29dafad7cdc

Pith citing papers

Observation 21cb1cac-046c-4efc-8a34-46571cda31a5 · inbound

SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment cites this paper.

SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:46:53.518176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-07T16:17:39.923337Z digest=sha256:a31ca71ccb00d12cbdd013ca3d2d4ad7aca859de0f54644eea032922778513b2

Observation 49492762-1d12-45f2-b3b5-26db7a6acfc9 · inbound

SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment cites this paper.

SymptomAI: Toward a Conversational AI Agent for Everyday Symptom Assessment Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:41:42.609389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-12T02:21:16.825681Z digest=sha256:f87c71f040a73d292f8a7164d1bd93a50679d75eaf0c320385e5f7736a183b10