Pith. sign in

Paper Citation Record · LEDGER

AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 12 inbound Pith citation observations for arXiv:2402.09742.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.09742 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 12 of 12 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:08:27.981124Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T09:55:40.453086Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 98b970bf-7059-4d7c-a7fb-c9d08f7c29e9 · inbound

Large Language Model Agent: A Survey on Methodology, Applications and Challenges cites this paper.

Large Language Model Agent: A Survey on Methodology, Applications and Challenges AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 137

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:52:10.745335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:51:34.309870Z digest=sha256:b14e81585a4cadccb6359f5f2432fe510c7a21b768004a6c004e7e4fb75b04e0

Observation 89d6027a-0986-455c-9000-fd762434b75c · inbound

A Survey of Scaling in Large Language Model Reasoning cites this paper.

A Survey of Scaling in Large Language Model Reasoning AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-22T21:22:09.176472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T21:20:07.238992Z digest=sha256:0446f2d5aa4b1881399ee8144770d4b2cad6f2a1df1fae44c25bfddc121a1e06

Observation 9e7061a3-daf2-4256-af21-4dd6aa659358 · inbound

R2MED: A Benchmark for Reasoning-Driven Medical Retrieval cites this paper.

R2MED: A Benchmark for Reasoning-Driven Medical Retrieval AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T13:54:53.084126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T13:52:15.835379Z digest=sha256:0ac03efcf8815603250068e158a7ec2ae0ab777e77433bad8616d016e6c2d3c6

Observation 3b5fe226-53a8-40ed-ab21-ac3cf60e143e · inbound

Real-World Doctor Agent with Proactive Consultation through Multi-Agent Reinforcement Learning cites this paper.

Real-World Doctor Agent with Proactive Consultation through Multi-Agent Reinforcement Learning AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T13:42:19.240874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T13:41:30.833840Z digest=sha256:2012c362f7188c5fae9ae2ada83e1d9e64b6c728cc6377b6e6b2c8feaa3d6119

Observation 806f0a48-f215-48fd-a548-b5c1636cbcd7 · inbound

AI4Research: A Survey of Artificial Intelligence for Scientific Research cites this paper.

AI4Research: A Survey of Artificial Intelligence for Scientific Research AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 195

Resolution
unresolved
no resolver link, observed 2026-08-06T20:45:12.571334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:45:12.571334Z digest=sha256:0750be7e6c6bc8c3d4cd134b51199fededf6385610ccc62dd76e9fff7eeea920

Observation e36b5d81-1974-42c3-9fea-ee005346e2db · inbound

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit cites this paper.

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T04:57:04.476242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T04:54:38.908451Z digest=sha256:da9f4e1b331b02fab7f018f1221da171a53fa21b1dfbd3283437801f820780ea

Observation 5e1d7268-7616-4105-b4d3-5b90d5d8fc1e · inbound

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit cites this paper.

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T17:51:07.983163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:51:07.983163Z digest=sha256:41165694d6c6674f758cef0a27d745b91369459e9dbff6687bb4e206f22ae69a

Observation 63a7c2f2-db37-41ad-8795-6bac1a1d14b1 · inbound

FinTeam: A Multi-Agent Collaborative Intelligence System for Comprehensive Financial Scenarios cites this paper.

FinTeam: A Multi-Agent Collaborative Intelligence System for Comprehensive Financial Scenarios AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:02:14.126538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:02:14.126538Z digest=sha256:3a314ffb4f991eaf3dfa7dcaaeb848680f0ca680e3ecc0b0c1b484d56f43a13e

Observation 24522d0a-1861-4ffb-9a96-a8e34b8b0448 · inbound

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting cites this paper.

Toward the Autonomous AI Doctor: Quantitative Benchmarking of an Autonomous Agentic AI Versus Board-Certified Clinicians in a Real World Setting AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T22:08:27.981124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:08:27.981124Z digest=sha256:2c4ee86b96f2c8688ef97c3c4a8d704a64f10769515e119c760f75c849195160

Observation a66b77e7-aa7a-4a2f-93cd-4fd24b3ed3fa · inbound

Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation cites this paper.

Inflated Excellence or True Performance? Rethinking Medical Diagnostic Benchmarks with Dynamic Evaluation AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T08:12:29.817129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T08:12:02.449352Z digest=sha256:90765c26492b6287739df3cb39002fbe75d522a734f83716962d8137e2d7111e

Observation 638c0e92-3f86-4a48-abac-54aedd7d2639 · inbound

RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical Diagnosis cites this paper.

RE-MCDF: Closed-Loop Multi-Expert LLM Reasoning for Knowledge-Grounded Clinical Diagnosis AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T08:42:37.034816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T08:40:45.666978Z digest=sha256:919404d08e5e500a402934068aca9efb79a459a4f9ee753596c34893e6ed27d0

Observation f74c3cef-a37e-47c7-85ac-30d12d2fa661 · inbound

HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents cites this paper.

HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents AI Hospital: Benchmarking Large Language Models in a Multi-agent Medical Interaction Simulator

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T09:55:40.454667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T06:07:06.879550Z digest=sha256:bbe7ca638e578a2b3bf40f6abb8902fc3ce2351a4805ecca638a4daa03600a9a