Pith. sign in

Paper Citation Record · LEDGER

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

As of 11 August 2026, this Paper Citation Record lists 100 of 108 outbound references and 2 inbound Pith citation observations for arXiv:2605.10286.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.10286 v1

Coverage vector

measured 100 of 108 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-12T05:10:02.941396Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T00:25:11.457757Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 108 outbound references displayed

  • verified exact44
  • verified fuzzy32
  • unresolved5
  • parse uncertain1
  • malformed identifier1
  • metadata mismatch17

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8590c3d1-714f-49e5-838f-229059352c8a · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.835066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:0cb04f5c5207f18a37e3510f851e0a29025b449797a472353886a6c4f9b51991

Observation 4eab7724-d9e2-4baf-83f9-0af0fd9caa07 · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.832517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:12a2a1c8ee730fc626b0da68f9faa31f16f2e0579fa2d72cfc8eada4bb57cbf7

Observation ddbf4b11-3073-48db-af19-6805e8da5973 · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.829683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:f04dcdcfc0f2656beffffe24ce5523f5261ebd7276c2474353eccee08bb0bc0a

Observation c7a09e7e-a47b-4673-8684-1e5cd8aa8f21 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 4

Resolution
parse uncertain
raw_fallback, observed 2026-05-12T12:06:33.827187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b9c582d2d996a870c0f8739dc2ce664f32061735d4720f81b6ecc756a0110958

Observation 7d666398-47d7-46fb-8c6f-ff6adf067ccf · outbound

This paper cites and Gupta, Vinayak and Althoff, Tim and Hartvigsen, Thomas , month = dec, year =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Gupta, Vinayak and Althoff, Tim and Hartvigsen, Thomas , month = dec, year =

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.786837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:c35f70c8cfa12f60adb5a0786e103110802f24c3424efd0fcef7e67e7f9f50de

Observation d2c00a8e-21e0-4ebf-9a83-14e8b5e0d1c2 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.760461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:2ced841f5b7a879240d268138b977574c49374f229517244753772d569ae3b22

Observation 65402a0d-98ee-4daa-855c-bfdd48ccf726 · outbound

This paper cites Proceedings of the 41st.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 41st

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.784135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:7ee735252646c0c290e92c1dbe6d0e1c4661822c5bab349809fb58a670998ec1

Observation f76aad0a-1487-4909-a43a-1d2b70b8900e · outbound

This paper cites Proceedings of the 10th Machine Learning for Healthcare Conference , year =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 10th Machine Learning for Healthcare Conference , year =

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.767612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:8e08c0c90ac5172c4d49fb00fbd319f795aa45d0e013d90be36eac05d807be90

Observation dfe0edfd-a566-4238-b994-6038f4a86492 · outbound

This paper cites and Shamout, Farah E.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Shamout, Farah E

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.765194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:c7caf95d9ebf5d2adadf5177062fb00540c0f3e64119d5d8caf7a809fee9035b

Observation 956986e3-92d4-40fa-9177-428da5858d64 · outbound

This paper cites Artsi, V.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Artsi, V

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.766348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b48e6261973b20c701a2b8b0e807ea4aecf4aa3ded1e93fe5353dc0005f71eec

Observation 2f205b71-3d31-42c5-bb6b-166afd34f280 · outbound

This paper cites Proceedings of the 38th.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 38th

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.778888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:7c7ebda3a53dfffda4da156f7139b3e24a2957900897aca4ccdde94a06b8e566

Observation cccb34cb-fa14-4b40-84f8-1aefc7e7f144 · outbound

This paper cites Informatics and Health , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Informatics and Health , author =

Reference 18

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.966505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:ad36738b60d91b9b2134fd162ca9b00a4edd66eb247776d7e42a75de8e4cb7b3

Observation 4ebe1f57-2ff1-4135-a8a4-5ad1fd229787 · outbound

This paper cites In: Intelligent Systems and Pattern Recognition.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks In: Intelligent Systems and Pattern Recognition

Reference 19

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.937555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5df4e8834aea9e5bdd4021e828f8f3d0dcce34a32d992d6f3944ca9534eedcf1

Observation fc7da582-45dd-41c5-9bd7-12f8a38694a1 · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.799859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:6b585f166dfcf6e514556cce2425a6967624f95eb75fee083d8be88d6f7fb537

Observation 54b51020-f2fd-410b-9653-cd87c3bb3286 · outbound

This paper cites arXiv.org , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks arXiv.org , author =

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.791673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:68f42aeee110c7ce1faaae72c3fe003da265f9a81ce2f5332c9ff256b59d3253

Observation 8af205ca-e01a-4bbd-8640-b3743a3499e6 · outbound

This paper cites and Le, Quoc V.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Le, Quoc V

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.775516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:7dd9e63c82166bb6b24f30fb4c110d65fa34bc5444041c628c8aea2d4e01778c

Observation 4a85e997-5c4d-4ab5-9eac-42f34287c9ce · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Retrieval-augmented generation for knowledge-intensive

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.781725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5bd7e255de4bb2e64105cf72728e7975d0cda1597cc90beb89f0f24e9ebdaf8e

Observation fa95bdd7-ead9-4742-80f6-8061dace8f19 · outbound

This paper cites Intelligence-Based Medicine , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Intelligence-Based Medicine , author =

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.803037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:7a6b8fd73c2ff3d012db307e19d4609bac4cc40c984afb1f946c496ac1cbf8eb

Observation 0a0daf09-b80f-4369-a7fd-e668a9407325 · outbound

This paper cites Hospitals , author =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Hospitals , author =

Reference 26

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.796355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:330536f531f9802b1c369231094592cafb74a56f4f5af068f28daa6e85b26bee

Observation 488f1228-f868-4569-a877-7aa79a7f38cc · outbound

This paper cites Proceedings of the 62nd.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 62nd

Reference 33

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.782852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:83d2faed35d046a6d481b3e9e9e50296203f7528aee4e0f1bb3441d50c51b34d

Observation a122d888-a304-4e87-867d-8314501f0cac · outbound

This paper cites and Mordatch, Igor , year =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Mordatch, Igor , year =

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.789182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:cc013a2f26678e619d34b7fb5b75b00d63b683564b2cb8a78101ff2e6029b6f9

Observation 905c4b48-4599-4961-9bd9-330ec76dcd29 · outbound

This paper cites Alistair Johnson, Tom Pollard, Steven Horng, Leo Anthony Celi, and Roger Mark.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Alistair Johnson, Tom Pollard, Steven Horng, Leo Anthony Celi, and Roger Mark

Reference 35

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.778821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:17d81b58af0a05031c768c17d9bcfa5bae88e94862403484200e7bc6c657d7a1

Observation 240c9a5d-a271-4a57-b63d-5f36682677b0 · outbound

This paper cites and Shamout, Farah E.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks and Shamout, Farah E

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.755061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b68a5dee4b73c2b82cdebfce10251e5dabdd51c5643091f1f2f6982207bc62e8

Observation cd872ba8-7a51-4d5f-8f33-506de68de883 · outbound

This paper cites 2023 , pages =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks 2023 , pages =

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.757832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:07e6c09c594f10dbcae91f14f6e9eaa1cc6f040b6391196469410c5edc173c9e

Observation f8d8895d-a904-447c-9dec-45b7e839ee71 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.772815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b3f065219a5b68959da0879c0d3e0c38ddb03ee2a7dd1a267a1d2fb8e61baf34

Observation 40a28629-120d-4401-98aa-a88c6afe9627 · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.960748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b8d22c8171d9cab7e92e786dab7d85aebab8e21e9779c4d2186b1c60c168067c

Observation ca4e5bde-3444-4c15-9ffc-148920d52e5d · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.929441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:bd3636e520157f22602009558b23b5165474e026c76478cd758adbad54000fe2

Observation 52437781-bb0c-4cf3-8340-61032040776b · outbound

This paper cites Qwen2.5-VL Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Qwen2.5-VL Technical Report

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:11:21.811686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5caf8db4cb69602e45218c3898fc3455806a9e1c7cd2cd127647f446c10e2707

Observation 86aaf8f5-e963-4ae2-9d27-b3378cb925d2 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:11:21.792229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:f8bb5ffb7791fbc4884c0562886c4d1312c76a5f7830026a8e55abfff8c35fc7

Observation ec767918-802c-497d-a451-f65b1ab606e5 · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.948174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:385961c0c82276a660bc5d20612d751b0ba965ef4414b7cbd8006e0fe92ae7f3

Observation 56e64423-ac75-48a6-8270-f1c120f20b7b · outbound

This paper cites Language.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Language

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.762843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:e013c90421490f774170e6c2da6d0f3f5d19485a3b2a5ffebbb41a8e164bd6db

Observation c77174ce-82b6-4854-867a-1ee4d4761337 · outbound

This paper cites Self-Refine: Iterative Refinement with Self-Feedback.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Self-Refine: Iterative Refinement with Self-Feedback

Reference 47

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T05:11:21.918839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:34de6d4843e6950660737cfa730e66ea5e52a18ad8ad92e0c5704585277249b9

Observation f195de01-cbd0-49f6-a370-356ad60be745 · outbound

This paper cites Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-20T00:00:25.943863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:798523764aa59539dc5af2eedf7a7052a6aee931cb7c4e8d680e69ab573e8727

Observation fcf00c6d-9c40-4731-a2a0-1773ae5a915a · outbound

This paper cites Proceedings of the 29th.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 29th

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.770268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:7efb254166cb108f8fd041d243f69839100f3482317e8b6659f08578159e0a11

Observation 33a81d47-ec96-4a4e-b5b6-882f7cd473b2 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.794149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:79a9433771b3508a662fa63b2180bd616f70a50370513e2b3069dd0f65989452

Observation 7f28da05-2d7c-4cd3-80df-b001161a214a · outbound

This paper cites ClinicalBERT: Modeling Clinical Notes and Predicting Hospital Readmission.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks ClinicalBERT: Modeling Clinical Notes and Predicting Hospital Readmission

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T09:26:53.041773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:d7df9e990f50aef30aecceefaf328bb815851d0d4a1b9ddb48ad3f9ecf629731

Observation 98ff3b63-8d61-4e03-94ad-e0a2a3ef00fc · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.796800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:84afb7d3cc08ed70a47a45699d20c6d046d35ca19ecb772bf670d9f4a9859ab3

Observation fdef00c9-901d-4245-bd75-1a75f80e2dec · outbound

This paper cites Warren, Lu Cheng, Haidar M.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Warren, Lu Cheng, Haidar M

Reference 56

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.883326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:29981ed4e96d6394c0b66495d865e37c1b6e453b4cd688501e030eb635ddc7a5

Observation 9eb3fa15-e481-4e43-9556-5eed798faa8a · outbound

This paper cites Dreher, T.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Dreher, T

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.773962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b967d8199d8c387260f5ae3c5f4d17731f83547816073d547eaccd012c9145b7

Observation 4b10719a-0204-4f9a-b429-d8178639a960 · outbound

This paper cites Guttag, and Adrian V.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Guttag, and Adrian V

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.895181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:fcf0dd912e76e2d8a0935b0355771b2901868fe7d4a4af23760dec9113d7efd8

Observation 416f1439-4d92-4719-b49c-02220db8ffe6 · outbound

This paper cites Enhancing diagnostic capability with multi-agents conversational large language models , volume =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Enhancing diagnostic capability with multi-agents conversational large language models , volume =

Reference 59

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.821234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:928b0368d59d41ab5dd5144d64bd914170ef2a1572025a6f89e8b50708790526

Observation 70b9a469-ffcd-4db9-b744-503d77b0c2bb · outbound

This paper cites Proceedings of the 15th.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Proceedings of the 15th

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:11:21.868133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:0187c02f2d6ee221ffffe6efcb0af3e73d3e7157aebf59a43e1f166b97541486

Observation c98274ef-ae7a-47d8-90fa-23d7b9ae7513 · outbound

This paper cites npj Digital Medicine , publisher =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks npj Digital Medicine , publisher =

Reference 61

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.902727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:6d575b76a5750cfd83854ee18d70ad9433c274e78442581193fa6fd2495bf9ec

Observation 27c91e59-f0fe-4ec6-a551-cde81b439828 · outbound

This paper cites Mmedagent: Learning to use medical tools with multi-modal agent.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Mmedagent: Learning to use medical tools with multi-modal agent

Reference 62

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.852360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:71caf64fbb2c9e8be9e4e0f878a12afac6478959f9c05c1866e328d5b3c090d1

Observation b14b127e-7069-46b9-9cd7-06700aedcde6 · outbound

This paper cites GPT-4 Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks GPT-4 Technical Report

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T05:11:21.750668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:4591f489bf6d3b8bbc8af4f2db8332914e09c93e2a9de53366eb3db77c8399a9

Observation 01f38518-7563-454c-b0b5-4941b0016f84 · outbound

This paper cites Learning.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Learning

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.808654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:c05d4f5652c24b36e9cebdc0fe25a2ff166780d08be459c887f88c0b344eaf93

Observation 209f307d-e45b-4695-8214-e9cc2fdf4615 · outbound

This paper cites 2025 , journal=.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks 2025 , journal=

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.802570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:a702c0f41f9440a616458d7ee633cd07ce7ddddaa53bff6a52fd097a703ab159

Observation 71ae8a11-63d2-4edb-b22f-63b0737f842a · outbound

This paper cites MedGemma Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MedGemma Technical Report

Reference 68

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T05:11:21.714853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:96940404d0335214e05650c5bcc5e64b78c41f674d4c5091ef9e7bdb2c7f6fce

Observation 219aebf7-7581-484f-b603-49a233befad6 · outbound

This paper cites Vision Language Models in Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Vision Language Models in Medicine

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.670645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:ec8023582dbc957865a5b39540f34d4dc0cc9e9a804febb3498c4520f0954e2d

Observation bd347140-57e0-412e-bc3c-a2c4120d35ad · outbound

This paper cites Multimodal.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Multimodal

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.858099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:1c5b02cc7b3ca9335550ae369518defd2c882ff9659449ef41b04e518c316d8a

Observation 9abd689d-4421-495c-9b34-1ffda86b1ecc · outbound

This paper cites Guyon and A.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Guyon and A

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.811845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:fb9dbb0688d7ae11630128a417becf8e126f7bfd76293b64382711eee0baa6b9

Observation 1ad8a2c5-ea43-42fa-8a0f-0aa296708d24 · outbound

This paper cites Guyon and C.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Guyon and C

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.805657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:ca3ca7ea5a6886afeabd244b8ec6871f8722e2206e7076e4e9a170debca9e44f

Observation 505d886b-643f-4855-89c6-9cca37dc03f1 · outbound

This paper cites MIMIC-III, a freely accessible critical care database.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MIMIC-III, a freely accessible critical care database

Reference 75

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.519615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5d904d175fbd0c5f6ca306f7123c6544cb4fe0d58f6c9943d0df8ac94755e8fb

Observation 68de056f-2ebf-4b76-9ed0-ec803bdc543e · outbound

This paper cites MIMIC-III clinical database.PhysioNet, September 2016.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MIMIC-III clinical database.PhysioNet, September 2016

Reference 76

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.564005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:24221c205d65cf2a1d9b84d6629caad5db457d911dfa172161b6f486601fd0bd

Observation ced4230b-1fa7-4f48-bcd7-a6e3bc42aecc · outbound

This paper cites Clinical risk prediction using language models: benefits and considerations.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Clinical risk prediction using language models: benefits and considerations

Reference 77

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.593576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:e0f267e14d8e735d0b1f16a01d4e22291f34988a0b03a8c2be1b880d22028f08

Observation 67e38d2e-4a99-437e-8656-dfd67e9ce55e · outbound

This paper cites A pragmatic randomized controlled trial of ambient artificial intelligence to improve health practitioner well-being.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks A pragmatic randomized controlled trial of ambient artificial intelligence to improve health practitioner well-being

Reference 78

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.649627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:4179fab0fb9623a223713136992a6d7b9e1946655df720cf8ba62bc147e3bfcb

Observation 6d4620d1-11c9-432e-9a7e-46450ad3879f · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.695874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:cc62ce67ba85f401a9e7017baaaaee8eabe15b11748a4a73354f948a1c65511d

Observation e23ac393-0a15-4091-91bb-094016ca50a3 · outbound

This paper cites EHRXQA : A Multi - Modal Question Answering Dataset for Electronic Health Records with Chest X -ray Images.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks EHRXQA : A Multi - Modal Question Answering Dataset for Electronic Health Records with Chest X -ray Images

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.818365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:2c883199cc749661a635b85e8753396485bc6baab59b94900cf66083dc096c57

Observation 909dd84a-8f62-49c1-b315-c0228c72bd94 · outbound

This paper cites Qwen2.5-VL Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Qwen2.5-VL Technical Report

Reference 81

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.830002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:381c6afe8f451e83a880e10111af9091753c439cfe4a81e92002224493999d51

Observation 68fac71c-0066-4905-949c-40c69955e1a3 · outbound

This paper cites Bicknell, Danner Butler, Sydney Whalen, James Ricks, Cory J.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Bicknell, Danner Butler, Sydney Whalen, James Ricks, Cory J

Reference 82

Resolution
malformed identifier
doi_truncated, observed 2026-05-12T05:11:21.550928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:337dec5176f484ac6d71e6eb155807352b616d3ab0f9cb939bac7a880a7dd358

Observation 48d31f79-ad17-49fd-8dd3-458f78165c3c · outbound

This paper cites Language Models are Few-Shot Learners.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Language Models are Few-Shot Learners

Reference 83

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.817942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5c004c1dc39e004ecdd4ea27eef48f0ce2892ba2b1b0d806c0f3130972febbd4

Observation 18b31044-ce80-4cf9-9db5-8a359be30395 · outbound

This paper cites Knowledge and perception of primary care healthcare professionals on the use of artificial intelligence as a healthcare tool.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Knowledge and perception of primary care healthcare professionals on the use of artificial intelligence as a healthcare tool

Reference 84

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.568768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:9912805bc0dbba6eafe3810cab9ab9e9406a0825e0d5b70f83ae2cff16ba6929

Observation d0d71ce1-715c-4037-815f-5e25acfa5a8f · outbound

This paper cites Why Do Multi-Agent LLM Systems Fail?.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Why Do Multi-Agent LLM Systems Fail?

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:42:59.242121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5439717914839f0861f9fa080add12d30a673b37b72921f6b9b8a49ce2ed63ac

Observation 2b5860c5-77d2-4320-a8c8-80f39e1c6c10 · outbound

This paper cites Multimodal Clinical Benchmark for Emergency Care ( MC - BEC ): A Comprehensive Benchmark for Evaluating Foundation Models in Emergency Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Multimodal Clinical Benchmark for Emergency Care ( MC - BEC ): A Comprehensive Benchmark for Evaluating Foundation Models in Emergency Medicine

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.824268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5f3a634e673bfc2624286bbb6006230e49e0c6ca3ea51bacd5589967b7d58485

Observation 60bd4685-df30-430c-950e-d8bd3138c0a9 · outbound

This paper cites Multi-modal learning for inpatient length of stay prediction.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Multi-modal learning for inpatient length of stay prediction

Reference 87

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.577644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:473f1518b052b01a256fd010ffe00d8ee37e1b19234e42a5be11cf2a3a2ddb83

Observation afc2741f-a9a7-4cdb-a1c9-2a62f8bc5c6a · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 88

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:31:25.823602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:9ea228bce2571f27cafba1c813f521c7ee44a661972aacdb8b1e9dd2f52bce57

Observation d5247dc3-f81b-4fa4-acac-2b9085bb4953 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 89

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.864501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:94d9f7bb6057f537ce884d16b0c9b0f489a4589b16804ef8f1d54ba43e4b7d6a

Observation c974c379-564d-4b0c-abe6-e3767e0588b8 · outbound

This paper cites Tenenbaum, and Igor Mordatch.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Tenenbaum, and Igor Mordatch

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.820841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:70a002c786d78c8c604782f6adc8003a87916b4a15c973c5ee79ac727be03afd

Observation d16899d3-b93c-4195-8e63-4aeef8255d9e · outbound

This paper cites Geras, and Farah E.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Geras, and Farah E

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.837974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:6bdefb95ac490be5310e5ea8d6298a0d79ddccb230ee53956c542116721f2a51

Observation e536169e-f640-4b34-afcf-23c003c59f29 · outbound

This paper cites Single-agent or Multi-agent Systems? Why Not Both?.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Single-agent or Multi-agent Systems? Why Not Both?

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:31:25.871484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:e7e26853a3fa1de5b7aea2fcf2736f2223507f61e6358202ffd5cac2b26a7515

Observation d3257e39-8c25-4988-9763-48db517776f3 · outbound

This paper cites Geras, and Farah E.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Geras, and Farah E

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.860483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:f423c5eaf0e198a1743cb64a6b5e49e86231a31aa178a51396c10b8f6a3d2fbe

Observation 880983be-084e-4578-8b73-e43f98306ae7 · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 94

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.809518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b7ed8353f8a1f121773921505bb6c16b364ee26fb2d342b589b1ebd2d4356113

Observation 17a8b695-b08d-4197-916b-0b217f956bfb · outbound

This paper cites MetaPrompting : Learning to Learn Better Prompts.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MetaPrompting : Learning to Learn Better Prompts

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.855144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:cf947664dcee32dbca1ea8561e4f61d6ce06eb1686fb52443e499643d63f8906

Observation 94f839cd-c5f8-4e39-a907-5a4ef5585db7 · outbound

This paper cites ClinicalBERT : Modeling Clinical Notes and Predicting Hospital Readmission , April 2019.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks ClinicalBERT : Modeling Clinical Notes and Predicting Hospital Readmission , April 2019

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.849589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:1d6fc8ef2624cf54f417b80080e14fb91ad7a258f8bc7702112f17eea1454b1f

Observation fc1b640b-e836-4188-8fb8-bb2495924a5d · outbound

This paper cites John Wilbur, Zhe He, R.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks John Wilbur, Zhe He, R

Reference 97

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.644022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:6c71217143466c6cd21fb040fe6dd12dd9c6117640bf0581a6ead2b2f113772a

Observation 78e19f98-e0d7-4cdf-8017-0fe91262e548 · outbound

This paper cites MIMIC - IV - Note : Deidentified free-text clinical notes, 2023 a.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MIMIC - IV - Note : Deidentified free-text clinical notes, 2023 a

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.852239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:782e87c79a2d8ca75c98584fc462d1066ac78574ac4ad5554d1bc88c4ab9da1b

Observation 7edca90c-5783-400f-9794-93b5a5035b54 · outbound

This paper cites URLhttps://www.nature.com/articles/s41597-019-0322-0.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks URLhttps://www.nature.com/articles/s41597-019-0322-0

Reference 99

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.585214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:6b2a033b60eef84c98a42b4a25e78580bd6a3b36e1401d01f06ab0b269474dd3

Observation a32491b8-8d1d-4e05-9295-0787a40eaf90 · outbound

This paper cites cc/paper_files/paper/2019/file/ ac52c626afc10d4075708ac4c778ddfc-Paper.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks cc/paper_files/paper/2019/file/ ac52c626afc10d4075708ac4c778ddfc-Paper

Reference 100

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.728236Z

Source-reported events for the cited work

correction dated 2023-01-16. Source: crossref record 10.1038/s41597-023-01945-2->10.1038/s41597-022-01899-x:correction, observed 2026-07-11T02:59:03.685257+00:00. This notice travels one citation hop only.

correction dated 2023-04-18. Source: crossref record 10.1038/s41597-023-02136-9->10.1038/s41597-022-01899-x:correction, observed 2026-07-11T02:59:35.186743+00:00. This notice travels one citation hop only.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:6de01e71e4d15bacee2950ff6a32571da8d62edfa079024a8a82677037c3c31a

Observation e41a7711-bb17-45cd-a70b-fd36762c3e95 · outbound

This paper cites an unresolved cited work.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Unresolved cited work

Reference 101

Resolution
unresolved
raw_fallback, observed 2026-05-12T12:06:33.814562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5f702df3dcb73fd63c68f7d627cf239b8570021b48f886a367aa8d22ca4416aa

Observation 8bb30ff7-d8eb-45f4-aa55-0aae2dfc2200 · outbound

This paper cites Voting or Consensus ? Decision - Making in Multi - Agent Debate.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Voting or Consensus ? Decision - Making in Multi - Agent Debate

Reference 102

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.709628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:8301458a830c57b7a9135c9803fd8d4cad9d0964693a23b8ad12adb3abd87e13

Observation 4c375437-dc30-42cc-8a32-a15a2edceb74 · outbound

This paper cites Vision Language Models in Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Vision Language Models in Medicine

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:31:25.805519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:7e8b621067c64f66e42cb75be26828474d1438b3582f0feccd559ce9c56ce14c

Observation eecf895d-669e-4506-8146-44d8c4b6b708 · outbound

This paper cites Clinical Risk Computation by Large Language Models Using Validated Risk Scores.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Clinical Risk Computation by Large Language Models Using Validated Risk Scores

Reference 104

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.660447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:86fb944381352feebefd67bf8008d9217fea0a9d571cd20d9461e13129b6b67f

Observation 623ab034-d58b-450b-8376-039eb80f1745 · outbound

This paper cites Medical transformer for multimodal survival prediction in intensive care: integration of imaging and non-imaging data.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Medical transformer for multimodal survival prediction in intensive care: integration of imaging and non-imaging data

Reference 105

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.664388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:8a01b4c144ad675ea7f1ccfaecb2c2fd5fc9eed0ee331c3900a9c6271c26d5ad

Observation 1b77f872-e55c-461f-9748-56824b728b98 · outbound

This paper cites MDAgents : an adaptive collaboration of LLMs for medical decision-making.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MDAgents : an adaptive collaboration of LLMs for medical decision-making

Reference 106

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.863540Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5c530fff55313e395ef4185c337d630ce91e3adb20427053967955b69e2f0730

Observation 8cd86349-524f-4bde-b1d0-62c2630dd209 · outbound

This paper cites Towards a Science of Scaling Agent Systems.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Towards a Science of Scaling Agent Systems

Reference 107

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.833946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:5138cf7f1a9d7093e047a3d360235809cb9697b47dfef255eab2280a25808db8

Observation ca0b8734-8b79-49af-8ea6-8bffd0194fde · outbound

This paper cites Bioinformatics , volume =.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Bioinformatics , volume =

Reference 108

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.604221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:376495eb1d09d8c3ad64909c43d4b3a7cfc2148d3d8d9798ecc76cabffd50ead

Observation 5b0f7f8c-ea5e-4f72-b302-019b290cd1bb · outbound

This paper cites Learning Missing Modal Electronic Health Records with Unified Multi -modal Data Embedding and Modality - Aware Attention.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Learning Missing Modal Electronic Health Records with Unified Multi -modal Data Embedding and Modality - Aware Attention

Reference 109

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T12:06:33.866329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:479ca4574278ecc494188f6177242d0cba3533d50c9b5983ecfedb5df03767ca

Observation d81d31eb-3a70-4359-acd7-6a34611eee8f · outbound

This paper cites A prompt framework for enhancing LLM -based explainability of medical machine learning models: an intensive care unit application.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks A prompt framework for enhancing LLM -based explainability of medical machine learning models: an intensive care unit application

Reference 110

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.634176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:0d0f4ee9fc3561b405164421656033abb04c950ab4ce69b611460946e8813d2e

Observation 9c33940e-220e-40d1-9c02-3c07270ed32c · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive NLP tasks.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Retrieval-augmented generation for knowledge-intensive NLP tasks

Reference 111

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.690309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:be8dfe34a25d42c7c438fee893e01fc60f826c472cd5b320325f09df5b4392a1

Observation 5f967fbc-ed65-43e8-b760-3e56eee4a028 · outbound

This paper cites LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks LLaVA-Med: Training a Large Language-and-Vision Assistant for Biomedicine in One Day

Reference 112

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:31:25.859967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:9e27b01f3a56baaa7e116bd43a93f9a5885a470294e3dc5b2a1ab97f28189ffb

Observation 2a97e58f-bcb2-44cb-a314-fcc5130719e6 · outbound

This paper cites In: Al-Onaizan, Y., Bansal, M., Chen, Y.-N.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks In: Al-Onaizan, Y., Bansal, M., Chen, Y.-N

Reference 113

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.616113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:457145f8a457e32d92d57d2850cad349094d66c949d57a372788acda48546cc1

Observation 6f9506d9-8c65-4988-b182-b5d58158b3fc · outbound

This paper cites Ovis2.5 Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Ovis2.5 Technical Report

Reference 114

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:30:17.288454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:74659a4f78894f3744591c61a3ab777c3438940d4ca683031812cf007cf8a620

Observation 3d247753-e3cc-4085-a355-485dbc5deb45 · outbound

This paper cites Lukac, William Turner, Sitaram Vangala, Aaron T.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Lukac, William Turner, Sitaram Vangala, Aaron T

Reference 115

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.705146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:fda15ff722cf57df1a509cecf5017c12b74aa58b29fe4519da4f7549cd7eff5e

Observation 957eaeca-ed5c-487a-b836-f0ae078f6602 · outbound

This paper cites Self-Refine: Iterative Refinement with Self-Feedback.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Self-Refine: Iterative Refinement with Self-Feedback

Reference 116

Resolution
verified exact
local_arxiv, observed 2026-05-12T05:31:25.800945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:dc754715ac2b69741251ad43df9908e9aeda752c8e28fc78c338d0e683e49d51

Observation ed19ff2b-9d15-4ee3-8e36-240d3f797754 · outbound

This paper cites Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine

Reference 117

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:31:25.879478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:ae2c50da53acc2b75a2365f3da3c41ff7effbc2fbd9a3d3540b63c1ea7ce1cde

Observation 86e46fa3-88c8-4a47-aaa2-b62cb7255e53 · outbound

This paper cites MedGemma Technical Report.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks MedGemma Technical Report

Reference 118

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T05:31:25.875394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:dc13dbb386c34d7934fd684bcd3981e27d5854c7ac6267901da734d87ff55877

Observation 817be877-f2e7-499a-b59d-a486c346e5d5 · outbound

This paper cites Transforming Healthcare with AI : Promises , Pitfalls , and Pathways Forward.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Transforming Healthcare with AI : Promises , Pitfalls , and Pathways Forward

Reference 119

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.555303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:0786bb86521abb19ef45571f5db4613b8cd4fa6b0d3dd4207a40785d60b30369

Observation 5198008f-61c7-48f3-95e0-fc1f6843e3d0 · outbound

This paper cites Large Language Models Encode Clinical Knowledge.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Large Language Models Encode Clinical Knowledge

Reference 120

Resolution
verified exact
doi, observed 2026-05-12T05:11:21.738778Z

Source-reported events for the cited work

correction dated 2023-07-27. Source: crossref record 10.1038/s41586-023-06455-0->10.1038/s41586-023-06291-2:correction, observed 2026-07-11T03:08:19.417011+00:00. This notice travels one citation hop only.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:b538df62e712d23b43abcbfb2267ea9a3e76af248319731919c949a5f82169ab

Observation 9631a6e4-7c32-4043-aee4-161770b7c85f · outbound

This paper cites Singhal, T.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Singhal, T

Reference 121

Resolution
metadata mismatch
doi, observed 2026-05-12T05:11:21.677554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:2a0bc830eaafe0aaec05706c029e95ba2de6029ffaf615e697854b3710d6fb6c

Observation 2baaec85-dd7c-41c4-8242-9cfcad0aae43 · outbound

This paper cites Baxter, Florin Vaida, Amanda Walker, Amy M.

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks Baxter, Florin Vaida, Amanda Walker, Amy M

Reference 122

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:21.640154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-12T05:10:02.941396Z digest=sha256:645e08761311b72ddef5f771eb415c8701cc933c8f3886dd6dab4f06944d0ec8

Pith citing papers

Observation 8453a3f8-b7b9-43d6-8e4d-98d4a510b9fd · inbound

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy cites this paper.

The Path to Self-Evolving Clinical Systems: Scaling Medical Agents from Assistance to Autonomy AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

Reference 111

Resolution
unresolved
no resolver link, observed 2026-07-14T06:30:16.612345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T06:30:16.612345Z digest=sha256:062eb57eef3b77c86dc8fa5661d2d3487704c263f2c8161ef394e483d3ef77c5

Observation 135a7814-95d7-487d-8dd8-857f1181faa7 · inbound

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures cites this paper.

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-03T00:25:11.457757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T00:25:11.457757Z digest=sha256:6405f5e0de49f82ef2fa4220a4d867be6de0ba1681eae01ee2f6f3f0ee95f962