Pith. sign in

Paper Citation Record · LEDGER

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning

As of 10 August 2026, this Paper Citation Record lists 81 of 81 outbound references and 3 inbound Pith citation observations for arXiv:2505.24105.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24105 v1

Coverage vector

measured 81 of 81 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:40:16.167357Z

measured 84 of 84 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:42:34.643403Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

81 of 81 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved56
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 88cd335b-b05b-4e69-9ea7-8938a1be17fd · outbound

This paper cites GPT-4 Technical Report.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:05.748193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:05.748193Z digest=sha256:f6b9abdce6d07b82c30e53e210492ad4028a8527b0c27d263ded97fbec357779

Observation d031cc18-80a3-4363-b2aa-beb9f1a67af2 · outbound

This paper cites Claude 3 haiku: our fastest model yet, March 2024.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Claude 3 haiku: our fastest model yet, March 2024

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:05.836155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:05.836155Z digest=sha256:f02035e1d637f2a2c5f5ba9218aa485ad4c28a2c4a1918d7098ab857227384e8

Observation ccb64938-c2e3-4146-adf7-1470961a05ba · outbound

This paper cites Claude 3.5 sonnet, June 2024.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Claude 3.5 sonnet, June 2024

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:24.573185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:05.989525Z digest=sha256:e72c42da22b7e4db04197c9e6a429e1e18b2f59feb76657d84b4163de59d15d6

Observation 25a76a8d-bac8-4dba-8c89-500edc05a857 · outbound

This paper cites Training a helpful and harmless assistant with reinforcement learning from human feedback, 2022.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Training a helpful and harmless assistant with reinforcement learning from human feedback, 2022

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:06.100300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:06.100300Z digest=sha256:01d558395e1eea456dd954c8cc8b3b53497fd8719e1b2b5262dd15ccd47b739c

Observation ccf6234e-e6e1-4535-84c1-6581ca90850e · outbound

This paper cites Rm-r1: Reward modeling as reasoning.arXiv preprint arXiv:2505.02387, 2025.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Rm-r1: Reward modeling as reasoning.arXiv preprint arXiv:2505.02387, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:06.195545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:06.195545Z digest=sha256:d24fc87f44992d5cbf12e22e32fe67e37c5cea9d52bc638f0c6f8630702303f0

Observation 650e90de-3aa1-40ab-93d0-bd5cfd21236a · outbound

This paper cites TIMER: Temporal Instruction Modeling and Evaluation for Longitudinal Clinical Records.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning TIMER: Temporal Instruction Modeling and Evaluation for Longitudinal Clinical Records

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:06.323564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:06.323564Z digest=sha256:dda70cec432f6a2212616f9753dd1bc984a50fadf42ad96a877cfc9a37a397fa

Observation 6e6a513b-0f05-4bc0-8d0e-4c93033ec4d1 · outbound

This paper cites A new paradigm for accelerating clinical data science at stanford medicine, 2020.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning A new paradigm for accelerating clinical data science at stanford medicine, 2020

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:24.207489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:06.481447Z digest=sha256:ca887383b48b33a2ba63a7405bfe74246d0f2998257c2e8502ae308921ce5666

Observation 326ae02a-724f-4313-8763-1e01c822d791 · outbound

This paper cites Electronic health records: then, now, and in the future.Yearbook of medical informatics, 25(S 01):S48–S61, 2016.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Electronic health records: then, now, and in the future.Yearbook of medical informatics, 25(S 01):S48–S61, 2016

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:06.646162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:06.646162Z digest=sha256:af24f86eda14b0fe3891f2345c318d6a531b291c2201cacaccff39a133a15ac8

Observation 20d7d28e-5b28-48fa-a794-14c19f6efdd7 · outbound

This paper cites Fleming, Alejandro Lozano, William J.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Fleming, Alejandro Lozano, William J

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:23.908708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:06.758263Z digest=sha256:dc617afb4e666798b5522a9ec15460f56cf96f213fc44d6ec37b53d105d7ba0d

Observation c33a26b2-608e-4b8e-b165-9dbfb4c71594 · outbound

This paper cites Metrics for multi-class classification: An overview.stat, 1050:13, 2020.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Metrics for multi-class classification: An overview.stat, 1050:13, 2020

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:23.653417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:06.880668Z digest=sha256:75e88956f93c8879ce652cec911a192d867f073133b0be67665949db3af52f20

Observation f6f7d809-fa99-4a90-8eff-08d7f966d1db · outbound

This paper cites The Llama 3 Herd of Models.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning The Llama 3 Herd of Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:07.001760Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:07.001760Z digest=sha256:07be991ed71cd74d5c90e2207f98fedbb19d60d6e5e84868981195989fc634d4

Observation 8e34339c-d7bc-46cd-a13f-f39629154c2d · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:07.140634Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:07.140634Z digest=sha256:d4207f3592ba31a9c020920997d370ba937ce2631d54637347b0865f8fd31c38

Observation 0c5e5c1d-46c5-4dd2-8e9d-b21bbf86b506 · outbound

This paper cites Kale, Greg Ver Steeg, and Aram Galstyan.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Kale, Greg Ver Steeg, and Aram Galstyan

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:23.408150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:07.316228Z digest=sha256:a8f0ec927e305105e69922ab2558b6533df506b16a3b94745e0b770ce3e219d4

Observation f99164e9-49d7-417a-800b-205e0852b455 · outbound

This paper cites GenCLS++: Pushing the Boundaries of Generative Classification in LLMs Through Comprehensive SFT and RL Studies Across Diverse Datasets.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning GenCLS++: Pushing the Boundaries of Generative Classification in LLMs Through Comprehensive SFT and RL Studies Across Diverse Datasets

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:07.437399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:07.437399Z digest=sha256:5f491e0fb678a31a6aecffab8d293ac1fd4b09ea8d1a39ee3e96ee2bd044ba2b

Observation 856cf426-b842-4c7a-9ba6-3811d2568764 · outbound

This paper cites Measuring massive multitask language understanding, 2021.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Measuring massive multitask language understanding, 2021

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:23.231577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:07.577410Z digest=sha256:bdbb0c2dbf25cc3b1604fcac2ef8cea20ad2b0d260482b211ecc2c4296dcbfdc

Observation 1ad0381e-b8c5-4f9d-a42d-14fa07e5d01c · outbound

This paper cites DeepRetrieval: Hacking Real Search Engines and Retrievers with Large Language Models via Reinforcement Learning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning DeepRetrieval: Hacking Real Search Engines and Retrievers with Large Language Models via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:07.713578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:07.713578Z digest=sha256:81622ad262e0c80c8be6c8992678539b0c017c3cf13dcc3330c9583cca7c9960

Observation daf9734b-6cb2-4714-a051-00d83dff8885 · outbound

This paper cites Reasoning-Enhanced Healthcare Predictions with Knowledge Graph Community Retrieval.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Reasoning-Enhanced Healthcare Predictions with Knowledge Graph Community Retrieval

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:07.860220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:07.860220Z digest=sha256:27433ed04a57c6b2373cc7139c7ddc0e042ecf96a8d4cdccb4e882f609234f1e

Observation 25ccaad0-916f-4dae-bbdd-6ece313ae861 · outbound

This paper cites Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:07.992856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:07.992856Z digest=sha256:124aa2dad51116937a560c84b20dbf423bc842071233e503b7c40a2cc3ff8f9e

Observation 7181a258-ea6a-4c20-bbf8-aacec9d18513 · outbound

This paper cites What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:08.110832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:08.110832Z digest=sha256:614950c5848ac1cd14852566cdbede90b808fbbdf8443e80d76937d6429d1211

Observation 81e95fa0-d63a-404d-b55a-74dc0d2175e0 · outbound

This paper cites Medcalc-bench: Evaluating large language models for medical calculations.Advances in Neural Information Processing Systems, 37:84730–84745, 2024.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Medcalc-bench: Evaluating large language models for medical calculations.Advances in Neural Information Processing Systems, 37:84730–84745, 2024

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:08.270915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:08.270915Z digest=sha256:600aefc96bcf7df553989009b20ec97ba82f0aa45080b78f7e395073ceb4d163

Observation 601f646c-4f78-4781-8a0d-ba13c52e6856 · outbound

This paper cites Enhancing LLMs' Clinical Reasoning with Real-World Data from a Nationwide Sepsis Registry.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Enhancing LLMs' Clinical Reasoning with Real-World Data from a Nationwide Sepsis Registry

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:40:16.829938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:08.444451Z digest=sha256:dc9b27dbb3ee5dbf86b01530248076fe8aa21ed3a014218e07adda30a0f91a19

Observation 6afc9b05-c40c-40cf-896b-767efb14c6d5 · outbound

This paper cites Med-r1: Reinforce- ment learning for generalizable medical reasoning in vision-language models.arXiv preprint arXiv:2503.13939, 2025.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Med-r1: Reinforce- ment learning for generalizable medical reasoning in vision-language models.arXiv preprint arXiv:2503.13939, 2025

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:08.591865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:08.591865Z digest=sha256:6b1611e45bae0d78b6d8a1a82d6e79333208208a2c6b11a180ea44421a2c0cc0

Observation 4e93ba54-3132-4bc9-b709-d718fdab05db · outbound

This paper cites ClinicalGPT-R1: Pushing reasoning capability of generalist disease diagnosis with large language model.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning ClinicalGPT-R1: Pushing reasoning capability of generalist disease diagnosis with large language model

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:08.713290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:08.713290Z digest=sha256:712827c870d5f3717a1834a041a9e6770f1de7c9df0fa9f0c6c7cb7f2a2d46c6

Observation 2fa7090f-a704-4064-b023-1af0552eff29 · outbound

This paper cites Rlaif vs.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Rlaif vs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:08.886612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:08.886612Z digest=sha256:f63cdea72fb9313cddb7a8fc734600aac17f3aecee252feb26a71c2deb956e6d

Observation 2975b656-7471-4252-88d4-89164de1da41 · outbound

This paper cites A scoping review of using Large Language Models (LLMs) to investigate Electronic Health Records (EHRs).

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning A scoping review of using Large Language Models (LLMs) to investigate Electronic Health Records (EHRs)

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:09.003800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:09.003800Z digest=sha256:bff0066db29d19b6f5033345bf3a24840e409ceaf24f1683bc31a866878d72de

Observation 94c35891-3a53-4a1c-a2bf-6535222ead76 · outbound

This paper cites Cls-rl: Image classifica- tion with rule-based reinforcement learning.arXiv preprint arXiv:2503.16188, 2025.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Cls-rl: Image classifica- tion with rule-based reinforcement learning.arXiv preprint arXiv:2503.16188, 2025

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:09.163179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:09.163179Z digest=sha256:fe1bbb4378b3bd6f746341170035b7f545a2aafb0e608936c88f493bbf20c744

Observation 441e5fc8-9d7d-4e00-a28d-9e403f6397ce · outbound

This paper cites Rec-r1: Bridging generative large language mod- els and user-centric recommendation systems via reinforcement learning.arXiv preprint arXiv:2503.24289, 2025.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Rec-r1: Bridging generative large language mod- els and user-centric recommendation systems via reinforcement learning.arXiv preprint arXiv:2503.24289, 2025

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:09.294209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:09.294209Z digest=sha256:67a314fd7197b6ee85be40aee477378f10d7b3ac39503d64bbd9f643bf8fc02c

Observation df6f4c4f-8547-405f-9a7c-128c6944d21e · outbound

This paper cites Panacea: A foundation model for clinical trial search, summarization, design, and recruitment.medRxiv, pages 2024–06, 2024.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Panacea: A foundation model for clinical trial search, summarization, design, and recruitment.medRxiv, pages 2024–06, 2024

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:23.004693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:09.445067Z digest=sha256:c68e6380c8caf08a7d8732b1898fcbcab3b292ac25f4736244673c1c57ebab65

Observation bf925afc-28e4-41b7-8cb3-7827c2d4e9dc · outbound

This paper cites Pisces: A cross-modal contrastive learning approach to synergistic drug combination prediction.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Pisces: A cross-modal contrastive learning approach to synergistic drug combination prediction

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:22.734561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:09.593187Z digest=sha256:f8205471552a1613c2af12f832b05c50a33d228d4e32fbbc0495859291d9ff4d

Observation 77f82f1d-90bb-424d-ad69-2b7662b4c4cf · outbound

This paper cites ReFT: Reasoning with Reinforced Fine-Tuning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning ReFT: Reasoning with Reinforced Fine-Tuning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:09.774303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:09.774303Z digest=sha256:b36f18143af99543c46bc3c9f920967c8948e1aebf9ba06e5395870337e61ec5

Observation 747e0839-1ec3-49c4-b4f8-1e95d895c7b6 · outbound

This paper cites Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:09.933663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:09.933663Z digest=sha256:8f6facb46b714e20a8f805df45f4021f58aad7aad53d23ff924e6ad3f1015f30

Observation 05492bca-953b-45f2-ba62-608ce5783f72 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:10.057892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:10.057892Z digest=sha256:99601ee9a6b06a3b0199f8cf811d72ed0565b963cdf8b7ae0574cccb9ca4f303

Observation 15441b47-519d-44f8-b74b-2f1a43c0cf99 · outbound

This paper cites Can generalist foundation models outcompete special-purpose tuning? case study in medicine.Medicine, 84(88.3):77–3, 2023.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Can generalist foundation models outcompete special-purpose tuning? case study in medicine.Medicine, 84(88.3):77–3, 2023

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:22.511312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:10.157196Z digest=sha256:4836a6061e7997461941cd42eb65f8b1eec4d05ff8f5e34def686e995a38982e

Observation d30ad90a-484b-4757-8c9b-68502463b61d · outbound

This paper cites Openai o3-mini.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Openai o3-mini

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:22.359460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:10.270598Z digest=sha256:07847331b5456d2332220d2f9f512baac499831b2dbf41db037868b1906472a3

Observation 4cb3e704-503c-471f-aa35-5a853328c76e · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:10.395468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:10.395468Z digest=sha256:3de3805fc4dd89bbc8bd08dabd5864ec563751db75dc597dc1be2400452da4eb

Observation f1baf4f4-88fa-417e-9e9b-0245d92c04ff · outbound

This paper cites Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744, 2022

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:10.512776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:10.512776Z digest=sha256:8952002c68e56763e994b978638dd52de75132a20d8a00a62e6e2f6562946b18

Observation e539073f-d59f-4e20-8372-62e44b4d1d5b · outbound

This paper cites Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Medmcqa: A large-scale multi-subject multi-choice dataset for medical domain question answering

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:10.620491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:10.620491Z digest=sha256:915717685dc874b6895c8cc9615a1e6c7defa5234873eac5fd4300d00ea61f63

Observation ea8fe538-4e8d-4330-8df7-7642bab30f10 · outbound

This paper cites MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:10.729078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:10.729078Z digest=sha256:5a7fd55dec6b213ade912388e89cad28d599e19f4623f71bdd0eab39df8147d2

Observation a5658105-bbc9-4f79-9842-5336daec8d31 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:10.817997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:10.817997Z digest=sha256:a5b6f033a88be5fad673020455cb570518dc6a4a19b72e89ca3494804e2db2d1

Observation a6d1740c-1ff0-4826-a24e-ffe0f51b25dc · outbound

This paper cites Open-Medical-R1: How to Choose Data for RLVR Training at Medicine Domain.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Open-Medical-R1: How to Choose Data for RLVR Training at Medicine Domain

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:10.939828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:10.939828Z digest=sha256:8d2df93834ad5109ff74737ef34d9f8d51a18a69db9dc3eaf2b22ad1edcc6cfd

Observation 89a6aaa2-52c3-446e-bdf2-aba93ce4177a · outbound

This paper cites Manning, and Chelsea Finn.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Manning, and Chelsea Finn

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:11.066282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:11.066282Z digest=sha256:1fc0e0603801b39222880bee866348ae2bb45725af6a026e32079c492cdd048d

Observation 41a07926-025e-42f1-ac89-d8f46f86f932 · outbound

This paper cites Overview of the trec 2021 clinical trials track.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Overview of the trec 2021 clinical trials track

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:22.244817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:11.129702Z digest=sha256:41c0058c31498a08c1af5fce24228ab5bd69b6482f3254071a99952ccbf7dc42

Observation ef5a72be-b684-48ee-bdcb-77fea2b15e24 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:11.264725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:11.264725Z digest=sha256:e63cfe9171f21dde3c1798d7c82c0bf16d677d4730a93cc92bb036aa7638b0bf

Observation 696f8fb7-b375-40b8-9f4d-88b6a322a84b · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:11.414091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:11.414091Z digest=sha256:9c08966432e841c11021edad229d8f4f1b645935cedda6e02ea9ff8d6f41831a

Observation cdcc5a0b-5908-41c5-bf23-98b54e24d8f8 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning HybridFlow: A Flexible and Efficient RLHF Framework

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:11.531409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:11.531409Z digest=sha256:fd766b8189227168afce138312037d9863624a6a8768ff1d285df2152cfc27b2

Observation 54a5c2ea-eec3-4fa9-9a1d-cce29d0c827d · outbound

This paper cites Large language models encode clinical knowledge.Nature, 620(7972):172–180, 2023.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Large language models encode clinical knowledge.Nature, 620(7972):172–180, 2023

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:11.665979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:11.665979Z digest=sha256:3d69df223e85dd31bf3e19b2ff8570b0b80d4a77bb75adaa0229c84b321f4679

Observation 7a7dded0-d419-47c2-9967-0ad3bdb0274f · outbound

This paper cites Toward expert-level medical question answering with large language models.Nature Medicine, pages 1–8, 2025.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Toward expert-level medical question answering with large language models.Nature Medicine, pages 1–8, 2025

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:22.139607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:11.804203Z digest=sha256:925956794b033bafa2831f2b16fc2928bfe0b32da30c4c567a2353425a2cfa72

Observation 8e4dc3a9-b1cb-43b2-bfe1-a9e3461b421e · outbound

This paper cites R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:11.914141Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:11.914141Z digest=sha256:b894f69f7d1ad4df029b32e886e7d462010560602c4d7bd9eaad8f666d823571

Observation cd557f50-9eff-4477-9534-8923451c3e2e · outbound

This paper cites GMAI-VL-R1: Harnessing Reinforcement Learning for Multimodal Medical Reasoning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning GMAI-VL-R1: Harnessing Reinforcement Learning for Multimodal Medical Reasoning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:12.027646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:12.027646Z digest=sha256:6e0512caf626dfe74e0b335a106a2a2f18c0ca7f9c8dc3010807a328edda94b2

Observation 645593a1-35a9-47e3-9965-95a705d71bbf · outbound

This paper cites Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Crossing the Reward Bridge: Expanding RL with Verifiable Rewards Across Diverse Domains

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:12.121976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:12.121976Z digest=sha256:8e8a761cd4900eabeb06bcbc3f7039a9cfa456f91313475f369cd52bcfcf5fa3

Observation 72ca7834-9d41-4ad0-9926-1b3fb3afe746 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:12.244157Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:12.244157Z digest=sha256:c2e3bb2909496ce600b62f23f14e49ba88f683668983ef76c6bd5268fa257c66

Observation 13ca0bb1-7e27-4cd9-a282-2bb9402481cf · outbound

This paper cites Yet another ICU benchmark: A flexible multi-center framework for clinical ML.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Yet another ICU benchmark: A flexible multi-center framework for clinical ML

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:21.914155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:12.354142Z digest=sha256:43c3f4e3f47d8c813d4c9a0ebc22f09c0f50ee77debb16f66504bdb9cb2e3715

Observation c12f8d70-63dd-4ca7-adbe-2b50f4f32d0b · outbound

This paper cites Drg-llama: tuning llama model to predict diagnosis-related group for hospitalized patients.npj Digital Medicine, 7(1):16, 2024.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Drg-llama: tuning llama model to predict diagnosis-related group for hospitalized patients.npj Digital Medicine, 7(1):16, 2024

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:21.641415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:12.471664Z digest=sha256:4e872f5d618e0b7f722acbf44f54dda0a982f6ad27b2bc78829c877c5656eaed

Observation 9a51b7fa-1c5c-49d7-92f2-0c1a3ea56ff4 · outbound

This paper cites Reinforcement Learning for Reasoning in Large Language Models with One Training Example.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Reinforcement Learning for Reasoning in Large Language Models with One Training Example

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:12.554505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:12.554505Z digest=sha256:282d67a4a0c6eef33955a4d3dc244e46b482ccf86e6ae754ec8b892b22bc6633

Observation c1ddffc4-6a31-42d1-9fad-801d8cd29507 · outbound

This paper cites Ehrshot: An ehr benchmark for few-shot evaluation of foundation models.Advances in Neural Information Processing Systems, 36:67125–67137, 2023.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Ehrshot: An ehr benchmark for few-shot evaluation of foundation models.Advances in Neural Information Processing Systems, 36:67125–67137, 2023

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:12.648627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:12.648627Z digest=sha256:f6e9190aff1d8fedc572a26677601c9c5ef98b2ff8c6a56dbe9cee3b584799ba

Observation a76b5e70-0439-4c0a-aa4e-d73a821e1bfa · outbound

This paper cites The Shaky Foundations of Clinical Foundation Models: A Survey of Large Language Models and Foundation Models for EMRs.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning The Shaky Foundations of Clinical Foundation Models: A Survey of Large Language Models and Foundation Models for EMRs

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:12.766149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:12.766149Z digest=sha256:8f7bf033a57b45e5ced877f99d0d8258528a3d017086b3eef61eebacc9258110

Observation 15631cde-5f52-4257-aac4-7dd590444ca3 · outbound

This paper cites PathVLM-R1: A Reinforcement Learning-Driven Reasoning Model for Pathology Visual-Language Tasks.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning PathVLM-R1: A Reinforcement Learning-Driven Reasoning Model for Pathology Visual-Language Tasks

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:12.932030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:12.932030Z digest=sha256:d2250c5ab174e5d09bd730d6a1973e08a9d12b3c09d2883d7f26389e97d2a1ed

Observation b08205ac-ba61-43a9-a7d5-dbcb2efed3b2 · outbound

This paper cites Sailing by the Stars: A Survey on Reward Models and Learning Strategies for Learning from Rewards.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Sailing by the Stars: A Survey on Reward Models and Learning Strategies for Learning from Rewards

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:13.070016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:13.070016Z digest=sha256:318c8f56e1aaaa558ea03ce7bbc56ca159db05ed7ce174fc7ff4580c904c8933

Observation e5edff52-49fb-4c3a-87aa-6758030fa9da · outbound

This paper cites Instruction tuning large language models to understand electronic health records.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Instruction tuning large language models to understand electronic health records

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:21.445973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:13.220372Z digest=sha256:17ff50c9a84a58c161b94fc785c40c020b764524f8f8728c8d6cd581e5da3993

Observation 0f342469-489a-4218-8dc4-7051792f570d · outbound

This paper cites Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:13.347732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:13.347732Z digest=sha256:8c3b038432b54cea3ab8fe98b2f05b39684c790285495628c625c2b44827e64e

Observation 84b1490f-7a21-4bb0-a6b9-faf8d27154c5 · outbound

This paper cites Qwen3 technical report, 2025.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Qwen3 technical report, 2025

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:21.277106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:13.455171Z digest=sha256:394695a952c3cf57dd43df8be07b94d9b8a50aa423a4e86bf46e987ca3c5ee55

Observation 9c2a96cf-38b2-41b9-9ea6-59fc39a26ca3 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:13.610491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:13.610491Z digest=sha256:62dc758f8581f97a7051b17c77c68004cd225e646943200b21ca08b11525778b

Observation c42987b4-5995-4040-8344-5dd7c2152bac · outbound

This paper cites Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Med-RLVR: Emerging Medical Reasoning from a 3B base model via reinforcement Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:13.760606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:13.760606Z digest=sha256:6d79e4462deb8dd3b0d6dd05820c1d090249ef90334d32a8ce6260c6dee1e681

Observation 2f092779-c71c-4252-97c3-7069f5184840 · outbound

This paper cites Rank-R1: Enhancing Reasoning in LLM-based Document Rerankers via Reinforcement Learning.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Rank-R1: Enhancing Reasoning in LLM-based Document Rerankers via Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:13.873620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:13.873620Z digest=sha256:67119c39a8346d04f95faaa173492fee35a41ddb0269bbada605e9117f92d826

Observation 2c36a95e-4922-4f60-ac37-678ff6b83b2e · outbound

This paper cites Details are provided in Section F.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Details are provided in Section F

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:21.135687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:14.010225Z digest=sha256:adacee1b68508f8cd3ffd96eb4697fe7b7d9501bd802861dbc2d79332d0a57eb

Observation b3e4d5c4-aeb4-4a14-ae2a-36027c04532f · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:40:21.008685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:14.133635Z digest=sha256:38ff9e706595b02b4c3a7989a13a0c6ad2e0626668b0820279c7a1a7e55a584f

Observation 89263a09-dd93-48a7-9451-70723d27474b · outbound

This paper cites See Section F.3.4.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning See Section F.3.4

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:20.770263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:14.231461Z digest=sha256:026ab9f5a58fe59cc9fccb9638c045c22833bd781f3855b4911d9c495d51864f

Observation 7f6da25b-063a-4bec-bcc1-748d665d2a26 · outbound

This paper cites Given a dose of Drug A, what is the equivalent dose of Drug B?.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Given a dose of Drug A, what is the equivalent dose of Drug B?

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:20.562596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:14.373677Z digest=sha256:f882f6b3415f3fea05264ee038077d9ea10639fd7dbd99c97a523fff6a73b340

Observation 3f082de1-aab8-42ea-9684-4b33f0fe3537 · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 69

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:40:20.376274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:14.559299Z digest=sha256:24ba87c1cd8b475bfd85b716976d69a864bc40ae84fb787060edb35a95482847

Observation 4180e1d6-005e-4cd2-8e52-377b263abc24 · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 70

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:40:20.075209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:14.725894Z digest=sha256:bf73a319a9b19d27f7a8da3645b1835ace305f97e9201bf4512264791d686a40

Observation 1538b0e0-2b9b-47d7-9e2e-13491bcfa142 · outbound

This paper cites Key Considerations: - Carefully **evaluate each inclusion and exclusion criterion individually**.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Key Considerations: - Carefully **evaluate each inclusion and exclusion criterion individually**

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:19.826437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:14.845286Z digest=sha256:b370ed2d92a72718fb9d8c6681f940c2c142d538ab8b9899a61564ed90bbe864

Observation 78876dcd-e8c1-425c-9b11-9bb14550e8ad · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:40:19.596276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:14.988132Z digest=sha256:cfdc90416e029ceb57f6246b245e7163d26a34187db5acf05b46b52785d24d2d

Observation 1f7bfa90-eef2-488c-b6be-8211fe1cac03 · outbound

This paper cites The reasoning should be comprehensive, medically sound, and clearly explain how the patient’s information leads to the predicted outcome.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning The reasoning should be comprehensive, medically sound, and clearly explain how the patient’s information leads to the predicted outcome

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:19.047445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:15.327917Z digest=sha256:79e98d3884c58d820833437dfffb89f1d5a99c9a4d2cf82091a510290867aedf

Observation 387b5318-5e61-4ea8-83b0-442d1ce234c6 · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:40:18.784235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:15.442336Z digest=sha256:e7ac9c97b9ac4d3913d85446642aeeed8eecd33589c88ab1a19c2192884ce0cf

Observation 5e44305c-dc9b-472a-a6dd-159556210ead · outbound

This paper cites Very Confident.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Very Confident

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:18.288998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:15.652511Z digest=sha256:1bbe21dd18782d821fe16195e63685d3680b9d20462f17033b7416f343bafa4e

Observation 92b791e6-2377-437c-a8b9-e4c34729e477 · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:40:18.047633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:15.731643Z digest=sha256:eb35b28402f5603ec41c3e5c3d3c237e9a6de3f28d1daf3b0f7693038a07d5b2

Observation 7b2ca135-5919-404b-be3d-7db4b0520553 · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:40:19.322411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:15.782070Z digest=sha256:420513d000da3c4771c3bf4066054d5ec75e29e77c90a5a9fd67cb3e9f9a8474

Observation 194b7b38-6e3b-4078-b552-77fe310ea1e4 · outbound

This paper cites The reasoning should be comprehensive, medically sound, and clearly explain how the patient’s information leads to the predicted outcome.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning The reasoning should be comprehensive, medically sound, and clearly explain how the patient’s information leads to the predicted outcome

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:17.796523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:15.870417Z digest=sha256:fc5541258a6cc7f42c39df1260ebec97aa303713f94284a96f7b86e4ac778c92

Observation 6e146c0a-7796-4a19-8648-cd5cd0d43b6b · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:40:17.503210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:15.940088Z digest=sha256:74d92d4edde0fbb61b5a3334f13077ad10f8dfec0833a730e5c30aecb993cef1

Observation 0b633e40-d006-4f11-9604-1435024ec891 · outbound

This paper cites an unresolved cited work.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:40:18.546105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:16.057297Z digest=sha256:2360f78bf8d22613bd9ff73954e162c6fecc8a551a4014a0eee7e34d2d2f6018

Observation 72c93ba4-f781-4338-a665-c3c17665769b · outbound

This paper cites Very Confident.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning Very Confident

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:40:17.248724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T12:40:16.167357Z digest=sha256:f7109366f998c19160fe3c361813711ebb019e193a03ee6029528a0cfd18b6ce

Pith citing papers

Observation 7e1fa655-5f32-4f99-853b-347455265ae0 · inbound

A Comprehensive Survey of Electronic Health Record Modeling: From Deep Learning Approaches to Large Language Models cites this paper.

A Comprehensive Survey of Electronic Health Record Modeling: From Deep Learning Approaches to Large Language Models Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning

Reference 170

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:34.643403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:34.643403Z digest=sha256:3d9e2954e1ef7be384fca0e99bbb1f5995d597056614b8e3cb8a1e8a65995a98

Observation 415c1475-4b7f-4c82-8317-a9fc9300d54b · inbound

Scalable Stewardship of an LLM-Assisted Clinical Benchmark with Physician Oversight cites this paper.

Scalable Stewardship of an LLM-Assisted Clinical Benchmark with Physician Oversight Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:23:23.350397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-16T20:21:40.867354Z digest=sha256:107b5ef54293582764eab86147078cd54aa8edaec4957e65022319d6951c08c6

Observation 336f4c34-8c09-48f6-abde-99a5db98e613 · inbound

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments cites this paper.

From Pixels to Digital Agents: An Empirical Study on the Taxonomy and Technological Trends of Reinforcement Learning Environments Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning

Reference 183

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:23:27.263072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T01:20:03.181903Z digest=sha256:cb9167c5bd80b778b030e47146301174bf63dc678caed71b4c8549075bdcb952