Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:28:11.158639Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 1 inbound Pith citation observation for arXiv:2504.18080.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T10:28:11.158639Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-12T04:32:16.930291Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T06:11:23.712115Z
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 73b2b576-4d87-421c-b95f-bf107ef1ed77 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Superhuman performance of a large language model on the reasoning tasks of a physician
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98110f26-a7fc-47bf-a384-64c93f1bf09c · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73d05cbc-4d28-44e1-a33b-8e24486b64ac · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization MEDITRON-70B: Scaling Medical Pretraining for Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10d34416-f6c7-421b-b9d6-3a57d230530a · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Beyond fine-tuning: Unleashing the potential of continuous pretraining for clinical LLM s
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 24ed6c3b-ce00-4dd1-99df-09b563b1a13d · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Towards a Personal Health Large Language Model
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6092b268-ad98-440c-9cee-d7759edc8355 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457ee354-005e-4c35-a054-5570fb6d34c9 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization The case for 4-bit precision: k-bit Inference Scaling Laws
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c58a6b2-b601-4b17-87fb-29f5ff648528 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Qlora: Efficient finetuning of quantized llms
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 917fb95d-8361-484a-a09e-1918df96bef9 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Does Fine-Tuning LLMs on New Knowledge Encourage Hallucinations?
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19c99389-363b-4da0-ad6c-b4fc0e67369e · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization The Llama 3 Herd of Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81d060c2-dbc1-4d9e-8c77-a889a1d130e7 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3060ebd3-1cd8-4034-b2f2-352701584d4b · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Measuring Massive Multitask Language Understanding
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c921b51-133c-41f9-a97d-a26f3cec6eab · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization LoRA: Low-Rank Adaptation of Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c261eda-d21e-4720-adf2-46af41367546 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization JMedBench: A Benchmark for Evaluating Japanese Biomedical Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eaf16e5-6cce-4550-b987-572c5adf0e95 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization What Disease does this Patient Have? A Large-scale Open Domain Question Answering Dataset from Medical Exams
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca91b0bb-b531-42d0-bbad-1c23b7753a27 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization P ub M ed QA : A dataset for biomedical research question answering
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb7c69c1-495e-4a23-b7b0-3af71d1f3876 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization An evaluation framework for clinical use of large language models in patient interaction tasks
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 48b23760-3ce1-4fc0-8875-80697cb08a23 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization MEDIC: Comprehensive Evaluation of Leading Indicators for LLM Safety and Utility in Clinical Applications
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b748f13-a5e6-4a0c-ba0e-904bb31ace17 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Evaluating GPT-4 and ChatGPT on Japanese Medical Licensing Examinations
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73cb7e17-8456-4916-863a-c19e0428d06a · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Medical Hallucinations in Foundation Models and Their Impact on Healthcare
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f15475b9-a3b8-44c3-a75f-052992fea374 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization ReFT: Reasoning with Reinforced Fine-Tuning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98a9ab13-58b5-4882-b961-1ad3d3c2f96c · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Radlink: Linking clinical entities from radiology reports
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b1b65886-20a3-42e7-9f14-eed8dcb70513 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f8dc36a-7d12-4725-9937-29c4223836e0 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Can Generalist Foundation Models Outcompete Special-Purpose Tuning? Case Study in Medicine
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31daaaa7-af17-491b-a9cb-9bf932f12744 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization From Medprompt to o1: Exploration of Run-Time Strategies for Medical Challenge Problems and Beyond
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72b9065e-cbe2-47b3-a690-5043ba97eb17 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization openai/MMMLU , 2024
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8105fb90-7260-4de8-bf8d-5d87853ef6f3 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization GPT-4o System Card
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e88db10-a75a-416c-959b-26fb15c1070b · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Training language models to follow instructions with human feedback
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90ebc75e-2d21-4aa1-ab54-f6d02fbe2b3f · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization MedMCQA : A Large-scale Multi-Subject Multi-Choice Dataset for Medical domain Question Answering
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a12ed395-812b-4182-b83d-ddb8330022a0 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Iterative Reasoning Preference Optimization
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1e73f53-7ab7-4ef8-9e3a-f2f32623bc35 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization PLaMo-100B: A Ground-Up Language Model Designed for Japanese Proficiency
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 76b0664a-0ccf-48f5-9665-455016c42998 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Qwen2.5 Technical Report
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77f88661-dd29-4745-9d94-a600beb808dc · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Direct Preference Optimization: Your Language Model is Secretly a Reward Model
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bd2fa28-c7df-427d-84d2-3e52dcf5aabf · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Capabilities of Gemini Models in Medicine
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 523c3ae4-07c3-431e-84ca-17c0ca95756d · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Toward expert-level medical question answering with large language models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a38297a-6470-4172-8e5f-cea04e9e3511 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Development and bilingual evaluation of Japanese medical large language model within reasonably low computational resources
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f7c7534-552b-4cc9-b1e7-63a1b779ac28 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization 70B-parameter large language models in Japanese medical question-answering
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb5034e7-d180-48f5-88a7-8867bba2d594 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Towards conversational diagnostic artificial intelligence
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7becc1bc-dc26-40de-92e7-45a1ff2c6117 · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization Adapted large language models can outperform medical experts in clinical text summarization
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c348976a-fe2c-4d1f-9b59-55dc43f7fabc · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization A Preliminary Study of o1 in Medicine: Are We Closer to an AI Doctor?
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5da5b1f8-702f-43bc-a4f8-3ac543734e9a · outbound
Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization HuatuoGPT, towards Taming Language Model to Be a Doctor
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a795170-6248-4603-bd60-8e1386c46c69 · inbound
CLR-voyance: Reinforcing Open-Ended Reasoning for Inpatient Clinical Decision Support with Outcome-Aware Rubrics Stabilizing Reasoning in Medical LLMs with Continued Pretraining and Reasoning Preference Optimization
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.