Pith. sign in

Paper Citation Record · LEDGER

Disentangling Reasoning and Knowledge in Medical Large Language Models

As of 18 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 5 inbound Pith citation observations for arXiv:2505.11462.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.11462 v2

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:57:56.132621Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:08:56.179861Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T08:16:47.669880Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved32
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 25c886f1-e267-4ca3-ac33-80c046f1868f · outbound

This paper cites an unresolved cited work.

Disentangling Reasoning and Knowledge in Medical Large Language Models Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:57:57.050512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:55.929498Z digest=sha256:b62a7006111b7adb0cd2bce5c50e13eb599b621b097ea2562f0b5bb09ecf83fc

Observation 2c5c92f9-357a-45f0-bdff-aa587e9c0dc3 · outbound

This paper cites G., and Sutphen, M.

Disentangling Reasoning and Knowledge in Medical Large Language Models G., and Sutphen, M

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:57.038182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:55.934322Z digest=sha256:b32b3a3b575f45d4f011b39a3fdaf653317952f4e548f3355fcca8de19780979

Observation 42562aad-0855-42f8-8431-cf4b568ca13f · outbound

This paper cites Lessons From Red Teaming 100 Generative AI Products.

Disentangling Reasoning and Knowledge in Medical Large Language Models Lessons From Red Teaming 100 Generative AI Products

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:55.938693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:55.938693Z digest=sha256:e13e4c20f07f05f9d9fd545b26a2e5a6dbdae04b5e9eb362246a7432631596d9

Observation ec64eca6-04bf-48fe-9e85-c88459f4df9f · outbound

This paper cites T., Farah, H., Gui, H., Rezaei, S.

Disentangling Reasoning and Knowledge in Medical Large Language Models T., Farah, H., Gui, H., Rezaei, S

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:57.026127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:55.943715Z digest=sha256:53e588a90c58538fd2e87fa5771a67ee5044ecbf4d63f56f8ead3abc1f53fdbe

Observation 3e3af9c9-8291-4b2c-9293-932ee956fe5d · outbound

This paper cites Benchmarking large language models on answering and explaining challenging medical questions.

Disentangling Reasoning and Knowledge in Medical Large Language Models Benchmarking large language models on answering and explaining challenging medical questions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:55.948153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:55.948153Z digest=sha256:fea87d2fbd6d6f57f7f2802233b75aeb8a4234463d945704933d5774194dad9d

Observation 87344c85-6269-4546-8a89-88cc6300acce · outbound

This paper cites SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models.

Disentangling Reasoning and Knowledge in Medical Large Language Models SFT or RL? An Early Investigation into Training R1-Like Reasoning Large Vision-Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:55.952413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:55.952413Z digest=sha256:e884db4e8f4040752de93639505d0c34ab73c1cf3fa4d61a9125d198dd7f7fe5

Observation 71c28014-abc5-4cd1-a784-3518fa783ac5 · outbound

This paper cites HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs.

Disentangling Reasoning and Knowledge in Medical Large Language Models HuatuoGPT-o1, Towards Medical Complex Reasoning with LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:55.957637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:55.957637Z digest=sha256:9a8c624d3c1983acd4e486f76dc8a34de1ce141540449142edfbd654f2d26efe

Observation 6248ec79-47b2-4d00-be1c-1560e2360dc9 · outbound

This paper cites SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training.

Disentangling Reasoning and Knowledge in Medical Large Language Models SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:55.962021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:55.962021Z digest=sha256:b7b2c6c2ffae009c80354fe279d7b36b0e4e16f78e7539c046300da7207d846f

Observation 08effff5-3236-4cc0-9951-2ff932a1644c · outbound

This paper cites A universal model of diagnostic reasoning.

Disentangling Reasoning and Knowledge in Medical Large Language Models A universal model of diagnostic reasoning

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:57.013694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:55.966821Z digest=sha256:08bfaa0d7ea00a69ccd556f8058727049e7357fcd0979edc2ba53b867ff412d6

Observation abf8c377-5417-425a-9835-0a4486263388 · outbound

This paper cites OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles.

Disentangling Reasoning and Knowledge in Medical Large Language Models OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:55.971155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:55.971155Z digest=sha256:aa2d1ef0d4a7216112d60670c37eb09f3de01138d2fdd0f611af8d0c3f363f0f

Observation d41c0ed5-6292-4a18-9e68-a11ecdc92e6f · outbound

This paper cites S., Shulman, L.

Disentangling Reasoning and Knowledge in Medical Large Language Models S., Shulman, L

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:57.001121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:55.975820Z digest=sha256:1f311d8f96ff16602761c3ab5be09665ab1640f5da32a243d426a36dcee3dcbc

Observation d8e52060-3a65-4073-bc63-b3304e3d0d67 · outbound

This paper cites Medgemma hugging face.

Disentangling Reasoning and Knowledge in Medical Large Language Models Medgemma hugging face

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.988953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:55.980393Z digest=sha256:258196393ebc12d2ea349de5a3cff885a58f1e2a2396920443fd2e3db78f8a96

Observation ca049a1b-3681-48b4-8900-39918e0e6467 · outbound

This paper cites The Llama 3 Herd of Models.

Disentangling Reasoning and Knowledge in Medical Large Language Models The Llama 3 Herd of Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:55.984400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:55.984400Z digest=sha256:e118a3650d5d3323766ef077dd36df0376eaf449ceaadb607aa450e198771c14

Observation bbd7a809-0d26-41a3-b4cf-9a7025ec5f6e · outbound

This paper cites Domain-specific language model pretraining for biomedical natural language processing.

Disentangling Reasoning and Knowledge in Medical Large Language Models Domain-specific language model pretraining for biomedical natural language processing

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.976818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:55.988501Z digest=sha256:2b4747df1953ffdababa2e0e76f15724e0100acc5ee5347cbd7e3c23c5610f57

Observation dcd2d587-e900-40db-ac9f-26e72d4e6f96 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Disentangling Reasoning and Knowledge in Medical Large Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:55.992357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:55.992357Z digest=sha256:4f1d0aa4bacaf460a1ef3b650e880faeddc9f530f91fb8e6498ce64a924b4886

Observation 9dd2dba2-847e-419c-9cad-bd4e6dceaaf8 · outbound

This paper cites K., et al.

Disentangling Reasoning and Knowledge in Medical Large Language Models K., et al

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.964732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:55.996246Z digest=sha256:7f4b3ddb51aa611a4876aa92b27ec26bb0a4910faaf28940d6bafaa071cc5c5d

Observation 80a06c1f-8567-48b8-b84a-172ec31e0619 · outbound

This paper cites m1: Unleash the potential of test-time scaling for medical reasoning with large language models.

Disentangling Reasoning and Knowledge in Medical Large Language Models m1: Unleash the potential of test-time scaling for medical reasoning with large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.000037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.000037Z digest=sha256:a51c2e3d73c4db9615cc0d5fdd2ef8c4673a3d20f0a40faf963ea5be25b468bb

Observation fd94fa28-0754-4359-854f-1548389cbd55 · outbound

This paper cites GPT-4o System Card.

Disentangling Reasoning and Knowledge in Medical Large Language Models GPT-4o System Card

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.003934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.003934Z digest=sha256:3f11ec0a86ad7ac6b7ea12fc68145ea1b2824c1c88bb228f0552c160560cbc36

Observation 45c08075-ed99-4f6b-89f6-cbc725cee737 · outbound

This paper cites OpenAI o1 System Card.

Disentangling Reasoning and Knowledge in Medical Large Language Models OpenAI o1 System Card

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.008077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.008077Z digest=sha256:f64457d9dc1f39d425bac4a5d69bb99423bc0cd0f9d537565e1f597a21b2a031

Observation 1848102d-2a74-4533-b4c0-5b38d261f2a2 · outbound

This paper cites What disease does this patient have? a large-scale open domain question answering dataset from medical exams.

Disentangling Reasoning and Knowledge in Medical Large Language Models What disease does this patient have? a large-scale open domain question answering dataset from medical exams

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.951711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.012307Z digest=sha256:53d6665b0a63cdafa88a0acd22207d0f1a9b946e1fd15373f97a4e5c0537421c

Observation fc580b76-bac7-404c-a072-9f72809b54db · outbound

This paper cites Disentangling Memory and Reasoning Ability in Large Language Models.

Disentangling Reasoning and Knowledge in Medical Large Language Models Disentangling Memory and Reasoning Ability in Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.016485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.016485Z digest=sha256:0615f8f90b0506bdafc5b82942625449c89ef3fcf6584fdc43f42b79ed050c8d

Observation 1ed913cc-4498-4a27-8f20-7d6b2f9984d1 · outbound

This paper cites PubMedQA: A Dataset for Biomedical Research Question Answering.

Disentangling Reasoning and Knowledge in Medical Large Language Models PubMedQA: A Dataset for Biomedical Research Question Answering

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.020780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.020780Z digest=sha256:509cf3d4e27a94a6cec56fa1248f9f81f9e37a4847a7534d890c3aaca77b26aa

Observation 919829a4-676a-4df7-bbf0-17245d51a1b9 · outbound

This paper cites Training Language Models to Self-Correct via Reinforcement Learning.

Disentangling Reasoning and Knowledge in Medical Large Language Models Training Language Models to Self-Correct via Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.024840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.024840Z digest=sha256:c53051f326a9c82033656d0621eadc5042ff10c3cf2d9239cbaf91e08e2c5878

Observation c793b0cf-68f1-4c80-853f-8e1dd4ea8e2d · outbound

This paper cites H., Gonzalez, J., Zhang, H., and Stoica, I.

Disentangling Reasoning and Knowledge in Medical Large Language Models H., Gonzalez, J., Zhang, H., and Stoica, I

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.832004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.029065Z digest=sha256:3e8c6e16ecc10dec6cebf4ff61da403d97db353743656737653a6e4d2ec3985d

Observation aec1cb67-43bd-49c0-92ac-1542533ea5cb · outbound

This paper cites Using script theory to cultivate illness script formation and clinical reasoning in health professions education.

Disentangling Reasoning and Knowledge in Medical Large Language Models Using script theory to cultivate illness script formation and clinical reasoning in health professions education

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.819468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.033146Z digest=sha256:9e1425b0c0bf4c40157d737d7fccfd951487ae1623bb072c23ec80996bfa2490

Observation 1de6b7fb-be87-4299-8821-d27cf50f450b · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.Advances in Neural Information Processing Systems , 36:46534–46594, 2023.

Disentangling Reasoning and Knowledge in Medical Large Language Models Self-refine: Iterative refinement with self-feedback.Advances in Neural Information Processing Systems , 36:46534–46594, 2023

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.806016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.037286Z digest=sha256:4182f4afaf54bad19f848c846fd70432dfd852c707942a14092e4a6cb39a19d7

Observation 8a9bd012-86a7-48ba-88a4-6fd1c0a45b27 · outbound

This paper cites an unresolved cited work.

Disentangling Reasoning and Knowledge in Medical Large Language Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:57:56.793235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.041290Z digest=sha256:ec942211128e75c9ca8d86fe790f7268ae939f466517ff2db700fd327129759f

Observation 381e8c61-29f0-4ba5-b7e7-5e3fc1d9dbe1 · outbound

This paper cites K., and Sankarasubbu, M.

Disentangling Reasoning and Knowledge in Medical Large Language Models K., and Sankarasubbu, M

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.780524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.045159Z digest=sha256:4be5bc2d8ec69d625e2b75393cee4e987836525d50a5e74b056c3f1ccfe3eb00

Observation 6b07d5b8-527e-4320-ad75-23b01e83cff9 · outbound

This paper cites an unresolved cited work.

Disentangling Reasoning and Knowledge in Medical Large Language Models Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:57:56.767690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.048976Z digest=sha256:ff698935788039727393dd4bff07a555dcf2633a8cf6bcc398da9661a3e3ee37

Observation 9aa8c74e-ec51-4088-a963-0edb2332cbdc · outbound

This paper cites Humanity's Last Exam.

Disentangling Reasoning and Knowledge in Medical Large Language Models Humanity's Last Exam

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.052948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.052948Z digest=sha256:a6da2d7ebe16f5da4fcc3fec4dd864c415208123769c3fadc4d2cee816e14c60

Observation 92a06937-9bbe-47a8-baa6-759e633db587 · outbound

This paper cites L., Stickland, A.

Disentangling Reasoning and Knowledge in Medical Large Language Models L., Stickland, A

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.754459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.056592Z digest=sha256:1ae36c2883104ebedc965ab7c7bf134464357f2f3e6b2805cd0cea90c8758f11

Observation d2015f21-5d21-4e5b-97bd-21d7b7e3640a · outbound

This paper cites an unresolved cited work.

Disentangling Reasoning and Knowledge in Medical Large Language Models Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:57:56.740590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.060561Z digest=sha256:92d36eae4bc5d0c40098ec94a0b53f128126d3dad6690def80fdeaa8ef6b3862

Observation 336b3cfa-207e-46d7-96e0-bc439302446b · outbound

This paper cites an unresolved cited work.

Disentangling Reasoning and Knowledge in Medical Large Language Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:57:56.727097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.064612Z digest=sha256:47aeb587751795373d4a321ba6a30e47bbb602fac065abdb3c146550b0bac6cb

Observation 41cff60e-bbae-4725-8de6-5ffc590c4e3f · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Disentangling Reasoning and Knowledge in Medical Large Language Models HybridFlow: A Flexible and Efficient RLHF Framework

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.068522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.068522Z digest=sha256:f06959de090b7a58092e171c53a5f80bf9e4265caf8bcbda9b0f1c292f5ed571

Observation a7d7f186-519a-431b-aea5-f1fb8dedf3d4 · outbound

This paper cites an unresolved cited work.

Disentangling Reasoning and Knowledge in Medical Large Language Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-15T20:57:56.714342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.072874Z digest=sha256:d15b340424f105822dff3af6245806526a0c4bab91f48572e7b2bb5ea631102d

Observation 928e991b-ae16-47cc-9919-bb09a063d528 · outbound

This paper cites Embracing complexity with systems thinking in general practitioners’ clinical reasoning helps handling uncertainty.

Disentangling Reasoning and Knowledge in Medical Large Language Models Embracing complexity with systems thinking in general practitioners’ clinical reasoning helps handling uncertainty

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.701149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.076676Z digest=sha256:2b4f1f73bdc12a96ad2ebbd5571d5759f5760ffbf06fc6be762e38bc87c0ba30

Observation be2bacc5-a788-41ef-a7cf-7a3e65752a7a · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Disentangling Reasoning and Knowledge in Medical Large Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.080460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.080460Z digest=sha256:7e1c674ba10557a6cc216e18272182065b0a3d0133835d62c51046152972b95b

Observation 41a1d290-2f33-455c-a71f-8c4621e5e815 · outbound

This paper cites HEAD-QA: A Healthcare Dataset for Complex Reasoning.

Disentangling Reasoning and Knowledge in Medical Large Language Models HEAD-QA: A Healthcare Dataset for Complex Reasoning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.084538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.084538Z digest=sha256:00e9f687c1d9f3f30976e8fc74ff9418be7836c7f0abba02f9fcae4d1dd6c26f

Observation 99bbe380-8c15-4ea5-9a63-06fd2bc697af · outbound

This paper cites Mmlu-pro: A more robust and challenging multi-task language understanding benchmark.

Disentangling Reasoning and Knowledge in Medical Large Language Models Mmlu-pro: A more robust and challenging multi-task language understanding benchmark

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.687725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.088734Z digest=sha256:f23216e5ba6b4783fbfad9d2eee124262133010100c14b97772e84fa3bcc22fa

Observation 3b492577-4f41-4d36-a32e-c5c8c3135c49 · outbound

This paper cites V ., Zhou, D., et al.

Disentangling Reasoning and Knowledge in Medical Large Language Models V ., Zhou, D., et al

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.092448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.092448Z digest=sha256:c5fdf2b07cb18d91b3ca50920ef703c64cd77a21c90fe2760db4abe9b9f4f1c8

Observation 472cceec-1574-497f-b478-430a8599104b · outbound

This paper cites Generating Sequences by Learning to Self-Correct.

Disentangling Reasoning and Knowledge in Medical Large Language Models Generating Sequences by Learning to Self-Correct

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.096318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.096318Z digest=sha256:a5ed5bb33d548888f98b6c8e09eb6a88bd92444474893f30438e24d90b732643

Observation 0078763b-4242-4659-9a27-4151b0d0c5ca · outbound

This paper cites MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs.

Disentangling Reasoning and Knowledge in Medical Large Language Models MedReason: Eliciting Factual Medical Reasoning Steps in LLMs via Knowledge Graphs

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.100584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.100584Z digest=sha256:3494c204baad31d585f57179bd88ac4e87ae18949dd9a37a2380749ebace3cd9

Observation 29473ed1-6c72-469d-8e0b-778378520ded · outbound

This paper cites Qwen2.5 Technical Report.

Disentangling Reasoning and Knowledge in Medical Large Language Models Qwen2.5 Technical Report

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.104776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.104776Z digest=sha256:b2a78775fa5233a5eb30c031c3b4f807f2d9a93eec997a633a0781b672dc16c5

Observation caa646ae-da9c-4e7e-9d66-66e9147286e2 · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

Disentangling Reasoning and Knowledge in Medical Large Language Models R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.108603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.108603Z digest=sha256:6d72d8200effbc9d192d88fd12489b2c96673b6e4e9727394a280d41e4fdd381

Observation 7570d88c-48a4-43b4-a53b-72612674e3db · outbound

This paper cites and Hoseini Abardeh, M.

Disentangling Reasoning and Knowledge in Medical Large Language Models and Hoseini Abardeh, M

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.665555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.112729Z digest=sha256:f6a8f85e56d2756e069593fabf39dde5ecdb65804859db84082a53b392460e23

Observation 77a1754e-84ad-4dae-a433-87cf075a1006 · outbound

This paper cites Backtracking Improves Generation Safety.

Disentangling Reasoning and Knowledge in Medical Large Language Models Backtracking Improves Generation Safety

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.116477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.116477Z digest=sha256:adccb0fd49cc1ec95e917f5c5981deacb6cf75d776c14d4a65d1ac9940247f3c

Observation 67c248f1-1b30-4300-9e07-7cedbadfc63a · outbound

This paper cites Easyr1: An efficient, scalable, multi-modality rl training framework, 2025.

Disentangling Reasoning and Knowledge in Medical Large Language Models Easyr1: An efficient, scalable, multi-modality rl training framework, 2025

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:57:56.651885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T20:57:56.120514Z digest=sha256:a20b54d1e778c49aacb52dce8396c71cb00b99e98555c118d01246f02a6c2e20

Observation 5d51348b-aca2-4dd7-9fc7-b40a7bc084e0 · outbound

This paper cites Evaluation of openai o1: Opportunities and challenges of agi.

Disentangling Reasoning and Knowledge in Medical Large Language Models Evaluation of openai o1: Opportunities and challenges of agi

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.124260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.124260Z digest=sha256:6a511bd02fde1d9a446972400b67155614cc8a404d91370ddb78f04f12cfcb85

Observation 664108b8-aa56-4930-bc1c-585558863c94 · outbound

This paper cites Navigating the Grey Area: How Expressions of Uncertainty and Overconfidence Affect Language Models.

Disentangling Reasoning and Knowledge in Medical Large Language Models Navigating the Grey Area: How Expressions of Uncertainty and Overconfidence Affect Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T20:57:56.128245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.128245Z digest=sha256:bd53da4848e4b07689f23dc3cebf40348c4391cc41a5a2d0d37706c7e886d8c6

Observation a6deb072-9fe2-4cc4-ba96-89e80df68f58 · outbound

This paper cites MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding.

Disentangling Reasoning and Knowledge in Medical Large Language Models MedXpertQA: Benchmarking Expert-Level Medical Reasoning and Understanding

Reference 50

Resolution
malformed identifier
no resolver link, observed 2026-08-15T20:57:56.132621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:57:56.132621Z digest=sha256:264d07c1246620775e26742f840cdb1c1c744e08e2540d963b97a4509976b784

Pith citing papers

Observation fa72c826-4858-41b7-8118-e88fec1060cb · inbound

Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs cites this paper.

Rethinking Test-Time Scaling for Medical AI: Model and Task-Aware Strategies for LLMs and VLMs Disentangling Reasoning and Knowledge in Medical Large Language Models

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T20:08:56.179861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:08:56.179861Z digest=sha256:c0d043c81cb567c77cbafff9dd7e9e7e5d9d90e84ebb4fb2e63eef97385f97ab

Observation a2fd42b1-9267-40cb-a7c9-dcdd005dd195 · inbound

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study cites this paper.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study Disentangling Reasoning and Knowledge in Medical Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:47.083790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:47.083790Z digest=sha256:6d5f276b2008d0377adac85e4a9a753798871d417da85520f98db9c321963abd

Observation d429c716-8372-44de-9094-d91b07dfceaa · inbound

Medical Reasoning with Large Language Models: A Survey and MR-Bench cites this paper.

Medical Reasoning with Large Language Models: A Survey and MR-Bench Disentangling Reasoning and Knowledge in Medical Large Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:25:26.828395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T10:21:39.892271Z digest=sha256:3ca743de72447aa667c17e1265926871ae8f8417d485e67251f6e79f9558b76d

Observation 7074bbad-d145-406e-9251-d0678ecfb4e8 · inbound

Evo-MedAgent: Beyond One-Shot Diagnosis with Agents That Remember, Reflect, and Improve cites this paper.

Evo-MedAgent: Beyond One-Shot Diagnosis with Agents That Remember, Reflect, and Improve Disentangling Reasoning and Knowledge in Medical Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:46:34.096890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T12:35:32.915685Z digest=sha256:d0d2ea2337351820cab466dccbfb9c37f321347278b2e06ca95ee863c1d21d9c

Observation 89a6fd6d-b127-4d3f-9c84-8991d5531de1 · inbound

Search-Time Contamination in Deep Research Agents: Measuring Performance Inflation in Public Benchmark Evaluation cites this paper.

Search-Time Contamination in Deep Research Agents: Measuring Performance Inflation in Public Benchmark Evaluation Disentangling Reasoning and Knowledge in Medical Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.671337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-28T06:16:54.162410Z digest=sha256:4898b0765871475be77904ae9fef533a4c42177aa6063a17a19b33a75013535b