Pith. sign in

Paper Citation Record · LEDGER

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints

As of 20 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 2 inbound Pith citation observations for arXiv:2506.06600.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06600 v2

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:57:22.751107Z

measured 35 of 35 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T19:39:46.922877Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T23:24:01.588708Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact1
  • verified fuzzy9
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c57bfb3f-f07d-4a98-a200-4525fd253797 · outbound

This paper cites Language Models are Few-Shot Learners.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Language Models are Few-Shot Learners

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.627032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.627032Z digest=sha256:997047d8eaf2cd6d75f2b63eb3900065d4d8f94170f0dceda07fb01eeed8ddb0

Observation 20461b7d-3f2e-4d9e-96ff-b7c3c5a2d8d5 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints LLaMA: Open and Efficient Foundation Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.632278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.632278Z digest=sha256:cfc222d93eadabec0b9eeaa9140328a301baf93636ec7db2d3c01f802a48d1d5

Observation 21a63095-f6ba-463f-9676-4a004d5f9353 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.637181Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.637181Z digest=sha256:897b9054050286c28cd227f6d414e1ea42f4dd486a3458b9ce28ae07f79a8e46

Observation 05346d2e-3ece-4cfa-93fd-f83fcabbd77e · outbound

This paper cites The Llama 3 Herd of Models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.641263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.641263Z digest=sha256:99762c452929c5cc5af3f36aebe6fff70a2d0df7f50efc2df9a7847abb770c1e

Observation d355fdae-063a-4c0b-890d-c6dbc462590b · outbound

This paper cites Qwen3, April 2025.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Qwen3, April 2025

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.193547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.646242Z digest=sha256:e390cbdbe2a3a2a531b09cfe2466209d085f34c6fed236672383a8f7449d5299

Observation a27ca818-e6c9-4536-ba0d-7a5c6713cedd · outbound

This paper cites Visual-language models for medical image analysis: A survey.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Visual-language models for medical image analysis: A survey

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.182550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.650068Z digest=sha256:08c0c21d18edb02af45019864ebfefcb0104a358f00b3c3f7765a334fe89b170

Observation 3d146fbb-eb46-4603-916d-570fde5b3fc8 · outbound

This paper cites Med-flamingo: a multimodal medical few-shot learner.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Med-flamingo: a multimodal medical few-shot learner

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.654852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.654852Z digest=sha256:ef6db3622fc830e00f3bf0c2a3250444079c24b4b6de816799e05f1ff14695bd

Observation 22fc5508-addf-460e-b57b-05dff12354d4 · outbound

This paper cites GEM3D: GEnerative Medial Abstractions for 3D Shape Synthesis.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints GEM3D: GEnerative Medial Abstractions for 3D Shape Synthesis

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:57:22.982969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.658151Z digest=sha256:4c177a54d1cc4c947ce9071580f285968ef4dab9e938ff1b48ddf351a5230be1

Observation e4811ea2-b637-4362-a90f-f8fe9f73e1ad · outbound

This paper cites BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints BiomedCLIP: a multimodal biomedical foundation model pretrained from fifteen million scientific image-text pairs

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.662714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.662714Z digest=sha256:2dbe4fc0f99b61de7476f425d90402184d9339c871eb08b1f7ffe04bd91ce62c

Observation b8a46e70-17ed-4b53-9668-28f92aec6133 · outbound

This paper cites Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Biomedgpt: A unified and generalist biomedical generative pre-trained transformer for vision, language, and multimodal tasks

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.666851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.666851Z digest=sha256:87e9653b8707c1175d50f3cdae9a2423ba9d71a88d9d8c47579b7ad2630a929d

Observation 621c8906-edb3-44d1-83ca-f6fdc282c21e · outbound

This paper cites Llava-med: Training a large language-and-vision assistant for biomedicine in one day.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Llava-med: Training a large language-and-vision assistant for biomedicine in one day

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.670308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.670308Z digest=sha256:8ce987aba71b037d7c6941aea8501dd2c7f276a08be82a0747e86f3f8038ed25

Observation 48f93aba-269a-4d53-b74d-ce843f717b7f · outbound

This paper cites Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Mimic-cxr, a de-identified publicly available database of chest radiographs with free-text reports

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.673917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.673917Z digest=sha256:573d434887241404723f24241ac10e2f122bc1f1f155032c057651b59f5bdc8e

Observation 1aad558e-94e1-447e-bc96-998d294aecf9 · outbound

This paper cites Robust Kalman Filters Based on the Sub-Gaussian $\alpha$-stable Distribution.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Robust Kalman Filters Based on the Sub-Gaussian $\alpha$-stable Distribution

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.677585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.677585Z digest=sha256:7189b23491f1e9a6eac89b7905c2110ca3ea532cabb3668b0ed7fe8e26a8e0a9

Observation fca66905-e254-4a17-a9af-fc70bf6fc588 · outbound

This paper cites Artificial intelligence in radiology.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Artificial intelligence in radiology

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.142062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.681298Z digest=sha256:588be3b11aa80ae65786065002d17c8f82f92bdd3f249b0284a8fd0055fa2f87

Observation 6349b17c-70c1-4041-a3ed-97f290a22c00 · outbound

This paper cites MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.684656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.684656Z digest=sha256:f77c918df197f56f2bba2273a978536890bf9801d038135f94c8eb92dd48a761

Observation eeb376da-144c-4f9e-bf4e-0ecab4f05ec3 · outbound

This paper cites Med-r1: Reinforce- ment learning for generalizable medical reasoning in vision-language models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Med-r1: Reinforce- ment learning for generalizable medical reasoning in vision-language models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.688075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.688075Z digest=sha256:2cb33b7b1d89cc5c9853f24343e11ebd15ed8e75341783c2d63c79bfe45863f3

Observation d591a105-dcc3-4bbb-b57a-fa8afe9d2838 · outbound

This paper cites Explainability for artificial intelligence in healthcare: a multidisciplinary perspective.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Explainability for artificial intelligence in healthcare: a multidisciplinary perspective

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.130793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.691442Z digest=sha256:a63fbf57b2ab3570a62e88460899a66ef5aacd6ecb7e7967864c6d9e62310290

Observation dd1ecbe9-32cb-47c8-b2ae-cec7f1c9e7aa · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Chain-of-thought prompting elicits reasoning in large language models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.118859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.695048Z digest=sha256:e1a3546f4f73cfc092ff0b3fa80c56b56396fad821e5a26497ec3a30cf15ea52

Observation c2ab2436-87e7-4f78-94c6-47d7493cd9ed · outbound

This paper cites MedCoT: Medical Chain of Thought via Hierarchical Expert.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints MedCoT: Medical Chain of Thought via Hierarchical Expert

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.698599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.698599Z digest=sha256:ba0f12096fc0f7a107c5498ab3ec7548df4fce920bc3e07a3d006109b0c67cd2

Observation b0fc92e6-b2f7-47af-b297-9e93e147bed8 · outbound

This paper cites Silvar-med: A speech-driven visual language model for explainable abnormality detection in medical imaging.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Silvar-med: A speech-driven visual language model for explainable abnormality detection in medical imaging

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.105655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.702267Z digest=sha256:066c73659a5b39d1302086f23e57a30df500859274f0cafa116ef21b7e8d6134

Observation c1813c07-a182-4df5-af8e-fb5791917559 · outbound

This paper cites Two-Stage Estimation and Variance Modeling for Latency-Constrained Variational Quantum Algorithms.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Two-Stage Estimation and Variance Modeling for Latency-Constrained Variational Quantum Algorithms

Reference 21

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:57:22.845736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.706142Z digest=sha256:27eff44eb4bb9bb0d2fa9717617c05b44b928a08e5c576c3cc93cf41f2105785

Observation 0f22b2dd-101a-4481-9ea9-42dce54c8cf9 · outbound

This paper cites Reinforcement learning for medical image analysis: Current progress and future directions.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Reinforcement learning for medical image analysis: Current progress and future directions

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.092508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.710524Z digest=sha256:ea46a8642e3d684cbe717f17438987c3f8fdb9045a10aa0366bb03a9f666d0af

Observation fe64f8fd-0251-4290-99d3-5e70ee5715fa · outbound

This paper cites Which shapes can appear in a Curve Shortening Flow Singularity?.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Which shapes can appear in a Curve Shortening Flow Singularity?

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.713892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.713892Z digest=sha256:8b5aa354a41d919398bbb6ece0e5513c68eee6bf1b4afb5e6294dc41754282b9

Observation a3f3f9be-363b-4006-9834-d829cd011408 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.717492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.717492Z digest=sha256:8698b5f7723bac2bb8971137a06fcc8d2964bc2481909d5fb1786df2a9ebd79b

Observation 34f1c314-0099-4ed5-96ae-1d085cc868c4 · outbound

This paper cites A dataset of clinically generated visual questions and answers about radiology images.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints A dataset of clinically generated visual questions and answers about radiology images

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.721052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.721052Z digest=sha256:cd824c0a441347d695575dfa161ea3063f526ef1b531298ce8b74ed2cddcd9b9

Observation 0758ad61-c5fa-4d6f-b6c8-001ca2c3f319 · outbound

This paper cites SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints SLAKE: A Semantically-Labeled Knowledge-Enhanced Dataset for Medical Visual Question Answering

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.724573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.724573Z digest=sha256:a5429eae591601aee85bbe9c4f3ee80db38cca99a44c15b896f0f40ab9c8a135

Observation 8c7b02f2-f134-4049-86c5-81af31d62f94 · outbound

This paper cites Hasan, Vivek V.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Hasan, Vivek V

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.074082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.728235Z digest=sha256:98cb90d8684deec2d8f06bf23b9890a778b3646cd14f9a90a46e0abbde0ef5f7

Observation c5d2c434-a1a8-4ce4-aaa3-5290bd92de62 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.731580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.731580Z digest=sha256:2338eb708c02077e43f0ac863086c0b7d8336d766eb3b47d5da301fb066ac8f6

Observation 55493f64-ddfc-4e72-b236-94359fc45321 · outbound

This paper cites Weinberger, and Yoav Artzi.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Weinberger, and Yoav Artzi

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:57:23.063149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.735194Z digest=sha256:425dc6ee7bca75b52d4a27ca1ff9a716de306b76837b27383867d7bc2698afc5

Observation cedffaf2-33a3-4aaf-a605-84f17b598694 · outbound

This paper cites Flashattention: Fast and memory-efficient exact attention with io-awareness.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Flashattention: Fast and memory-efficient exact attention with io-awareness

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.742808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.742808Z digest=sha256:e40d59efcdb05c27626941ceb1cb8260f6be9e23dff1aca10360a55d62bea688

Observation d107493a-58e3-434b-8b63-5de0e80d7d93 · outbound

This paper cites Cambrian-1: A fully open, vision-centric exploration of multimodal llms.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Cambrian-1: A fully open, vision-centric exploration of multimodal llms

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.747458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.747458Z digest=sha256:e268fd437d189f273b4b61c369b56d1fdf200fe8a81a8e82948d590e43140e9e

Observation a1a0afb0-d962-481f-9147-dd3b48a2456a · outbound

This paper cites R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.751107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.751107Z digest=sha256:a243b5e67945eb3b6de3b9af2be68a8a755a265fca2560b212817886161c46fd

Observation 04006296-ec90-4526-8845-f0df9fa6fb01 · outbound

This paper cites an unresolved cited work.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints Unresolved cited work

Reference 2020

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:57:23.051792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:57:22.738698Z digest=sha256:6bb914f947f2dcaa35fc111ed8e17e6fb7d6139cbbef79b47604d0c6bd6f35e0

Pith citing papers

Observation a7916707-232e-4e32-99ff-9c13587ee46b · inbound

LLM-as-a-Judge in Healthcare: A Scoping Analysis of Applications, Methods, and Human Alignment cites this paper.

LLM-as-a-Judge in Healthcare: A Scoping Analysis of Applications, Methods, and Human Alignment RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints

Reference 116

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:24:01.590061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T23:18:59.283834Z digest=sha256:5ef75d0c8467b432180661c9807a396053de40a9e531a63b725e173500c65f4a

Observation 03ad372f-6803-41f7-bc9d-8f8783da5a6d · inbound

Improving Heart-Focused Medical Question Answering in LLMs via Variance-Aware Rubric Rewards with GRPO cites this paper.

Improving Heart-Focused Medical Question Answering in LLMs via Variance-Aware Rubric Rewards with GRPO RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T19:39:46.922877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:39:46.922877Z digest=sha256:5a231de62417c8d2a6fb9ca8990e9f2354e41205e12dfa719305f0eb300f7d41