Pith. sign in

Paper Citation Record · LEDGER

MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 42 inbound Pith citation observations for arXiv:2502.19634.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.19634 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 42 of 42 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:41:37.084607Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:39:30.318199Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 08026dea-3f74-48ef-8d4e-966a96f8e7c6 · inbound

Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning cites this paper.

Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:37:27.808901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T03:35:56.594095Z digest=sha256:3ffb158a96f528647d0bafde6c86c6df1f287c7c0b7b5059370b3b8647dc4020

Observation c610166b-951f-4814-b62a-e0dd348f9059 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 258

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:36:24.450656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:40fa57a09a11ba0d67d3d602824ca9e70559afba731a915a21a31ffa9450f172

Observation f3a791f7-a624-48ac-8ffd-269dbf708893 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 154

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T17:18:53.259697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:729a206875be7f9b50bdac5da2a65042b419fd1bed2979d6e4209bba506f9e1d

Observation c49900b3-024a-4617-b6d4-97c206c7ac88 · inbound

FractalMamba++: Scaling Vision Mamba Across Resolutions via Hilbert Fractal Geometry cites this paper.

FractalMamba++: Scaling Vision Mamba Across Resolutions via Hilbert Fractal Geometry MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T14:01:38.168898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T13:58:28.333444Z digest=sha256:6218a09fb36cb255dcf24f07ab4f769309305e44162b813816fa295de4152b8e

Observation f99ccc4a-99b6-43ff-8516-31e3ba326ad1 · inbound

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL cites this paper.

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:37.084607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:37.084607Z digest=sha256:cb4fcbcdec12911543ba1cc4a1e5d7618ee95f4573775cbe4539f5835bd08a86

Observation 71364c43-cfd6-4c3d-b3d9-5d776cd13b8f · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:18.761499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:18.761499Z digest=sha256:dd9b302661a3d6573f68bd355c110061c33dcd31fa59eb6c203f36cf0e0d6e2e

Observation 16f03c70-230e-4852-847a-cf1b281a323c · inbound

Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning cites this paper.

Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:42.458642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:42.458642Z digest=sha256:36217d6a37a54157b866b3b0ad50205931a934a4a2807ef0f94dafca601073ab

Observation 9fd95f8c-1951-43c2-b50c-9516296cf75d · inbound

Real-World Doctor Agent with Proactive Consultation through Multi-Agent Reinforcement Learning cites this paper.

Real-World Doctor Agent with Proactive Consultation through Multi-Agent Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T13:42:19.273147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T13:41:30.833840Z digest=sha256:7ac520797baf0b1beb6a048843e26ef73b5e2eef2d545ce43dd93d22a63deb06

Observation 5e90f5b5-2338-45ef-b97c-806ceb17304c · inbound

Collision- and Reachability-Aware Multi-Robot Control with Grounded LLM Planners cites this paper.

Collision- and Reachability-Aware Multi-Robot Control with Grounded LLM Planners MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:23.945486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:23.945486Z digest=sha256:6dd85572ea04e6d6cd38dda2ccb29ce6645ec2e41d3e6178dc69372480052afd

Observation ea8fe538-4e8d-4330-8df7-7642bab30f10 · inbound

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning cites this paper.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:10.729078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:10.729078Z digest=sha256:16deb14502da47ce185cbfa874848ac1ff1c03f20e73ab612f84071b339aa917

Observation 09bbc576-4bbc-4fe1-b9a6-f2bd2e90c95e · inbound

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book cites this paper.

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:04.483839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:04.483839Z digest=sha256:1017b07bf6f80c86d39abd5f1a3ced5a62ccda37d32108c3e920d3885a912697

Observation c9cc2324-a571-4895-822c-83e6c158173a · inbound

Q-Ponder: A Unified Training Pipeline for Reasoning-based Visual Quality Assessment cites this paper.

Q-Ponder: A Unified Training Pipeline for Reasoning-based Visual Quality Assessment MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:22:46.567623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:22:46.567623Z digest=sha256:a3469a875188010399e4a204e1982fdc2ded7c63b03f4a1d232046d345ec8a02

Observation 1d4cbf67-d28b-4623-8c4a-68fa5d80421c · inbound

HeartcareGPT: A Unified Multimodal ECG Suite for Dual Signal-Image Modeling and Understanding cites this paper.

HeartcareGPT: A Unified Multimodal ECG Suite for Dual Signal-Image Modeling and Understanding MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T10:47:15.084237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T10:44:01.880405Z digest=sha256:8f09fd793906f27314cf7ab026a09ed3f560c9a6ec70d2c0f989fd63eba14ad5

Observation 6349b17c-70c1-4041-a3ed-97f290a22c00 · inbound

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints cites this paper.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.684656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.684656Z digest=sha256:41e58d2390d86e6a28d4aef22539f4f16ae7b6d4548c6d96f0bc1f3cd4daa654

Observation d1bb6f39-0c47-464b-afce-d848a1151154 · inbound

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning cites this paper.

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T10:37:14.982596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T10:34:48.849524Z digest=sha256:c5cd3a98158953cc84d0838e30f0332fe25de6a0238b6bc7bc262d61f5efc9d2

Observation 04b3b00c-b072-4a79-9734-767ad16a5a2c · inbound

Interpretable and Reliable Detection of AI-Generated Images via Grounded Reasoning in MLLMs cites this paper.

Interpretable and Reliable Detection of AI-Generated Images via Grounded Reasoning in MLLMs MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:51.106955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:51.106955Z digest=sha256:3b916eb2f42727e7a19392f7999c88ddf61c37961a9979b24fa4003094cfeacc

Observation 1e46f3ab-1a81-4f92-9c2a-d227c6a7c6c6 · inbound

ComfyUI-R1: Exploring Reasoning Models for Workflow Generation cites this paper.

ComfyUI-R1: Exploring Reasoning Models for Workflow Generation MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T04:45:36.557572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:45:36.557572Z digest=sha256:e571a15c5d239ef0caf6b6f0f8a6122af38b35491bda6215e2e87ce7a275d483

Observation 0148c9c3-ae20-46f6-a515-aafe76cc86d6 · inbound

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought cites this paper.

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:33:02.077129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T08:32:20.566798Z digest=sha256:cebac9ff3794d698fbba18805553de756161911b69f135032e19a7dd1a9b91a3

Observation dbde38d3-4b95-4ad0-9163-63e5af4152cf · inbound

Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning cites this paper.

Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:02.328497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:55:02.328497Z digest=sha256:910aab9a828c817eb54c071dfa71d305f9013f77f453296242ff2bf5940687d1

Observation 9c0eb45d-968e-4b77-abab-ae40351bd9f6 · inbound

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study cites this paper.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.982626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.982626Z digest=sha256:45f3f1b45c8ed670e4bd85d25779e94d0bde1c1079b2f529cd5912bfff393d19

Observation e06e925f-1684-4ceb-85d4-210433870bd8 · inbound

Addressing Benchmarking Gaps in Large Language Models for Health and Medicine with Dynamic Red-Teaming cites this paper.

Addressing Benchmarking Gaps in Large Language Models for Health and Medicine with Dynamic Red-Teaming MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T11:45:41.058150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:45:41.058150Z digest=sha256:d282cc808b8dfe5216288014a28a3d7c4338b058abba2b9aeeeba1bfa9de77db

Observation 7d1ce7c5-0092-47fe-974a-428a5f9be061 · inbound

The Philosophy and Physics of Duality cites this paper.

The Philosophy and Physics of Duality MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T05:32:07.529497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:32:07.529497Z digest=sha256:f649b4684a3596651215e4ce31cef2ad0d5f13ba6b515c77e8c73dbab838b4f6

Observation bb4a5f50-d6f3-4315-978e-cbdbfe02f87b · inbound

CX-Mind: A Pioneering Multimodal Large Language Model for Interleaved Reasoning in Chest X-ray via Curriculum-Guided Reinforcement Learning cites this paper.

CX-Mind: A Pioneering Multimodal Large Language Model for Interleaved Reasoning in Chest X-ray via Curriculum-Guided Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T11:00:49.732799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:00:49.732799Z digest=sha256:979fd6772a6571e66d8dc3f50f1fbce354f3f291ab608c08fc0ac3db7dcb7c8a

Observation e3e3ac75-b9d8-45bd-b950-8994d508b0df · inbound

MedFoundationHub: A Lightweight and Secure Toolkit for Deploying Medical Vision Language Foundation Models cites this paper.

MedFoundationHub: A Lightweight and Secure Toolkit for Deploying Medical Vision Language Foundation Models MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:09:18.778484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:09:18.778484Z digest=sha256:c714641775fabb10406b6bc69659476d46ba5a3119368bd53f074ffb68fa73a8

Observation 791e385e-640a-49d3-85db-aebc084a5437 · inbound

Towards Better Dental AI: A Multimodal Benchmark and Instruction Dataset for Panoramic X-ray Analysis cites this paper.

Towards Better Dental AI: A Multimodal Benchmark and Instruction Dataset for Panoramic X-ray Analysis MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T19:27:40.746625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:27:40.746625Z digest=sha256:863204daafe1b800e49625a07f752f12f823df99b2b897248fbeeda15fbc731b

Observation 61857e1c-7081-479a-a066-a429f9b75eb3 · inbound

Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images cites this paper.

Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:16:14.423376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T10:13:08.112072Z digest=sha256:0fdbb03eb79fa6ce45e1027f50130dc823c314318c89c1764e73a8162f60a992

Observation 54b3cfac-6a56-4408-b77d-2953c92349ba · inbound

FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound cites this paper.

FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:08:23.042354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T20:05:27.320122Z digest=sha256:f1d08b55d4eb7e6a51ecc06bc4ea566dd5428c2999efe5038251f1df7e2a54a4

Observation a120570f-f138-467d-b61e-fa89357efa06 · inbound

Weather-R1: Logically Consistent Reinforcement Fine-Tuning for Multimodal Reasoning in Meteorology cites this paper.

Weather-R1: Logically Consistent Reinforcement Fine-Tuning for Multimodal Reasoning in Meteorology MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:37:53.025564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T12:36:23.661583Z digest=sha256:d6f8f36296619db1c76885ee1a9f0401a6637cdaa3146e6e550d9e5bd0677e7f

Observation b0ee57f9-fc5b-4c1c-b1ab-46d0f6de3f2f · inbound

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space cites this paper.

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T18:16:32.946219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:16:32.946219Z digest=sha256:7bdb49efe5f3e0bf4f1f9c863a8af532ace2ad62c2d2399bba235b86195e53fd

Observation e1bb637b-14cf-4b5d-9987-4d92f921f90b · inbound

MedRCube: A Multidimensional Framework for Fine-Grained and In-Depth Evaluation of MLLMs in Medical Imaging cites this paper.

MedRCube: A Multidimensional Framework for Fine-Grained and In-Depth Evaluation of MLLMs in Medical Imaging MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:50:25.199996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-10T12:47:27.551670Z digest=sha256:340a4aa632de07fa775c5d1868908b5bc9d9e29047ba140d1e86da52f76848ee

Observation f6b2504d-bff0-49fa-9020-f58f474016e6 · inbound

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems cites this paper.

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:06:05.434288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-09T23:27:54.704794Z digest=sha256:1a76589cb290a7530cb3f2493b06155e2b7633b05b1aa011017e6dcbae6fa0e2

Observation 4c0e7059-6ff1-40b2-bc9b-faf01211a785 · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.647995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:6f430512b34dcaef3ac119098876bd90df1602d466fca5f379bb99c0f83a5e59

Observation d6a3139c-1b42-4e71-8d77-74e597963c95 · inbound

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding cites this paper.

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T18:43:38.841158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T18:39:11.904941Z digest=sha256:8daee1f4dab340096b81e645488fec66e56d325f23b8308df79d9872b63e6ff5

Observation 6d0cdeff-5f88-43bb-af6a-401979889ef3 · inbound

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming cites this paper.

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T09:14:45.421856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T09:13:45.830192Z digest=sha256:475b6a5e0261f7a2a9932938902b0d7b970a157208a2ba4dda7f49bb8d2f669e

Observation a4957f2e-dcab-423e-b975-ee2863c70602 · inbound

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming cites this paper.

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T16:54:58.805483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T16:48:55.898188Z digest=sha256:4ba3f706eb76310b94e5b193086facb09de361bdced1d7e9fcb747a5b2989fd5

Observation 52ba3a10-6510-4da7-8117-e8f853cc6d58 · inbound

Mechanistically Interpreting the Role of Sample Difficulty in RLVR for LLMs cites this paper.

Mechanistically Interpreting the Role of Sample Difficulty in RLVR for LLMs MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:13:26.541175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T12:11:00.402276Z digest=sha256:2e41286b41dddbff7c3a40f8d8c512f12b1e13cc58f892d930560e9f195f1637

Observation 1f94e389-9146-4977-a77c-0cc35b798a99 · inbound

Does Language Shift Break Medical Vision-Language Models? Indonesian Radiology Visual Question Answering Case Study cites this paper.

Does Language Shift Break Medical Vision-Language Models? Indonesian Radiology Visual Question Answering Case Study MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:46:28.064890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T10:42:50.961855Z digest=sha256:bc6866b7503f4a64e3d6979078e02176e6034fd0125a84a10b43c6af8f5a1bb3

Observation b071d916-9eb2-4dcf-bfdc-0ba60644be2d · inbound

PlanBench-V: A Spatial Planning Map Benchmark for Vision-Language Models cites this paper.

PlanBench-V: A Spatial Planning Map Benchmark for Vision-Language Models MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:26:56.321942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T02:10:12.974244Z digest=sha256:3cff8bfb506f8cb1beea326d0fad9cc7bf5a0b2ad13c0557f8b23257b24d5e92

Observation 531ca0e1-481c-4e74-ab98-55e08e2cc31f · inbound

DIYHealth Suite: Dataset, Model, and Benchmark for Health Management at Home cites this paper.

DIYHealth Suite: Dataset, Model, and Benchmark for Health Management at Home MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:05:31.550453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-01T07:40:03.204845Z digest=sha256:441d0c7b416b88ef8674fda28982e7c8416d756906191ce57817f78ead27cc64

Observation 1553cc7f-5a39-4e8a-a21c-37dcf4a3e62f · inbound

Confidence Calibration for Multimodal LLMs: An Empirical Study through Medical VQA cites this paper.

Confidence Calibration for Multimodal LLMs: An Empirical Study through Medical VQA MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:39:30.320275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:49:38.151224Z digest=sha256:be184051c5186d3a51146574525f2d2b826d832942c5b65bbbf4dc61c2ccfe67

Observation f8263665-5bc3-474a-beb1-b78d698eaf4c · inbound

EchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in Echocardiography cites this paper.

EchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in Echocardiography MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T16:55:51.511954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T04:22:07.273815Z digest=sha256:511eafdb57dc0f13cc4504dc9ccdaba0f0a6cada1f129047e4ca5436cf2304fa

Observation 27b95a77-2560-4073-9f5d-e33643a84344 · inbound

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA cites this paper.

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T05:36:46.863899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:36:46.863899Z digest=sha256:099d701f4c39809cd0d395f09093191d9adc324963bb529e684708f6b95b15af