Pith. sign in

Paper Citation Record · LEDGER

MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 42 inbound Pith citation observations for arXiv:2502.19634.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.19634 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 42 of 42 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:41:37.084607Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:39:30.318199Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 08026dea-3f74-48ef-8d4e-966a96f8e7c6 · inbound

Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning cites this paper.

Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-23T03:37:27.808901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T03:35:56.594095Z digest=sha256:23df799f759fc9aa6b52016dd09210d4c4fd48129f777d2f759eb1199fd641e8

Observation c610166b-951f-4814-b62a-e0dd348f9059 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 258

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:36:24.450656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:ea4077fade8214e231b6047d22143b1e8a97dc7c52c06030aa21fcaf288b57a7

Observation f3a791f7-a624-48ac-8ffd-269dbf708893 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 154

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T17:18:53.259697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:362e7bfae24e5e3a10e1e2e7163a5b808f2b41de7a88e5460dd1ca614332e25d

Observation c49900b3-024a-4617-b6d4-97c206c7ac88 · inbound

FractalMamba++: Scaling Vision Mamba Across Resolutions via Hilbert Fractal Geometry cites this paper.

FractalMamba++: Scaling Vision Mamba Across Resolutions via Hilbert Fractal Geometry MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T14:01:38.168898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T13:58:28.333444Z digest=sha256:8426bab8a1c846f06c0ba1982c144e3970406905f29475624d8bdb66f3210774

Observation f99ccc4a-99b6-43ff-8516-31e3ba326ad1 · inbound

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL cites this paper.

Beyond Distillation: Pushing the Limits of Medical LLM Reasoning with Minimalist Rule-Based RL MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:41:37.084607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:41:37.084607Z digest=sha256:cb4fcbcdec12911543ba1cc4a1e5d7618ee95f4573775cbe4539f5835bd08a86

Observation 71364c43-cfd6-4c3d-b3d9-5d776cd13b8f · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 112

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:18.761499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:18.761499Z digest=sha256:dd9b302661a3d6573f68bd355c110061c33dcd31fa59eb6c203f36cf0e0d6e2e

Observation 16f03c70-230e-4852-847a-cf1b281a323c · inbound

Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning cites this paper.

Improving Medical Reasoning with Curriculum-Aware Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:42.458642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:24:42.458642Z digest=sha256:36217d6a37a54157b866b3b0ad50205931a934a4a2807ef0f94dafca601073ab

Observation 9fd95f8c-1951-43c2-b50c-9516296cf75d · inbound

Real-World Doctor Agent with Proactive Consultation through Multi-Agent Reinforcement Learning cites this paper.

Real-World Doctor Agent with Proactive Consultation through Multi-Agent Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T13:42:19.273147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T13:41:30.833840Z digest=sha256:0256524b9d8e671fdc138a11976e0835387b677144396e2e135a4742ea554f40

Observation 5e90f5b5-2338-45ef-b97c-806ceb17304c · inbound

Collision- and Reachability-Aware Multi-Robot Control with Grounded LLM Planners cites this paper.

Collision- and Reachability-Aware Multi-Robot Control with Grounded LLM Planners MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:57:23.945486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:57:23.945486Z digest=sha256:6dd85572ea04e6d6cd38dda2ccb29ce6645ec2e41d3e6178dc69372480052afd

Observation ea8fe538-4e8d-4330-8df7-7642bab30f10 · inbound

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning cites this paper.

Training LLMs for EHR-Based Reasoning Tasks via Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:40:10.729078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:40:10.729078Z digest=sha256:5a7fd55dec6b213ade912388e89cad28d599e19f4623f71bdd0eab39df8147d2

Observation 09bbc576-4bbc-4fe1-b9a6-f2bd2e90c95e · inbound

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book cites this paper.

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:04.483839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:04.483839Z digest=sha256:9ef8d50c9494c2b322cda176dfad3f3e485c220f93b8bad9ac08b27039879550

Observation c9cc2324-a571-4895-822c-83e6c158173a · inbound

Q-Ponder: A Unified Training Pipeline for Reasoning-based Visual Quality Assessment cites this paper.

Q-Ponder: A Unified Training Pipeline for Reasoning-based Visual Quality Assessment MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:22:46.567623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:22:46.567623Z digest=sha256:a3469a875188010399e4a204e1982fdc2ded7c63b03f4a1d232046d345ec8a02

Observation 1d4cbf67-d28b-4623-8c4a-68fa5d80421c · inbound

HeartcareGPT: A Unified Multimodal ECG Suite for Dual Signal-Image Modeling and Understanding cites this paper.

HeartcareGPT: A Unified Multimodal ECG Suite for Dual Signal-Image Modeling and Understanding MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T10:47:15.084237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T10:44:01.880405Z digest=sha256:a38da6770a9539e6fff92f8585cf352d7532fd557f6e2a2449d0328cf14af65c

Observation 6349b17c-70c1-4041-a3ed-97f290a22c00 · inbound

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints cites this paper.

RARL: Improving Medical VLM Reasoning and Generalization with Reinforcement Learning and LoRA under Data and Hardware Constraints MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:57:22.684656Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:57:22.684656Z digest=sha256:bfb608d237223b10ebbdff76baa7753ac302ccf1917e1777e267bd8cb7c00c3c

Observation d1bb6f39-0c47-464b-afce-d848a1151154 · inbound

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning cites this paper.

Vision-EKIPL: External Knowledge-Infused Policy Learning for Visual Reasoning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T10:37:14.982596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T10:34:48.849524Z digest=sha256:8b6ba45e9e86a6d52092bd5b5974207a2056106295f4ab98dc312df89fc1c324

Observation 04b3b00c-b072-4a79-9734-767ad16a5a2c · inbound

Interpretable and Reliable Detection of AI-Generated Images via Grounded Reasoning in MLLMs cites this paper.

Interpretable and Reliable Detection of AI-Generated Images via Grounded Reasoning in MLLMs MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:48:51.106955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:48:51.106955Z digest=sha256:77b2f0cdf94f01dd60c06739dae9dfa1d958b98fe8a1ca0b3ffc3edd15c4da2d

Observation 1e46f3ab-1a81-4f92-9c2a-d227c6a7c6c6 · inbound

ComfyUI-R1: Exploring Reasoning Models for Workflow Generation cites this paper.

ComfyUI-R1: Exploring Reasoning Models for Workflow Generation MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T04:45:36.557572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:45:36.557572Z digest=sha256:c10c2df6a8128a76bb0e2c08781034a48cf046c2f5a97934d2d9a418202e8592

Observation 0148c9c3-ae20-46f6-a515-aafe76cc86d6 · inbound

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought cites this paper.

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:33:02.077129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-19T08:32:20.566798Z digest=sha256:194fe29f4bc2b1aae15d1bcac9874e4ae84d9908c587e5e842257c2f0fdb4644

Observation dbde38d3-4b95-4ad0-9163-63e5af4152cf · inbound

Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning cites this paper.

Unleashing Embodied Task Planning Ability in LLMs via Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T21:55:02.328497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:55:02.328497Z digest=sha256:4f07444fa9dd19f5937b15806da483ef34c7f40e702a33ba74436921172218da

Observation 9c0eb45d-968e-4b77-abab-ae40351bd9f6 · inbound

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study cites this paper.

How Far Have Medical Vision-Language Models Come? A Comprehensive Benchmarking Study MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T17:19:46.982626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:19:46.982626Z digest=sha256:f4ccea583c76253ba43bccf396e1896bff26ca6605edf88f82263646092342fc

Observation e06e925f-1684-4ceb-85d4-210433870bd8 · inbound

Addressing Benchmarking Gaps in Large Language Models for Health and Medicine with Dynamic Red-Teaming cites this paper.

Addressing Benchmarking Gaps in Large Language Models for Health and Medicine with Dynamic Red-Teaming MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T11:45:41.058150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:45:41.058150Z digest=sha256:d282cc808b8dfe5216288014a28a3d7c4338b058abba2b9aeeeba1bfa9de77db

Observation 7d1ce7c5-0092-47fe-974a-428a5f9be061 · inbound

The Philosophy and Physics of Duality cites this paper.

The Philosophy and Physics of Duality MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T05:32:07.529497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:32:07.529497Z digest=sha256:3fce1f871c9584d04399b7960f65979203e9e53fffe23ad2f9134149f4c16214

Observation bb4a5f50-d6f3-4315-978e-cbdbfe02f87b · inbound

CX-Mind: A Pioneering Multimodal Large Language Model for Interleaved Reasoning in Chest X-ray via Curriculum-Guided Reinforcement Learning cites this paper.

CX-Mind: A Pioneering Multimodal Large Language Model for Interleaved Reasoning in Chest X-ray via Curriculum-Guided Reinforcement Learning MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T11:00:49.732799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:00:49.732799Z digest=sha256:979fd6772a6571e66d8dc3f50f1fbce354f3f291ab608c08fc0ac3db7dcb7c8a

Observation e3e3ac75-b9d8-45bd-b950-8994d508b0df · inbound

MedFoundationHub: A Lightweight and Secure Toolkit for Deploying Medical Vision Language Foundation Models cites this paper.

MedFoundationHub: A Lightweight and Secure Toolkit for Deploying Medical Vision Language Foundation Models MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:09:18.778484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:09:18.778484Z digest=sha256:c714641775fabb10406b6bc69659476d46ba5a3119368bd53f074ffb68fa73a8

Observation 791e385e-640a-49d3-85db-aebc084a5437 · inbound

Towards Better Dental AI: A Multimodal Benchmark and Instruction Dataset for Panoramic X-ray Analysis cites this paper.

Towards Better Dental AI: A Multimodal Benchmark and Instruction Dataset for Panoramic X-ray Analysis MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T19:27:40.746625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T19:27:40.746625Z digest=sha256:7343df869b7b196005c467aa36f20a5e6709173c09a2b342305fb6b99c89c73a

Observation 61857e1c-7081-479a-a066-a429f9b75eb3 · inbound

Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images cites this paper.

Locate-Then-Examine: Grounded Region Reasoning Improves Detection of AI-Generated Images MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T10:16:14.423376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-18T10:13:08.112072Z digest=sha256:3d8988ebadafd54531c2e4ebc18e2725c5712494c3ce119598e597e5db76c403

Observation 54b3cfac-6a56-4408-b77d-2953c92349ba · inbound

FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound cites this paper.

FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T20:08:23.042354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T20:05:27.320122Z digest=sha256:fb7c8fedd138f7ca8357cb5380201e32f97b9ef5ed4f94f1efeb976238e6d84c

Observation a120570f-f138-467d-b61e-fa89357efa06 · inbound

Weather-R1: Logically Consistent Reinforcement Fine-Tuning for Multimodal Reasoning in Meteorology cites this paper.

Weather-R1: Logically Consistent Reinforcement Fine-Tuning for Multimodal Reasoning in Meteorology MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T12:37:53.025564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T12:36:23.661583Z digest=sha256:9e1a347a196de3e4fa1f1397838f7bd0ad049da7c038c7a2da7b2c3dfe247e2b

Observation b0ee57f9-fc5b-4c1c-b1ab-46d0f6de3f2f · inbound

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space cites this paper.

Beyond Medical Diagnostics: How Medical Multimodal Large Language Models Think in Space MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T18:16:32.946219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:16:32.946219Z digest=sha256:8d8b8007e84c7920df1aecd65e262a1db7917a8d3bcc6d9e7eac040c4a61e60a

Observation e1bb637b-14cf-4b5d-9987-4d92f921f90b · inbound

MedRCube: A Multidimensional Framework for Fine-Grained and In-Depth Evaluation of MLLMs in Medical Imaging cites this paper.

MedRCube: A Multidimensional Framework for Fine-Grained and In-Depth Evaluation of MLLMs in Medical Imaging MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-10T12:50:25.199996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T12:47:27.551670Z digest=sha256:5c4c008f4b3e899d2c3246bd47cc1ecdf16655f88ec2dbf36ee7e31503dad692

Observation f6b2504d-bff0-49fa-9020-f58f474016e6 · inbound

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems cites this paper.

Navigating the Clutter: Waypoint-Based Bi-Level Planning for Multi-Robot Systems MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 64

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:06:05.434288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-09T23:27:54.704794Z digest=sha256:ca95583e9772162a5e492f4ff412fb6daaede51959a58459e93453719faa47e9

Observation 4c0e7059-6ff1-40b2-bc9b-faf01211a785 · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.647995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:ff0d9838e28d3844a987e75651fe4bcf37bc63c57cf39ecefb817674eeb954da

Observation d6a3139c-1b42-4e71-8d77-74e597963c95 · inbound

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding cites this paper.

From Failure to Feedback: Group Revision Unlocks Hard Cases in Object-Level Grounding MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 59

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T18:43:38.841158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T18:39:11.904941Z digest=sha256:ff3280f874676fb0f0b8d72fa4f1a452fd90a806fc94f86c54a7b471bae35049

Observation 6d0cdeff-5f88-43bb-af6a-401979889ef3 · inbound

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming cites this paper.

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T09:14:45.421856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T09:13:45.830192Z digest=sha256:d42ee60e60975057ddf8cf493702a34ed2e6828d759419c277e2bb8907dfdf50

Observation a4957f2e-dcab-423e-b975-ee2863c70602 · inbound

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming cites this paper.

Look-Closer-Then-Diagnose: Confidence-Aware Ultrasound VQA via Active Zooming MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T16:54:58.805483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T16:48:55.898188Z digest=sha256:3d82f02ac90441e19769ea61ee328eaa87b92e73f90dfadc3f55beff14b06a0b

Observation 52ba3a10-6510-4da7-8117-e8f853cc6d58 · inbound

Mechanistically Interpreting the Role of Sample Difficulty in RLVR for LLMs cites this paper.

Mechanistically Interpreting the Role of Sample Difficulty in RLVR for LLMs MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T12:13:26.541175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T12:11:00.402276Z digest=sha256:488f35f2a87c9d43eb1fc315475b60485c2fb130fb545a5e749c8f8948810ec4

Observation 1f94e389-9146-4977-a77c-0cc35b798a99 · inbound

Does Language Shift Break Medical Vision-Language Models? Indonesian Radiology Visual Question Answering Case Study cites this paper.

Does Language Shift Break Medical Vision-Language Models? Indonesian Radiology Visual Question Answering Case Study MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:46:28.064890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T10:42:50.961855Z digest=sha256:c18bc79c3bac1efeefa9538921c22de8c7ce622017065b469b9411e4c40df084

Observation b071d916-9eb2-4dcf-bfdc-0ba60644be2d · inbound

PlanBench-V: A Spatial Planning Map Benchmark for Vision-Language Models cites this paper.

PlanBench-V: A Spatial Planning Map Benchmark for Vision-Language Models MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:26:56.321942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T02:10:12.974244Z digest=sha256:68dfd10a6bc08b48c6e90e41693d73046eb0c292cd8105682a93863cb06bed11

Observation 531ca0e1-481c-4e74-ab98-55e08e2cc31f · inbound

DIYHealth Suite: Dataset, Model, and Benchmark for Health Management at Home cites this paper.

DIYHealth Suite: Dataset, Model, and Benchmark for Health Management at Home MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 85

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T08:05:31.550453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T07:40:03.204845Z digest=sha256:99764fa223793b9951e5ad20c70746ec31a8593bb0d633979255c4902d36ca9a

Observation 1553cc7f-5a39-4e8a-a21c-37dcf4a3e62f · inbound

Confidence Calibration for Multimodal LLMs: An Empirical Study through Medical VQA cites this paper.

Confidence Calibration for Multimodal LLMs: An Empirical Study through Medical VQA MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:39:30.320275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T17:49:38.151224Z digest=sha256:fd88afa4be383c7fa6bfcaa23bfa6b50a39f770148b33c4ae69bfe2374bb5776

Observation f8263665-5bc3-474a-beb1-b78d698eaf4c · inbound

EchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in Echocardiography cites this paper.

EchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in Echocardiography MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T16:55:51.511954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T04:22:07.273815Z digest=sha256:717fd81e75ebb28eaaf840350fa16fd8f830cf625fd28891c16fff540d9eee30

Observation 27b95a77-2560-4073-9f5d-e33643a84344 · inbound

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA cites this paper.

Objective-Aligned Direct Answer SFT for Robust Multi-Frame Medical VQA MedVLM-R1: Incentivizing Medical Reasoning Capability of Vision-Language Models (VLMs) via Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T05:36:46.863899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T05:36:46.863899Z digest=sha256:099d701f4c39809cd0d395f09093191d9adc324963bb529e684708f6b95b15af