Pith. sign in

Paper Citation Record · LEDGER

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models

As of 18 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2505.13973.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.13973 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:34.049540Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c993af9c-0799-4b55-bc75-1b967876b1d8 · outbound

This paper cites online" 'onlinestring :=.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:26.837843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:26.837843Z digest=sha256:76298ddf06013abebca3a2baaa2e569ebb89bbc947d0020155420130426d0c77

Observation 0086c278-2456-4593-9740-5be1e49a590d · outbound

This paper cites write newline.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.116351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.116351Z digest=sha256:cee43199438d2b80936a6e9efc9fb4d2dd01cb1804d5478be45068b26326339d

Observation d2625a23-56a5-45cf-900f-5be9c59664b9 · outbound

This paper cites Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.245226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.245226Z digest=sha256:32cced85a4b652717f9a199ff7ffbbda78478d11762c69d385aa4200ae1cfda5

Observation e585a3a3-fa6d-4b21-9c28-ec369a8b097a · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:35.830709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T15:43:30.351330Z digest=sha256:bbcb33456f6812b3f3aa1e0fa7c9389f994f36a369460cc284bc6570cf3e9262

Observation eca15f64-1a27-4119-821e-080a268450fd · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.454995Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.454995Z digest=sha256:0cdcfbf737ab7b2c5190361593d812b94c24f86795fe9c6dc4036a1d1205bd5d

Observation 39147750-bcef-4a0f-885f-f886e0ed0be5 · outbound

This paper cites Vision-Language Models Can Self-Improve Reasoning via Reflection.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Vision-Language Models Can Self-Improve Reasoning via Reflection

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.548612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.548612Z digest=sha256:05202c4a340d48781e83b3c05e6699bac7c00cada965d62d91e4c6b9809d3768

Observation 0353c2e0-0333-4069-944a-6ff0ea14d591 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.665481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.665481Z digest=sha256:094f2ff16d20fea39398abd9aea7cc2051d2ec42d8ee492c88395d37238e9d6b

Observation 7c12872c-2522-40a6-970b-18415343905b · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:35.614745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T15:43:30.752803Z digest=sha256:400ac5d070a17c5ddc030dc5de0e1be2b7b30a3e22283a2a99b8a4d362028c1c

Observation 66edf0f7-809f-4720-b248-3123724d02d7 · outbound

This paper cites Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.841430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.841430Z digest=sha256:98845f01166442ec1012e705b6f7914e0f7b6a991ae3c9d4bfae8ff4a597719c

Observation 2a0aab34-6632-4de3-b56f-189c9c036697 · outbound

This paper cites The Llama 3 Herd of Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:30.931668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:30.931668Z digest=sha256:3f2a11b463fa9566e3f94e9180764ea4eeda2b7fde9be662e6afa9aadf0f10a0

Observation 4cdb0920-2151-4686-94cf-331907972674 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.075196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.075196Z digest=sha256:efbf4a89b44b1f65daf30b40003e898658b3251350ed91aebfa7009786f9a5af

Observation ccd43bf7-d09b-4d95-9382-1b4d39545788 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.214749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.214749Z digest=sha256:cf1b09bb720a3b23e41fffef0c412c7c977af5acb2e90051ff8f600764096bfa

Observation 2c09e7ae-55fe-4bb5-bc2d-714a569c4bb0 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.317058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.317058Z digest=sha256:ae3acdfcb9e0ddf2da5b742cd38cce081b6e1b98f86a7beba313c64b43c7d100

Observation 98b9ac47-5b2f-41f1-a935-572a3768597d · outbound

This paper cites Mistral 7B.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Mistral 7B

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.423898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.423898Z digest=sha256:d1e5994cef02ea77a95701fde4f428067df4ee9d118035c92a5c4b43ce7c69ab

Observation 9a7f238e-3dc0-49c4-a70a-ea85c494cf12 · outbound

This paper cites LLM Post-Training: A Deep Dive into Reasoning Large Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models LLM Post-Training: A Deep Dive into Reasoning Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.517756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.517756Z digest=sha256:8e9366fd5c1a28a5e9f37a64904975626d4b75d24fb44dae60743778d62d1466

Observation d3224970-441f-48be-9339-575cec1021d1 · outbound

This paper cites BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.609026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.609026Z digest=sha256:d49941771d09a17f7db27ffec7d146d8da9284f2e09e9e320c302c800b04de15

Observation 3db54ed8-b633-4a5d-8721-1bd9472187d9 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.706196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.706196Z digest=sha256:efa469b3ad27ef23dbc14901927dbf94c83a9d06c7bfd407e65749c4b62ef91e

Observation c59230b9-8a55-4f08-a5e8-635c00080c36 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:35.372647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T15:43:31.790430Z digest=sha256:10ed8caad94b84292946c858cbd3b42ab1e4a5cd4d40a182a7dc77a3084ca853

Observation 2bbb2501-e7b7-4296-9881-50c504fd81a5 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.871537Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.871537Z digest=sha256:f719d20a6b85e2f68a0cb4fc05b607bab12f84698dc1e24e0a36878ad4804fa3

Observation ff045b84-e5de-4954-a5aa-fab2900f7ee1 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Understanding R1-Zero-Like Training: A Critical Perspective

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:31.932837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:31.932837Z digest=sha256:a785605dd0b4a2957a757ca8256dc06fd5d8b107b9e179c8dd59d61cfa6d7826

Observation c37a488c-14a6-4e37-930d-03174f64e814 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:35.154763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T15:43:32.056482Z digest=sha256:9385f992f15d21c47eab9ced407c61321ba15f49fce30fc0c7a047a3c320d524

Observation 85bc92cc-e12d-40b5-bf4b-85aacc88b35b · outbound

This paper cites UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.164373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.164373Z digest=sha256:2028a91188007f337f2cf0041cfa5ec83c80791be3e66a02a02113550d025daf

Observation 10625c65-f303-48e7-9c90-4ced387d8e12 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.291467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.291467Z digest=sha256:0e3e057ab698d655be34ea2265a524478d157c8389658eae03670183fa24531a

Observation cc79f665-816d-45ab-98a2-834e7b962796 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.391411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.391411Z digest=sha256:62a56db7d37ad640689b291059ad729b3126f7cd1602fd6573415821652a8bfc

Observation 8a1e9c79-23c0-470c-9cb4-92c09170525a · outbound

This paper cites Proximal Policy Optimization Algorithms.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Proximal Policy Optimization Algorithms

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.491701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.491701Z digest=sha256:2b22c6d4e3905b5962810f6048b9455bbc420d3c9ceb89cb45d8edb523547387

Observation 2d97a5d0-1e19-4bfa-a1dc-496777c6c5cd · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.625665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.625665Z digest=sha256:cac12e02d3c8d4f39d0cb103ec268a4ce5f24a805e193f25fb76ee8bc97d14f0

Observation b69e7c64-86ed-48bb-8779-7f4c0fb824f7 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:32.721672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:32.721672Z digest=sha256:9341524222ef01b4dfbb2e1726cd5f0ce771e55008afc588c3945ffff311d070

Observation e0ffa0d3-726f-4542-871f-6a653604d65a · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:34.968000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T15:43:32.849962Z digest=sha256:3c324c6eedeff29b055c4928ca945922a53b4346aac05cb7708face8b1deac31

Observation e21db740-c644-4fd7-bdcd-e6fd90737047 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.001856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.001856Z digest=sha256:784a33fa7b208e6426fa41a45b1f4a4faf37e1027cce8bf003456b1a0c14ae6b

Observation 1947b600-f53c-449e-843c-2386d4acd9ef · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.082866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.082866Z digest=sha256:39597193f1c736a19fa1286e6f4d1310767460b791aead8ac16da8c05ec01f9f

Observation 9a6b834a-b81e-450e-8489-5efb40855b7e · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:34.834256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T15:43:33.221448Z digest=sha256:33830a9dfa5a142f8d8d19e85acd52a19845053114cc932bf925003608530964

Observation 9e3019e2-026c-4072-bb84-e23a7b886cd7 · outbound

This paper cites GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.341166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.341166Z digest=sha256:ffa182a04c930ceac364b35b17e3f42b847ce474bce1e539cfb58c2c7d92a434

Observation 06a2df12-52da-4aee-95e4-e9c4a025b307 · outbound

This paper cites Improve Vision Language Model Chain-of-thought Reasoning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Improve Vision Language Model Chain-of-thought Reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.432167Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.432167Z digest=sha256:024a3b5701f6c55a5a45a4d857373bc913b49bdad2c2e709f8a78e6f4b99d3c3

Observation 45eb7a77-1544-45be-9be2-8045541f163e · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.540806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.540806Z digest=sha256:b504e225fe13c31567cb34c1f771a12d8106821a5013bee9affcef1fd9eeeeb4

Observation 66a8f6d7-957d-4702-89d6-8ebcfda7b726 · outbound

This paper cites PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models PMC-VQA: Visual Instruction Tuning for Medical Visual Question Answering

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.648501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.648501Z digest=sha256:09f14e4f05305eca161dfa33166ef3f4470b750219cae0412213fb583ac916fe

Observation e46c66b0-777a-4121-bda9-b6f6f8947210 · outbound

This paper cites R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.729004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.729004Z digest=sha256:d31f5b814da93b36c991cbdb0b7e24c775b917e65e9ea0c0b1be06770816827b

Observation 4591e73e-4ce8-40f9-ae02-e5d880d8d484 · outbound

This paper cites an unresolved cited work.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:43:34.611405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-07T15:43:33.787424Z digest=sha256:eb5852ceef9cdba274abdb1622870e6e56d6c1ee304af8546587bcfc0ece1c98

Observation 52a61e03-e4ca-4a96-86a6-fee5261db3a2 · outbound

This paper cites R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models R1-Zero's "Aha Moment" in Visual Reasoning on a 2B Non-SFT Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:33.968558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:33.968558Z digest=sha256:018302e1e87064cc9ec145832f7ed55c2d3f66ffbc74096f215521f35023e0b2

Observation 247b3255-68d8-431b-8ceb-1ecf28d49abd · outbound

This paper cites RetinalGPT: A Retinal Clinical Preference Conversational Assistant Powered by Large Vision-Language Models.

Toward Effective Reinforcement Learning Fine-Tuning for Medical VQA in Vision-Language Models RetinalGPT: A Retinal Clinical Preference Conversational Assistant Powered by Large Vision-Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:34.049540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:43:34.049540Z digest=sha256:d8bbc73905ee922aeb69f12b347f730fde8c22b69c0b5c6557ba32ea1dddf71b

Pith citing papers

No inbound Pith citation observations are available.