Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-01T06:03:12.182207Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 1 inbound Pith citation observation for arXiv:2606.31825.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-01T06:03:12.182207Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-08T05:22:50.462331Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-08T05:24:31.802601Z
13 of 13 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 391de103-519a-486b-b9ea-c60ed180301b · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning arXiv preprint arXiv:2503.13939 , year=
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bef98d89-4b53-4285-9591-c3ef4a30bfed · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Quilt-LLaVA: Visual Instruction Tuning by Extracting Localized Narratives from Open-Source Histopathology Videos
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1d1d2a86-09f8-4388-80bf-ea728822a117 · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Gtpo and grpo-s: Token and sequence-level reward shaping with policy entropy.arXiv preprint arXiv:2508.04349
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 847f69e5-aa7a-4c5d-9380-c41f42571d61 · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning arXiv preprint arXiv:2506.13793 (2025)
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4c54032d-1369-4b42-9250-a271c57e6cd2 · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Medgr ^ 2 : Breaking the data barrier for medical reasoning via generative reward learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8f5369d4-37b6-40b6-973f-e20775bf7b0c · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Our imple- mentation builds on VLM-R1 (Shen et al., 2025), an open-source GRPO framework for VLMs
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 880a61e7-868c-4a23-ab28-0d2b83701c18 · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning consistent with
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f1b47ad8-bb2a-49e2-911e-0153f571c989 · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Step2 It has a deep reddish-brown color, a lobulated shape, and a granular, nodular cut surface
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ea312527-37a9-4c37-8b32-21d584feb6f5 · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1a381f95-1bc9-498a-890a-ef5224ae6e2b · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Step2 The dark regions in and around the liver look like fluid, suggesting a fluid- sensitive T2 sequence
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9d09fcc0-7c04-4877-94b4-a857eddab13d · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6d724102-70fa-4c29-b98e-7a25e18d61aa · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning lobulated,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 435e1e5a-80ec-4845-b10a-d888ae7869d8 · outbound
Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning lobulated
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e32c882c-8e53-4c7e-a5b8-867115386abb · inbound
From Voting to Agent Collaboration: Answer-Type-Aware LLM Pipelines for BioASQ 14b Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.