Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-10T05:07:40.106099Z
Paper Citation Record · LEDGER
As of 10 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2607.08572.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-10T05:07:40.106099Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
25 of 25 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 75206606-2d14-4ee8-9766-409d162a943f · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Qwen3-VL Technical Report
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation f47b38ad-c95f-4e43-9eb5-56ae2e3c5701 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Ares: Multimodal adaptive reasoning via difficulty-aware token-level entropy shaping
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cb889607-63a3-4f79-8594-f425090b6a7c · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Video-R1: Reinforcing Video Reasoning in MLLMs
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 73c4187c-fa4e-477a-9202-0e1a3dd727f4 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 0596a9b0-fd0d-41ec-9eaa-bc89c8382b7e · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation acad75f2-78db-4485-91aa-9a6d2cca548a · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Boosting MLLM Reasoning with Text-Debiased Hint-GRPO
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6182bec2-599a-4f4a-aaa2-ab55463b8b88 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning LLaVA-OneVision: Easy Visual Task Transfer
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 3a2061e6-be90-4961-b661-0e466d79821e · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 445bfd8a-f823-46ed-a448-fdd75fb93c32 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Visual-RFT: Visual Reinforcement Fine-Tuning
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 60df4eee-bb17-4688-a3ee-162763452667 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning IconQA: A New Benchmark for Abstract Diagram Understanding and Visual Language Reasoning
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 23937e72-3d26-4014-8c13-480738d92280 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning TERMINATOR: Learning Optimal Exit Points for Early Stopping in Chain-of-Thought Reasoning
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 32d2dd63-4520-40eb-93b5-0f6064373c65 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Safegrpo: Self-rewarded multimodal safety alignment via rule-governed policy optimization.arXiv preprint arXiv:2511.12982,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 6f8f9413-c154-4aec-8677-78b768dce8ad · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 85f4df1b-6102-4138-b5ce-8e10c571ab83 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation caec8aa2-2cc6-4a14-8bfc-28456e20a3fa · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation cbdb5770-6e00-4341-98d8-226e05c37194 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning AdaptR1: Reinforcement Learning Based Adaptive Interleaved Thinking in Multi-hop Question Answering
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 5e486c2b-ec80-4531-a2e7-5dde2d114f28 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning When More is Less: Understanding Chain-of-Thought Length in LLMs
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation ede9c4e3-60de-4973-8663-4d65770c8b26 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Dynamic early exit in reasoning models.arXiv preprint arXiv:2504.15895, 2025a
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 172c5265-982d-45bd-88fc-ec64add2b199 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Annealing and Reinforce Learning
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 86302e5a-d931-4bed-a8b0-c36aaa99e3f9 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 03dc2a73-379d-46d3-b1bd-3180698589b6 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning A Survey of Safety on Large Vision-Language Models: Attacks, Defenses and Evaluations
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation 1a8593ca-a459-4bfa-bb0a-a2a8f5dcbca8 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e61c7e6f-ad04-4b8a-a5a2-404dba7ec733 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation dc21e25c-2802-4e4a-977c-f18ea9d45bc3 · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Think in Blocks: Adaptive Reasoning from Direct Response to Deep Reasoning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
Observation e9edec57-32d4-4589-99aa-0d0537938e3c · outbound
Switch-Reasoner: Learn When to Think in Multitask Mixtures via Reinforcement Learning Mathe- matics includes Geometry3k and MathVista; Chart/Doc comprises ChartQA and DocVQA; Ground- ing consists of RefAdv; and the remaining benchmarks are grouped under General
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.
No inbound Pith citation observations are available.