Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T11:17:40.289683Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 0 inbound Pith citation observations for arXiv:2607.19962.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T11:17:40.289683Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
22 of 22 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cb855382-5745-4d62-98f3-66af94cc6b88 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Sketch-of-thought: Efficient LLM rea- soning with adaptive cognitive-inspired sketching
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b0eb710-8d34-4a72-aadd-14e25293efe4 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Optimizing Length Compression in Large Reasoning Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1a0ecba-065e-4ee2-9176-c567615d9f1c · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f07cc2cf-ff5d-46e0-af9f-dc1de20e1401 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85cb031e-917b-4801-a3ca-3083afc79230 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Token- budget-aware llm reasoning
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40955301-8fd3-4990-85b5-e3de72f17479 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Measuring mathemat- ical problem solving with the math dataset
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b2bb410-a8d6-4927-bb03-a392866cb7eb · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization TACO: Topics in Algorithmic COde generation dataset
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda6ec71-ed17-418c-86d5-1055d6e3b994 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fdbc693-ab41-4104-9c81-3107b4d03b6b · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Reasoning Models Can Be Effective Without Thinking
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d2a359f-7229-4e3d-86bb-070bc164659f · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization s1: Simple test-time scaling
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e1b069d-8157-49fc-b506-eadc9e2da592 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization DynaThink: Fast or slow? a dynamic decision-making framework for large language models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 302028a2-a39b-41e6-9985-742e17983424 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Direct preference optimization: Your lan- guage model is secretly a reward model.Advances in neural information processing systems, 36:53728–53741,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23dbda9e-4e0d-4b86-b021-aec00b4b198b · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization DAST: Difficulty-adaptive slow-thinking for large reasoning mod- els
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30e244d6-2465-44ff-871a-4d845bddc279 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Hybridflow: A flexible and efficient rlhf framework
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e7a233f3-2f83-4bfc-bda6-5be9a1c1717c · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Dualformer: Controllable fast and slow thinking by learning with ran- domized reasoning traces
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3ba12e7-a5dc-432a-a7e5-b3dc1b750f5b · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 810116fd-8204-4e41-ae84-f2225e54129a · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3aaf4ff-c0b0-4533-bbfe-1c9a075eb16c · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Qwen2.5 Technical Report
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d0c6014-004c-4ee5-af98-36920e3e0d16 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization ThinkPrune: Pruning Long Chain-of-Thought of LLMs via Reinforcement Learning
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41d58f43-6f89-4885-aa6a-76213e83e346 · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization ThinkSwitcher: When to think hard, when to think fast
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4746e66a-7f39-47a5-98a7-f34fbf85603e · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization The overthinker’s DIET: Cutting token calories with DIfficulty-aware training
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4682635f-18ea-4e05-af52-36b22b0d716f · outbound
EvoThink: Evolving Thinking in Large Reasoning Models via Self-Pruning and Aha-Moment Preference Optimization Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.