Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:20:19.609052Z
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 36 inbound Pith citation observations for arXiv:2505.13379.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:20:19.609052Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T15:50:30.535082Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T13:29:51.951052Z
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 3549abb0-1ed0-4506-9c89-dc08d8517cc0 · outbound
Thinkless: LLM Learns When to Think L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0739fbfa-8df9-43d8-90b8-5a1a695d054a · outbound
Thinkless: LLM Learns When to Think Claude 3.7 Sonnet
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 63f97f2b-fe1f-4da9-be21-f0b85cab4747 · outbound
Thinkless: LLM Learns When to Think Sketch-of-thought: Efficient llm reasoning with adaptive cognitive-inspired sketching
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afb52ade-f5d9-4775-823b-7f24a0074784 · outbound
Thinkless: LLM Learns When to Think Llama-nemotron: Efficient reasoning models, 2025
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation aae95e0e-c178-4fe7-8340-dd6aa7203a03 · outbound
Thinkless: LLM Learns When to Think ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07575248-a8f2-4122-9cd9-3cdb2ccaaff7 · outbound
Thinkless: LLM Learns When to Think Distilling Reasoning Ability from Large Language Models with Adaptive Thinking
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16383140-0ea8-483c-8c5b-7413ac0f137c · outbound
Thinkless: LLM Learns When to Think Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3edf7244-d894-4412-90b8-661b17b5c710 · outbound
Thinkless: LLM Learns When to Think Training Verifiers to Solve Math Word Problems
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7f4a4c74-5368-4c51-8aa5-53c77003f58d · outbound
Thinkless: LLM Learns When to Think The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4023a631-6fd5-4edf-965d-045eca0df593 · outbound
Thinkless: LLM Learns When to Think Open r1: A fully open reproduction of deepseek-r1, January 2025
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a160fda-eaf2-4bed-ae2a-e5bad2638398 · outbound
Thinkless: LLM Learns When to Think Efficient reasoning models: A survey
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25be10f0-0e5c-479e-924e-0edfb7afa20b · outbound
Thinkless: LLM Learns When to Think DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4527e8d-3ce6-4cf2-b160-b8105b570bd1 · outbound
Thinkless: LLM Learns When to Think Token-Budget-Aware LLM Reasoning
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f1e7d7f-250a-4159-bb17-2bc7a9879eab · outbound
Thinkless: LLM Learns When to Think DNR Bench: Benchmarking Over-Reasoning in Reasoning LLMs
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58762963-db4d-447f-9146-aeb084b068d0 · outbound
Thinkless: LLM Learns When to Think Measuring mathematical problem solving with the math dataset
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0e3ed912-d38d-4265-8468-96bd9b813c1f · outbound
Thinkless: LLM Learns When to Think Distilling the Knowledge in a Neural Network
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5350e465-f236-4214-9a56-6d538275b34c · outbound
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfe1e0b4-81c8-499b-b276-a3e8100a7fde · outbound
Thinkless: LLM Learns When to Think C3oT: Generating Shorter Chain-of-Thought without Compromising Effectiveness
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b3a235d-27d7-4f89-95c4-f4f80d600fe4 · outbound
Thinkless: LLM Learns When to Think Mixed Distillation Helps Smaller Language Model Better Reasoning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb7ac746-8b96-4272-a805-fa9c2a6dcdeb · outbound
Thinkless: LLM Learns When to Think Small models struggle to learn from strong reasoners
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9886e641-06b5-49a4-8537-079a74175c48 · outbound
Thinkless: LLM Learns When to Think Reward-Guided Speculative Decoding for Efficient LLM Reasoning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 094d3205-2ebb-4a17-9b3a-927223114517 · outbound
Thinkless: LLM Learns When to Think Let's Verify Step by Step
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b63a05a-2386-4b42-ab37-8bf82d46e761 · outbound
Thinkless: LLM Learns When to Think Understanding R1-Zero-Like Training: A Critical Perspective
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6be362e5-ec4a-4741-85bf-b51008340dd0 · outbound
Thinkless: LLM Learns When to Think O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d385ba2-2110-493d-9f44-93bd27126dfc · outbound
Thinkless: LLM Learns When to Think Deepscaler: Surpassing o1-preview with a 1.5 b model by scaling rl
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6ce3ae65-f8dc-4125-a099-fdfd07f644d4 · outbound
Thinkless: LLM Learns When to Think CoT-Valve: Length-Compressible Chain-of-Thought Tuning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19c9988b-c9fa-4eca-bb16-c5da06d5d596 · outbound
Thinkless: LLM Learns When to Think Teaching Small Language Models to Reason
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea1ed517-940b-4910-b110-f741139bc29a · outbound
Thinkless: LLM Learns When to Think RouteLLM: Learning to Route LLMs with Preference Data
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f67bb2a3-ab42-4bc5-8ad4-e90447f4cdc1 · outbound
Thinkless: LLM Learns When to Think Reasoning with latent thoughts: On the power of looped transformers
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5e2bd26a-c325-45ac-a72f-95429c206f80 · outbound
Thinkless: LLM Learns When to Think DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b14ebac2-4633-471d-80c9-fa7387f8477e · outbound
Thinkless: LLM Learns When to Think HybridFlow: A Flexible and Efficient RLHF Framework
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a31032c-3ed7-465f-b488-deffbf693c98 · outbound
Thinkless: LLM Learns When to Think Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7505c744-9322-4684-9ed6-da3ae807855e · outbound
Thinkless: LLM Learns When to Think Towards reasoning ability of small language models
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd072823-7a31-48f5-918d-4312f8c598f2 · outbound
Thinkless: LLM Learns When to Think Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47174539-7f7c-41aa-8302-aa65d71cc4b8 · outbound
Thinkless: LLM Learns When to Think Open thoughts, January 2025
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a7f05102-f52c-46a4-b9ab-dd33f7aff274 · outbound
Thinkless: LLM Learns When to Think Qwen3, April 2025
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6043dd08-4a97-4861-a327-c9db7bab9439 · outbound
Thinkless: LLM Learns When to Think Aime problem set 1983-2024, 2023
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7df49e20-7550-43a4-a10a-068ec507bfac · outbound
Thinkless: LLM Learns When to Think Chain-of-thought prompting elicits reasoning in large language models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f73032e7-a3e2-492b-935a-10265f918540 · outbound
Thinkless: LLM Learns When to Think Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e9b6396-131b-4c12-8c5b-37556c95cc88 · outbound
Thinkless: LLM Learns When to Think Chain of Draft: Thinking Faster by Writing Less
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b9d49cf-0538-4a3a-834d-b41eaf729b02 · outbound
Thinkless: LLM Learns When to Think Qwen2.5 Technical Report
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6c6ada6-a173-4deb-8bc1-07bdca586f08 · outbound
Thinkless: LLM Learns When to Think Distilling System 2 into System 1
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e4027dd-7c9d-4f87-8d03-0cfa0a584f4f · outbound
Thinkless: LLM Learns When to Think SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68ddf33c-0af0-4c0b-98e8-09b54dbfa625 · outbound
Thinkless: LLM Learns When to Think Improving Mathematical Reasoning Capabilities of Small Language Models via Feedback-Driven Distillation
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5805c16f-6451-464d-8c15-ce71414fc658 · inbound
VeriThinker: Learning to Verify Makes Reasoning Model Efficient Thinkless: LLM Learns When to Think
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24297eed-6dbe-462b-9d27-58e0c3675a5e · inbound
How Far Are We from Optimal Reasoning Efficiency? Thinkless: LLM Learns When to Think
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9314b0c2-61ef-45fb-95fb-099941b95cb7 · inbound
Schema-R1: A reasoning training approach for schema linking in Text-to-SQL Task Thinkless: LLM Learns When to Think
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65dfbec1-48ed-4253-ab04-18b7438fc778 · inbound
Do Thinking Tokens Help or Trap? Towards More Efficient Large Reasoning Model Thinkless: LLM Learns When to Think
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d898a92-237a-487e-9dd3-77011c2e65ed · inbound
SmartThinker: Learning to Compress and Preserve Reasoning by Step-Level Length Control Thinkless: LLM Learns When to Think
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abe65655-e6ef-4284-bd57-2d0231e6b8fd · inbound
KAT-V1: Kwai-AutoThink Technical Report Thinkless: LLM Learns When to Think
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cc5bc8c-2fbe-4908-8efc-8ad752250969 · inbound
Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey Thinkless: LLM Learns When to Think
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c2ae177-ebe1-478c-a172-115c2560f6d9 · inbound
LAPO: Internalizing Reasoning Efficiency via Length-Adaptive Policy Optimization Thinkless: LLM Learns When to Think
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c0aaae6-8018-48ee-8579-0d863a965da8 · inbound
Hierarchical Budget Policy Optimization for Adaptive Reasoning Thinkless: LLM Learns When to Think
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9490e6c-3a40-47f8-b36b-55aaedaecd12 · inbound
Think in Blocks: Adaptive Reasoning from Direct Response to Deep Reasoning Thinkless: LLM Learns When to Think
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1f49abb-7da4-47c4-8008-68286d13377b · inbound
Implicit Reasoning in Large Language Models: A Comprehensive Survey Thinkless: LLM Learns When to Think
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b052fbbe-70e8-48c3-a101-990c49522374 · inbound
Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle Thinkless: LLM Learns When to Think
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ed3392b-55b4-4bb4-84e2-54a430e19043 · inbound
Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts Thinkless: LLM Learns When to Think
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 28a8a6f2-cf2c-473d-8eb2-606ec986f069 · inbound
Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards Thinkless: LLM Learns When to Think
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a9630c31-87df-423b-bd84-2c1ea5043bdc · inbound
Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards Thinkless: LLM Learns When to Think
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e8f611a-3ac5-4afd-8c0b-aa806a81f98f · inbound
AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification Thinkless: LLM Learns When to Think
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 060c60f2-3ee7-43e3-9add-38dba8d67927 · inbound
Verifying Meta-Awareness via Predictive Rewards in Reasoning Models Thinkless: LLM Learns When to Think
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ca94dd9-b465-4765-ae7d-c22765a14997 · inbound
Probing the Difficulty Perception Mechanism of Large Language Models Thinkless: LLM Learns When to Think
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30502e58-4101-4174-ab32-97238c0d9876 · inbound
MixReasoning: Switching Modes to Think Thinkless: LLM Learns When to Think
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04c5c3d1-5b28-40bb-a0d6-0a850f68ae6a · inbound
Rectifying LLM Thought from Lens of Optimization Thinkless: LLM Learns When to Think
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 87ae3283-f0da-420f-b089-ec5dbb03a541 · inbound
Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers Thinkless: LLM Learns When to Think
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e2c4434-3505-44c1-ba00-5bedda487da8 · inbound
ConPress: Learning Efficient Reasoning from Multi-Question Contextual Pressure Thinkless: LLM Learns When to Think
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1130b5d-0ecd-4eee-b58f-8a9e13aace82 · inbound
Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression Thinkless: LLM Learns When to Think
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 42201804-09fc-49d4-b13d-f17754432bd8 · inbound
Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression Thinkless: LLM Learns When to Think
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f49898b-951b-4d7f-9318-04d2fb1d0a32 · inbound
Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning Thinkless: LLM Learns When to Think
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df110d54-c29b-4a9f-8b7f-73924b8abdf2 · inbound
Shorter, but Still Trustworthy? An Empirical Study of Chain-of-Thought Compression Thinkless: LLM Learns When to Think
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 27b12289-0739-46de-8e32-76732aec461b · inbound
Learning to Interrupt in Language-based Multi-agent Communication Thinkless: LLM Learns When to Think
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 23ba1d01-8a16-434a-a672-35a0a99b5740 · inbound
When Less is Enough: Efficient Inference via Collaborative Reasoning Thinkless: LLM Learns When to Think
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0732046b-2927-46aa-9cc6-ea73bb53c093 · inbound
Efficient Agentic Reasoning Through Self-Regulated Simulative Planning Thinkless: LLM Learns When to Think
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 28ce41e8-a0a0-4d6a-80db-0917324f2e7f · inbound
From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Thinkless: LLM Learns When to Think
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7b4dd15a-7162-40f7-bc02-d32b33428e20 · inbound
Finding the Time to Think: Learning Planning Budgets in Real-Time RL Thinkless: LLM Learns When to Think
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c98d59f4-8e26-401e-84f8-f3af8e19d670 · inbound
Finding the Time to Think: Learning Planning Budgets in Real-Time RL Thinkless: LLM Learns When to Think
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2ffaa845-bf7a-4845-80b9-3816d5477890 · inbound
When to Plan: Learning to Select Between Reactive Control and Deliberative Planning Thinkless: LLM Learns When to Think
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51e39b87-e1fb-45dd-968a-7e15bdbab9f5 · inbound
AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning Thinkless: LLM Learns When to Think
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e249b7d-bd6a-40e0-871e-0056f67c682c · inbound
Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning Thinkless: LLM Learns When to Think
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 521a6c05-246e-4cea-aafb-9239b739d344 · inbound
ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling Thinkless: LLM Learns When to Think
Reference 114
Source-reported events for the cited work
Unavailable: canonical work link unavailable.