Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:38:09.532937Z
Paper Citation Record · LEDGER
As of 12 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 11 inbound Pith citation observations for arXiv:2506.10446.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:38:09.532937Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T17:54:17.219235Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T11:37:03.402806Z
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0069c815-bd16-4231-8173-7a5934c47051 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty URL: " 'urlintro :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4980ab87-9059-4904-b725-1fbe1c83ae99 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6aeaa226-188c-4a21-a082-b5543e7e2a48 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty GPT-4 Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3402a25a-1763-4e0c-a88b-a64da857ae46 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 908d5555-fa5d-4423-b0e8-8c01887fd963 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Back to Basics: Revisiting REINFORCE Style Optimization for Learning from Human Feedback in LLMs
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b665ca90-4556-469d-b170-c5d81f76ea51 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c99523a-ece3-43b6-88ec-2b64a9fcaa15 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7398259e-816c-4ee9-9ffa-ec0a5567b183 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Training Verifiers to Solve Math Word Problems
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d3c70e8-74e4-4cf2-80e1-95f6f15f6072 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Break the Chain: Large Language Models Can be Shortcut Reasoners
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e67df85b-d344-4f51-8833-ae66fee42643 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Alphazero-like Tree-Search can Guide Large Language Model Decoding and Training
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d298be2c-dca1-4534-b93a-5f0ed09b9553 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c883f271-b8c6-46de-be2f-abf7ad3f223f · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Token-Budget-Aware LLM Reasoning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7de45b5-aa0c-48ee-85a2-47fc2c432ccc · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Qwen2.5-Coder Technical Report
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95c9409d-08a4-436d-ab34-5940fff60271 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1effa212-a6a7-4461-8958-7e997de45ed1 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty How Well do LLMs Compress Their Own Chain-of-Thought? A Token Complexity Approach
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1c23248-35b3-439c-91e7-dedaa67ed2da · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29ea50cc-b6d1-4a6c-8b6a-4dc85ddca9ab · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Can Language Models Learn to Skip Steps?
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37e8aeb8-dd22-4af2-b8fc-8bd33ca73664 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e617a0c-4d1c-40c4-9b1c-089b198ab836 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty CoT-Valve: Length-Compressible Chain-of-Thought Tuning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29ebbd21-de7d-44f0-b298-a49443504355 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Self-Training Elicits Concise Reasoning in Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc246e1e-f44a-4d10-9d85-744785eb390a · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation ef1838da-9d3b-4dcc-b568-afde3d122de3 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90e46486-f059-4c37-a975-434082a63560 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Optimizing Test-Time Compute via Meta Reinforcement Fine-Tuning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd99e05-0f91-4449-858d-59837883ed03 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Proximal Policy Optimization Algorithms
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b5b2642-3fe4-4aff-9dc8-97c1a30d7232 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4479b550-e857-41f0-be56-da4d091dd45f · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 35933934-069a-4963-bb2a-00c124f287a4 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Kimi k1.5: Scaling Reinforcement Learning with LLMs
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00c920ab-74c4-4ebe-a797-247384b583c4 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 0b57aabd-de9b-4a89-b2e8-dc5e9ad2cece · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d873ba20-899e-44a3-a97c-bd6f33e6daff · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cebeac8f-9693-428e-a446-8303a767d6db · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b38dbbbf-32d0-4f04-b678-3fbd2a42fdc2 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Chain of Draft: Thinking Faster by Writing Less
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c8a539-fe39-48dc-9a2c-88b63d6ab88d · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Unresolved cited work
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d12d43a8-1b04-49a0-86fe-b96238f7adae · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty LIMO: Less is More for Reasoning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c333e166-6cdc-4f61-8eb9-33e7b9cf86b2 · outbound
Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty Least-to-Most Prompting Enables Complex Reasoning in Large Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b70cf35-6073-4738-a278-6306d0f96f9e · inbound
Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53d96970-d5c3-4493-a872-35f66a92d696 · inbound
Baichuan-M2: Scaling Medical Capability with Large Verifier System Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 466f95ac-15c9-42fd-8d4e-9166aa94374e · inbound
Baichuan-M2: Scaling Medical Capability with Large Verifier System Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8665faff-1f48-4338-90cc-d2a77ebdebf2 · inbound
Learning to Reason Efficiently with Discounted Reinforcement Learning Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a3c8790-f987-4c3a-9320-c06007867427 · inbound
Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51a4b97b-b9c2-44c7-921f-bab23716752b · inbound
Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 357fa96c-7cad-4d73-ab57-e01601f11fd8 · inbound
PR-CAD: Progressive Refinement for Unified Controllable and Faithful Text-to-CAD Generation with Large Language Models Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation e22b7477-10a0-4a34-b13e-281a7de1e7e1 · inbound
Reinforcement Learning for Tool-Calling Agents in Fast Healthcare Interoperability Resources (FHIR) Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 9b9a0fb6-a743-4d4a-bbc2-5ae534685e3d · inbound
SLAT: Segment-Level Adaptive Trimming for Efficient CoT Reasoning Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation a82d171d-7af1-424e-9a02-6588c10dc8c2 · inbound
Beyond Penalizing Mistakes: Stabilizing Efficiency Training in Large Reasoning Models via Adaptive Correct-Only Rewards Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.
Observation 7aa84793-094d-4c27-bc20-0c9d0a2e9eed · inbound
A First-Principles Theory of Slow Thinking and Active Perception Fast on the Easy, Deep on the Hard: Efficient Reasoning via Powered Length Penalty
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.