Pith. sign in

Paper Citation Record · LEDGER

Thinkless: LLM Learns When to Think

As of 20 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 36 inbound Pith citation observations for arXiv:2505.13379.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.13379 v2

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:20:19.609052Z

measured 80 of 80 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 36 of 36 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:50:30.535082Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T13:29:51.951052Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved39
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3549abb0-1ed0-4506-9c89-dc08d8517cc0 · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Thinkless: LLM Learns When to Think L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.422776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.422776Z digest=sha256:26da4b4cb6fd0764e1f018f5f5469fa639fc467d57893dddda9d2401dfa89db0

Observation 0739fbfa-8df9-43d8-90b8-5a1a695d054a · outbound

This paper cites Claude 3.7 Sonnet.

Thinkless: LLM Learns When to Think Claude 3.7 Sonnet

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:20:20.566584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T20:20:19.428812Z digest=sha256:771f3d7d46e3158a024b33101f32896ab632c758a2100632966c60e3792c4a10

Observation 63f97f2b-fe1f-4da9-be21-f0b85cab4747 · outbound

This paper cites Sketch-of-thought: Efficient llm reasoning with adaptive cognitive-inspired sketching.

Thinkless: LLM Learns When to Think Sketch-of-thought: Efficient llm reasoning with adaptive cognitive-inspired sketching

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.433529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.433529Z digest=sha256:66ca256a1ed416ccf9346950f25968ead5cd3e7d7a453f220d7583cbef91e598

Observation afb52ade-f5d9-4775-823b-7f24a0074784 · outbound

This paper cites Llama-nemotron: Efficient reasoning models, 2025.

Thinkless: LLM Learns When to Think Llama-nemotron: Efficient reasoning models, 2025

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:20:20.551089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T20:20:19.438012Z digest=sha256:94c8bbfa638d8481d9ba832ecdb2600c8bd07799eb7f59e73c8ea40c3d1440e1

Observation aae95e0e-c178-4fe7-8340-dd6aa7203a03 · outbound

This paper cites ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning.

Thinkless: LLM Learns When to Think ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.443034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.443034Z digest=sha256:dab9069ac2a36d142c733a012de4dac90a56717884c8e59ef7f64c1bfd5cbcf7

Observation 07575248-a8f2-4122-9cd9-3cdb2ccaaff7 · outbound

This paper cites Distilling Reasoning Ability from Large Language Models with Adaptive Thinking.

Thinkless: LLM Learns When to Think Distilling Reasoning Ability from Large Language Models with Adaptive Thinking

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.447284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.447284Z digest=sha256:0a1898ea9a79f347a022111f0e34496d428b8e459d8d0a38cc0eec7edb31e8b3

Observation 16383140-0ea8-483c-8c5b-7413ac0f137c · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

Thinkless: LLM Learns When to Think Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.451908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.451908Z digest=sha256:137037b4595e657d2a78da0f503e2911c70c8b0dfda254633d07ba1cdb51619d

Observation 3edf7244-d894-4412-90b8-661b17b5c710 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Thinkless: LLM Learns When to Think Training Verifiers to Solve Math Word Problems

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.456028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.456028Z digest=sha256:cdbd9ea89725b96408dd18fa128b54a4e50643e6a454346d5274b2f267b8a457

Observation 7f4a4c74-5368-4c51-8aa5-53c77003f58d · outbound

This paper cites The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks.

Thinkless: LLM Learns When to Think The Danger of Overthinking: Examining the Reasoning-Action Dilemma in Agentic Tasks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.460532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.460532Z digest=sha256:8d2928fac9d9bf8f83eda4743c40288606e91b33161548a4a38a7d1d8a0c835a

Observation 4023a631-6fd5-4edf-965d-045eca0df593 · outbound

This paper cites Open r1: A fully open reproduction of deepseek-r1, January 2025.

Thinkless: LLM Learns When to Think Open r1: A fully open reproduction of deepseek-r1, January 2025

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.465136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.465136Z digest=sha256:a1b94fb26c2d28dbd89f18411b81be5057ee05624aae124240e27491ec9b43ed

Observation 8a160fda-eaf2-4bed-ae2a-e5bad2638398 · outbound

This paper cites Efficient reasoning models: A survey.

Thinkless: LLM Learns When to Think Efficient reasoning models: A survey

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.468663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.468663Z digest=sha256:07e6a7b5644fd9d754e747c78158baebd142eb76ead75e2c1e1a8f371154af5f

Observation 25be10f0-0e5c-479e-924e-0edfb7afa20b · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Thinkless: LLM Learns When to Think DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.472575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.472575Z digest=sha256:3cf3db202243a181a200b362b973d627a983e56f15b13f92d9e7dd4cd6c72768

Observation c4527e8d-3ce6-4cf2-b160-b8105b570bd1 · outbound

This paper cites Token-Budget-Aware LLM Reasoning.

Thinkless: LLM Learns When to Think Token-Budget-Aware LLM Reasoning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.476509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.476509Z digest=sha256:db36671f4f08d74fe52c397363469c70b3d2fb2b67778b2df5922b90af51da8a

Observation 6f1e7d7f-250a-4159-bb17-2bc7a9879eab · outbound

This paper cites DNR Bench: Benchmarking Over-Reasoning in Reasoning LLMs.

Thinkless: LLM Learns When to Think DNR Bench: Benchmarking Over-Reasoning in Reasoning LLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.480349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.480349Z digest=sha256:c89fe6fa491df4f8b7907ac9fdf681a3b121014c960ec97921dda2e3279c80a8

Observation 58762963-db4d-447f-9146-aeb084b068d0 · outbound

This paper cites Measuring mathematical problem solving with the math dataset.

Thinkless: LLM Learns When to Think Measuring mathematical problem solving with the math dataset

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.484371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.484371Z digest=sha256:c1a9fa4294dc56390d9bdf353343c0026d271a332df970c70ea303dd1cbe8d27

Observation 0e3ed912-d38d-4265-8468-96bd9b813c1f · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Thinkless: LLM Learns When to Think Distilling the Knowledge in a Neural Network

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.488104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.488104Z digest=sha256:11f8c3a2bea5122c43483d165dbbf389c0aa1f939ba14963463df05241509dbd

Observation 5350e465-f236-4214-9a56-6d538275b34c · outbound

This paper cites OpenAI o1 System Card.

Thinkless: LLM Learns When to Think OpenAI o1 System Card

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.491794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.491794Z digest=sha256:e61690b13392617316a52f2800328e705a586d6dbb85d9e62dd4270ab839c79b

Observation bfe1e0b4-81c8-499b-b276-a3e8100a7fde · outbound

This paper cites C3oT: Generating Shorter Chain-of-Thought without Compromising Effectiveness.

Thinkless: LLM Learns When to Think C3oT: Generating Shorter Chain-of-Thought without Compromising Effectiveness

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.496387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.496387Z digest=sha256:1ebe80bfaf124d1844ea6e6d6156a480fefbb84a99cc6c2af531b58b8b597a66

Observation 1b3a235d-27d7-4f89-95c4-f4f80d600fe4 · outbound

This paper cites Mixed Distillation Helps Smaller Language Model Better Reasoning.

Thinkless: LLM Learns When to Think Mixed Distillation Helps Smaller Language Model Better Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.500946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.500946Z digest=sha256:5c62f5488c25db9a92e1fb6036b482824b0d7068b2c4e3c6519187bab223e06c

Observation cb7ac746-8b96-4272-a805-fa9c2a6dcdeb · outbound

This paper cites Small models struggle to learn from strong reasoners.

Thinkless: LLM Learns When to Think Small models struggle to learn from strong reasoners

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.505273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.505273Z digest=sha256:c29c2b90f4ba22eff250dc391478bd707ea86b364f7d658efc1df4763a96b0b7

Observation 9886e641-06b5-49a4-8537-079a74175c48 · outbound

This paper cites Reward-Guided Speculative Decoding for Efficient LLM Reasoning.

Thinkless: LLM Learns When to Think Reward-Guided Speculative Decoding for Efficient LLM Reasoning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.509764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.509764Z digest=sha256:9b747437b55c32fbaae6d2c5e80645128cfe7070bc90dc55f9d232f3aff7b802

Observation 094d3205-2ebb-4a17-9b3a-927223114517 · outbound

This paper cites Let's Verify Step by Step.

Thinkless: LLM Learns When to Think Let's Verify Step by Step

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.514360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.514360Z digest=sha256:e5eb2df7f36bd8f855a43e1d8af89f735770383119d3996e5873e61fc76ceff6

Observation 8b63a05a-2386-4b42-ab37-8bf82d46e761 · outbound

This paper cites Understanding R1-Zero-Like Training: A Critical Perspective.

Thinkless: LLM Learns When to Think Understanding R1-Zero-Like Training: A Critical Perspective

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.518752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.518752Z digest=sha256:d7ebd79ca607e18cc45f30d9b063e870fb48c789c86e73c022ff24971432f15d

Observation 6be362e5-ec4a-4741-85bf-b51008340dd0 · outbound

This paper cites O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning.

Thinkless: LLM Learns When to Think O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.522736Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.522736Z digest=sha256:556aabc33e35f63010db31bc6180352359bb792948c760e4bf8db4a61a7d461b

Observation 3d385ba2-2110-493d-9f44-93bd27126dfc · outbound

This paper cites Deepscaler: Surpassing o1-preview with a 1.5 b model by scaling rl.

Thinkless: LLM Learns When to Think Deepscaler: Surpassing o1-preview with a 1.5 b model by scaling rl

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:20:20.514695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T20:20:19.527543Z digest=sha256:91d8fe01dcc9f72613ff822e1d32be80357677e4aa9d9c6e4cfdf03999d289b6

Observation 6ce3ae65-f8dc-4125-a099-fdfd07f644d4 · outbound

This paper cites CoT-Valve: Length-Compressible Chain-of-Thought Tuning.

Thinkless: LLM Learns When to Think CoT-Valve: Length-Compressible Chain-of-Thought Tuning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.532146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.532146Z digest=sha256:8d2aa5314c630fea0ca560f8be1a280a02b59efecc574c6e57bf46ef90ea127e

Observation 19c9988b-c9fa-4eca-bb16-c5da06d5d596 · outbound

This paper cites Teaching Small Language Models to Reason.

Thinkless: LLM Learns When to Think Teaching Small Language Models to Reason

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.535956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.535956Z digest=sha256:85374fd231c10b0867fcd151b7790af84634d7d5eec8ee79bd58cbc2f9db7380

Observation ea1ed517-940b-4910-b110-f741139bc29a · outbound

This paper cites RouteLLM: Learning to Route LLMs with Preference Data.

Thinkless: LLM Learns When to Think RouteLLM: Learning to Route LLMs with Preference Data

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.540076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.540076Z digest=sha256:8fa239999e3aaca3b2c19d6e3c5dd669be7dcc1a1dcdd096817689ae9a166e62

Observation f67bb2a3-ab42-4bc5-8ad4-e90447f4cdc1 · outbound

This paper cites Reasoning with latent thoughts: On the power of looped transformers.

Thinkless: LLM Learns When to Think Reasoning with latent thoughts: On the power of looped transformers

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:20:20.498253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T20:20:19.544158Z digest=sha256:0444c80e235e0024d7f2262c9f1719d9e6502570919a84549f66b7492679cf21

Observation 5e2bd26a-c325-45ac-a72f-95429c206f80 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Thinkless: LLM Learns When to Think DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.548578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.548578Z digest=sha256:3986344d049aea10374976bd2e8de00766cbca2ce130ce15f6a54086b77a9451

Observation b14ebac2-4633-471d-80c9-fa7387f8477e · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Thinkless: LLM Learns When to Think HybridFlow: A Flexible and Efficient RLHF Framework

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.552365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.552365Z digest=sha256:9cae09fd461fb3a2b9ee4fdfc0d709748fb4844d5303801bb0ab61cef70c998c

Observation 8a31032c-3ed7-465f-b488-deffbf693c98 · outbound

This paper cites Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism.

Thinkless: LLM Learns When to Think Megatron-LM: Training Multi-Billion Parameter Language Models Using Model Parallelism

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.556246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.556246Z digest=sha256:e7310c20208b7fd335b1e5e4f51975438146d57c404ca53b7a5f73820333bb20

Observation 7505c744-9322-4684-9ed6-da3ae807855e · outbound

This paper cites Towards reasoning ability of small language models.

Thinkless: LLM Learns When to Think Towards reasoning ability of small language models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.559968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.559968Z digest=sha256:cafec9449d31772f1a269b7b6ed5cb1fde13c829c8f51176e728e1acc5fa3793

Observation cd072823-7a31-48f5-918d-4312f8c598f2 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

Thinkless: LLM Learns When to Think Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.563711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.563711Z digest=sha256:4fb520616bd425dd50b2f3c98d19e0fe6368efaec54ed1f6153b1407fb6fe783

Observation 47174539-7f7c-41aa-8302-aa65d71cc4b8 · outbound

This paper cites Open thoughts, January 2025.

Thinkless: LLM Learns When to Think Open thoughts, January 2025

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:20:20.482508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-15T20:20:19.567365Z digest=sha256:a6239229651100649b54c311128edc591ed09ea49d833aeb0dcca61eb9bfb230

Observation a7f05102-f52c-46a4-b9ab-dd33f7aff274 · outbound

This paper cites Qwen3, April 2025.

Thinkless: LLM Learns When to Think Qwen3, April 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.571387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.571387Z digest=sha256:f7e65f14740c967c3c0adade83a82a40abea4e630d51e515510c8757d45a3df9

Observation 6043dd08-4a97-4861-a327-c9db7bab9439 · outbound

This paper cites Aime problem set 1983-2024, 2023.

Thinkless: LLM Learns When to Think Aime problem set 1983-2024, 2023

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.575279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.575279Z digest=sha256:9e34a1a523da2e9e421ba985475d97dad003d59ce39d6fb594fa25d3eebc6318

Observation 7df49e20-7550-43a4-a10a-068ec507bfac · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

Thinkless: LLM Learns When to Think Chain-of-thought prompting elicits reasoning in large language models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.579799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.579799Z digest=sha256:8b056827a3ba06cb483d93e264907fbf85a052bcb93da7199fbfee057059b9f8

Observation f73032e7-a3e2-492b-935a-10265f918540 · outbound

This paper cites Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools.

Thinkless: LLM Learns When to Think Agentic Reasoning: A Streamlined Framework for Enhancing LLM Reasoning with Agentic Tools

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.583886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.583886Z digest=sha256:503b67d7ce4707cae5491d9757f7fd7ba736ffadb896a1bbdebb7f3f2f0b3be0

Observation 3e9b6396-131b-4c12-8c5b-37556c95cc88 · outbound

This paper cites Chain of Draft: Thinking Faster by Writing Less.

Thinkless: LLM Learns When to Think Chain of Draft: Thinking Faster by Writing Less

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.588106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.588106Z digest=sha256:c860ba52b1d64f3bf21d215a35c5e4ee2cfc07084e111dde51563edf6c7c484b

Observation 0b9d49cf-0538-4a3a-834d-b41eaf729b02 · outbound

This paper cites Qwen2.5 Technical Report.

Thinkless: LLM Learns When to Think Qwen2.5 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.592590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.592590Z digest=sha256:4edac630c1cee92d62e4d5602083fa956ff182fe2176dc7c1296ba8f91d4a1d5

Observation e6c6ada6-a173-4deb-8bc1-07bdca586f08 · outbound

This paper cites Distilling System 2 into System 1.

Thinkless: LLM Learns When to Think Distilling System 2 into System 1

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.598477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.598477Z digest=sha256:f3bcf04210a7a81eea0992c5c75c92ad0dc07789ee82b1dbaa1b38e63664b619

Observation 9e4027dd-7c9d-4f87-8d03-0cfa0a584f4f · outbound

This paper cites SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild.

Thinkless: LLM Learns When to Think SimpleRL-Zoo: Investigating and Taming Zero Reinforcement Learning for Open Base Models in the Wild

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.603841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.603841Z digest=sha256:8148e423688dc4824bd74c06d04f41cd604450a5493be96bf7e3ceb39c6dcc75

Observation 68ddf33c-0af0-4c0b-98e8-09b54dbfa625 · outbound

This paper cites Improving Mathematical Reasoning Capabilities of Small Language Models via Feedback-Driven Distillation.

Thinkless: LLM Learns When to Think Improving Mathematical Reasoning Capabilities of Small Language Models via Feedback-Driven Distillation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T20:20:19.609052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:20:19.609052Z digest=sha256:b2e163d1872a28132aac66bb5ca77c0b51d1d22ebe579f0d580bdbb6d7d73746

Pith citing papers

Observation 5805c16f-6451-464d-8c15-ce71414fc658 · inbound

VeriThinker: Learning to Verify Makes Reasoning Model Efficient cites this paper.

VeriThinker: Learning to Verify Makes Reasoning Model Efficient Thinkless: LLM Learns When to Think

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:42:11.153934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:42:11.153934Z digest=sha256:32c525bec1fd82923ae7a8aa59f4cccedd496387fbb50eae9b5435e6a92b1eff

Observation 24297eed-6dbe-462b-9d27-58e0c3675a5e · inbound

How Far Are We from Optimal Reasoning Efficiency? cites this paper.

How Far Are We from Optimal Reasoning Efficiency? Thinkless: LLM Learns When to Think

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:49:35.129268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:49:35.129268Z digest=sha256:b4fc7450cde94fc6479b9841e7962f2bdbef4b3d725d8d442f2d407b92fed9d2

Observation 9314b0c2-61ef-45fb-95fb-099941b95cb7 · inbound

Schema-R1: A reasoning training approach for schema linking in Text-to-SQL Task cites this paper.

Schema-R1: A reasoning training approach for schema linking in Text-to-SQL Task Thinkless: LLM Learns When to Think

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T01:05:26.831361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:05:26.831361Z digest=sha256:b79e24f34b487fd1ee028e6c81bbb4ea9893c9c432fc108189135eb7eca8cef5

Observation 65dfbec1-48ed-4253-ab04-18b7438fc778 · inbound

Do Thinking Tokens Help or Trap? Towards More Efficient Large Reasoning Model cites this paper.

Do Thinking Tokens Help or Trap? Towards More Efficient Large Reasoning Model Thinkless: LLM Learns When to Think

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T21:45:07.494215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:45:07.494215Z digest=sha256:07c094ed8213306744e6b3523813150104391b9ef6939fa15a3ed1de3e116fe2

Observation 6d898a92-237a-487e-9dd3-77011c2e65ed · inbound

SmartThinker: Learning to Compress and Preserve Reasoning by Step-Level Length Control cites this paper.

SmartThinker: Learning to Compress and Preserve Reasoning by Step-Level Length Control Thinkless: LLM Learns When to Think

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:56:08.680429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:56:08.680429Z digest=sha256:99fc9e0b3bec6ed4702485a77877435494fa060d94c1ab06badf7731902ff96f

Observation abe65655-e6ef-4284-bd57-2d0231e6b8fd · inbound

KAT-V1: Kwai-AutoThink Technical Report cites this paper.

KAT-V1: Kwai-AutoThink Technical Report Thinkless: LLM Learns When to Think

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T18:28:29.779954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:28:29.779954Z digest=sha256:e7e07904c926b6bc780c9d4931224f4e090164c44d0cef3e9ad6c2077d6c12b2

Observation 8cc5bc8c-2fbe-4908-8efc-8ad752250969 · inbound

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey cites this paper.

Towards Concise and Adaptive Thinking in Large Reasoning Models: A Survey Thinkless: LLM Learns When to Think

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T17:53:46.856179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:53:46.856179Z digest=sha256:cc500bdd3ce57df746a5eaa5310567a9e18f385e3e5c181b8f74260e2ff9371c

Observation 0c2ae177-ebe1-478c-a172-115c2560f6d9 · inbound

LAPO: Internalizing Reasoning Efficiency via Length-Adaptive Policy Optimization cites this paper.

LAPO: Internalizing Reasoning Efficiency via Length-Adaptive Policy Optimization Thinkless: LLM Learns When to Think

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T15:29:42.379745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:29:42.379745Z digest=sha256:b0c7833fbb0e4346c982817bc6d157a376c5a0e73f3de5a1712c83380e9994a3

Observation 6c0aaae6-8018-48ee-8579-0d863a965da8 · inbound

Hierarchical Budget Policy Optimization for Adaptive Reasoning cites this paper.

Hierarchical Budget Policy Optimization for Adaptive Reasoning Thinkless: LLM Learns When to Think

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T15:31:27.442287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:31:27.442287Z digest=sha256:ff01c7df9ea236fccfe559b0bfd6f24a50225781bcfd9cd6ee2c0fa7cd2b4714

Observation b9490e6c-3a40-47f8-b36b-55aaedaecd12 · inbound

Think in Blocks: Adaptive Reasoning from Direct Response to Deep Reasoning cites this paper.

Think in Blocks: Adaptive Reasoning from Direct Response to Deep Reasoning Thinkless: LLM Learns When to Think

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T17:52:51.135256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T17:52:51.135256Z digest=sha256:c9066952c770ab2b2e1e4ab77dd39c1f68e4e0ed4cc9761b1b26b216cdf827ad

Observation c1f49abb-7da4-47c4-8008-68286d13377b · inbound

Implicit Reasoning in Large Language Models: A Comprehensive Survey cites this paper.

Implicit Reasoning in Large Language Models: A Comprehensive Survey Thinkless: LLM Learns When to Think

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T11:39:36.526564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:39:36.526564Z digest=sha256:cae4cfbc0603489fe900c10e8b786d26ceca18c2eda3944b08fbabe8a0e58575

Observation b052fbbe-70e8-48c3-a101-990c49522374 · inbound

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle cites this paper.

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle Thinkless: LLM Learns When to Think

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T16:07:28.955737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:07:28.955737Z digest=sha256:6e4230400eeded0ee572c2cf7ef975b88332dccf74d4d3aed77a009f9adc567f

Observation 6ed3392b-55b4-4bb4-84e2-54a430e19043 · inbound

Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts cites this paper.

Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts Thinkless: LLM Learns When to Think

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:42:38.713577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T13:42:07.883909Z digest=sha256:4859e8c814961d94deda8a1378de25239daccf77360d28cbdd5f4e21a2f88327

Observation 28a8a6f2-cf2c-473d-8eb2-606ec986f069 · inbound

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards cites this paper.

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards Thinkless: LLM Learns When to Think

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:26:28.293067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T14:24:48.666197Z digest=sha256:2975fac39f37da22c2fd788bd516686f6b1628f4f00e7a9c3fd7db67fc8f34d0

Observation a9630c31-87df-423b-bd84-2c1ea5043bdc · inbound

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards cites this paper.

Position: The Hidden Costs and Measurement Gaps of Reinforcement Learning with Verifiable Rewards Thinkless: LLM Learns When to Think

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:48:29.032685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:48:29.032685Z digest=sha256:56057328b27b5ecc1c7bacd83e2a98cee77ec3deb4cd026640537ea132041896

Observation 2e8f611a-3ac5-4afd-8c0b-aa806a81f98f · inbound

AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification cites this paper.

AdaThink-Med: Optimizing Inference-Time Compute for Medical Reasoning via Uncertainty Quantification Thinkless: LLM Learns When to Think

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T13:51:50.649862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:51:50.649862Z digest=sha256:8751593f2acb7984b7ce2a95f57f364b9424139d338384b82f90fd9ce08151d8

Observation 060c60f2-3ee7-43e3-9add-38dba8d67927 · inbound

Verifying Meta-Awareness via Predictive Rewards in Reasoning Models cites this paper.

Verifying Meta-Awareness via Predictive Rewards in Reasoning Models Thinkless: LLM Learns When to Think

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:50:30.535082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:50:30.535082Z digest=sha256:df3f2b963aed8a34e82061ace0361bdd93085552dfc07a6e9873eca9d0e4abbc

Observation 6ca94dd9-b465-4765-ae7d-c22765a14997 · inbound

Probing the Difficulty Perception Mechanism of Large Language Models cites this paper.

Probing the Difficulty Perception Mechanism of Large Language Models Thinkless: LLM Learns When to Think

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T11:17:47.339495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:17:47.339495Z digest=sha256:9c675d8638ca3cdc2dcd39a87c974ffd53cec00225459e6dbc7c0b15ec5ceec0

Observation 30502e58-4101-4174-ab32-97238c0d9876 · inbound

MixReasoning: Switching Modes to Think cites this paper.

MixReasoning: Switching Modes to Think Thinkless: LLM Learns When to Think

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-04T11:16:35.529003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:16:35.529003Z digest=sha256:cfc4c1e297d738e8e763d8aabcdedfafa13186019e9033ebec7cbeb3d4680283

Observation 04c5c3d1-5b28-40bb-a0d6-0a850f68ae6a · inbound

Rectifying LLM Thought from Lens of Optimization cites this paper.

Rectifying LLM Thought from Lens of Optimization Thinkless: LLM Learns When to Think

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T02:38:53.803631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-17T02:38:12.769600Z digest=sha256:92e7d752e2d996f37eb83ddb6e2746abd9f93a0e8cc1d7ca3b0a638ca839a534

Observation 87ae3283-f0da-420f-b089-ec5dbb03a541 · inbound

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers cites this paper.

Mid-Think: Training-Free Intermediate-Budget Reasoning via Token-Level Triggers Thinkless: LLM Learns When to Think

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T11:16:00.360878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:16:00.360878Z digest=sha256:6e6b6b8d2532ea4357b8f73e693ac694b69df51d81049e958a58b33976da8b5f

Observation 8e2c4434-3505-44c1-ba00-5bedda487da8 · inbound

ConPress: Learning Efficient Reasoning from Multi-Question Contextual Pressure cites this paper.

ConPress: Learning Efficient Reasoning from Multi-Question Contextual Pressure Thinkless: LLM Learns When to Think

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-03T05:43:53.202735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T05:43:53.202735Z digest=sha256:348aff8cfaf17b2d9cc212e2800efe7c180adfbe9c5d5b07d2fb0e6499a9aa74

Observation b1130b5d-0ecd-4eee-b58f-8a9e13aace82 · inbound

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression cites this paper.

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression Thinkless: LLM Learns When to Think

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-21T13:44:11.433304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T13:43:51.127429Z digest=sha256:f93d7516df84f9c6332d161a79362cc2ffcfd39a150fc608455aed6d78ff1873

Observation 42201804-09fc-49d4-b13d-f17754432bd8 · inbound

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression cites this paper.

Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression Thinkless: LLM Learns When to Think

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T03:25:22.412034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:25:22.412034Z digest=sha256:bbcd82f6b1f6b2512e69e40f60fc3da24195ec1d77b727f8d649b8c1262ccf5d

Observation 6f49898b-951b-4d7f-9318-04d2fb1d0a32 · inbound

Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning cites this paper.

Compress the Easy, Explore the Hard: Difficulty-Aware Entropy Regularization for Efficient LLM Reasoning Thinkless: LLM Learns When to Think

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-02T20:44:44.540785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:44:44.540785Z digest=sha256:3091c6663c97e60ab9c98b8f9fd4633b25c9932e51651ad3a58793754709be3c

Observation df110d54-c29b-4a9f-8b7f-73924b8abdf2 · inbound

Shorter, but Still Trustworthy? An Empirical Study of Chain-of-Thought Compression cites this paper.

Shorter, but Still Trustworthy? An Empirical Study of Chain-of-Thought Compression Thinkless: LLM Learns When to Think

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:18:01.247867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-13T17:16:47.284564Z digest=sha256:9842db397164e199961b12d23c0c27e74c61e99ae94b8130718bf62b96a4728a

Observation 27b12289-0739-46de-8e32-76732aec461b · inbound

Learning to Interrupt in Language-based Multi-agent Communication cites this paper.

Learning to Interrupt in Language-based Multi-agent Communication Thinkless: LLM Learns When to Think

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:25:53.825958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T19:08:45.818851Z digest=sha256:aaf1ebbcd5072dca78e01ba8a06343887c0cee534575769c59bdfda5e51f805e

Observation 23ba1d01-8a16-434a-a672-35a0a99b5740 · inbound

When Less is Enough: Efficient Inference via Collaborative Reasoning cites this paper.

When Less is Enough: Efficient Inference via Collaborative Reasoning Thinkless: LLM Learns When to Think

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:41:42.578914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-09T19:27:04.267404Z digest=sha256:ee39c675c5969c1f326bb1650d671482a0646fa4bacad05ced2480e17e2415fe

Observation 0732046b-2927-46aa-9cc6-ea73bb53c093 · inbound

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning cites this paper.

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning Thinkless: LLM Learns When to Think

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:34:40.994678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-22T06:33:36.846345Z digest=sha256:28a94f39afbe76c92cd55a6213381ee0ac6ccebdd47f34f414265a98c133bdf0

Observation 28ce41e8-a0a0-4d6a-80db-0917324f2e7f · inbound

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning cites this paper.

From Reasoning Traces to Reusable Modules: Understanding Compositional Generalization in Language Model Reasoning Thinkless: LLM Learns When to Think

Reference 71

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:38:55.923992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T01:13:11.483599Z digest=sha256:cc202ebe080ba752e02c69f27681fc4c8c055d66f88e65531f0979c05708cd5b

Observation 7b4dd15a-7162-40f7-bc02-d32b33428e20 · inbound

Finding the Time to Think: Learning Planning Budgets in Real-Time RL cites this paper.

Finding the Time to Think: Learning Planning Budgets in Real-Time RL Thinkless: LLM Learns When to Think

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-04T13:29:51.952467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T05:08:19.504454Z digest=sha256:56bf7b4f050e5eab7d7252213def110aca06e393a359bc7e525963230269de81

Observation c98d59f4-8e26-401e-84f8-f3af8e19d670 · inbound

Finding the Time to Think: Learning Planning Budgets in Real-Time RL cites this paper.

Finding the Time to Think: Learning Planning Budgets in Real-Time RL Thinkless: LLM Learns When to Think

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:34:34.893647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T09:26:31.944405Z digest=sha256:65f061db1dad5a1f729fd01190bb43dc32fa117ea62384051af64943b1789b23

Observation 2ffaa845-bf7a-4845-80b9-3816d5477890 · inbound

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning cites this paper.

When to Plan: Learning to Select Between Reactive Control and Deliberative Planning Thinkless: LLM Learns When to Think

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-01T21:04:09.787016Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:04:09.787016Z digest=sha256:cc037b488d9987e904a69d75e4831f313dded76f68982dc0c57dd01984a7d5e9

Observation 51e39b87-e1fb-45dd-968a-7e15bdbab9f5 · inbound

AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning cites this paper.

AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning Thinkless: LLM Learns When to Think

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T17:24:13.842512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T17:24:13.842512Z digest=sha256:bc21c5b1fd8c0a06ceba65be8ef3b3563a13391e8ff146bee385881e07eb84ee

Observation 4e249b7d-bd6a-40e0-871e-0056f67c682c · inbound

Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning cites this paper.

Toward Skill-Native LLMs: Skill Entropy for Benchmarking and Training Long-Horizon Reasoning Thinkless: LLM Learns When to Think

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T04:30:39.238053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:30:39.238053Z digest=sha256:7527226e70997417176d2bd5656a23c395f37d8e1524485c97b8ad7540dab337

Observation 521a6c05-246e-4cea-aafb-9239b739d344 · inbound

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling cites this paper.

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling Thinkless: LLM Learns When to Think

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-12T14:10:45.409749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T14:10:45.409749Z digest=sha256:d4dd6288bfa21388ab023fcc5de2c5ec7b8eb5b91cb6557dfb210041f1d83a8d