Pith. sign in

Paper Citation Record · LEDGER

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning

As of 10 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.04771.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04771 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T17:07:02.537106Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3f7401b1-e66e-4a35-82f9-b95cfafdcffa · outbound

This paper cites xKV: Cross-Layer KV-Cache Compression via Aligned Singular Vector Extraction.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning xKV: Cross-Layer KV-Cache Compression via Aligned Singular Vector Extraction

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.541065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.541065Z digest=sha256:7135019c00880ebc512374e38e4fa6b63c15009b7e785b14831874408f4b9858

Observation 23052acf-81c3-4cd9-a423-071ef0845d51 · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.657571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.657571Z digest=sha256:369ddb389c49dec8c46f9c6e4c7ef397b1f9b291add5b1f2b5d818ee71aad32f

Observation fc845da2-56de-4b98-bd43-00d02756bb98 · outbound

This paper cites InFindings of the Association for Computational Linguistics: ACL 2025, 21539–21564.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InFindings of the Association for Computational Linguistics: ACL 2025, 21539–21564

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:04.195709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:07:00.768651Z digest=sha256:1fc197f2c642d454b4041c90e999e32305b0eb719a799693ce98290ad79b5ff9

Observation 2eb70d07-703f-49c1-860c-461436066125 · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.917179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.917179Z digest=sha256:b0c41b5aefb338aa72a32736f2be79a0b6bb98bc3d32ac6217fdaf8283bf342f

Observation a0cfa119-adc9-4e40-ae43-174ab5186bf9 · outbound

This paper cites InFindings of the Asso- ciation for Computational Linguistics: ACL 2026, 17051– 17064.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InFindings of the Asso- ciation for Computational Linguistics: ACL 2026, 17051– 17064

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:03.795551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:07:01.067843Z digest=sha256:fc6c7699a99e178e6dfb98ed4cbe7af0c6e6492a28b14a63407717bfdaf51587

Observation 49a984e6-6c44-4cab-b23a-d2eec4bfe0fe · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.139839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.139839Z digest=sha256:af50dac84622772cf3bdda704197352a99580124ea687b555caed4ac3c05fbe9

Observation 4910617c-f844-4da1-8e5b-cac80449faa1 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.200943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.200943Z digest=sha256:e712a6918a453276ec689dc316a1398577e717372b26ce7dbd6d8aa34320c116

Observation 1c088ec5-3752-46d1-85fc-26178c07292f · outbound

This paper cites Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Hierarchical Sparse Attention Done Right: Toward Infinite Context Modeling

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:07:03.397949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:07:01.285318Z digest=sha256:c59fd718a221eb9d7c3cd12f857ca1bba8fd039a75e42075c8c312f1ef48b56c

Observation e5d6b42c-067e-41e5-9e1e-efc67620bace · outbound

This paper cites OpenAI o1 System Card.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning OpenAI o1 System Card

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.346455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.346455Z digest=sha256:eabfb139476570fbdaec28929480e1528db1f99ab9b6857d7e80928b503aa47d

Observation 4a0002bf-6639-46c2-bd20-44e0d630bdf3 · outbound

This paper cites Concise Thoughts: Impact of Output Length on LLM Reasoning and Cost.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Concise Thoughts: Impact of Output Length on LLM Reasoning and Cost

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.424636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.424636Z digest=sha256:f466010956c1b249491f5d958efb73a94054f5c907c8a818da6e9640eaad2dfc

Observation 59fc4700-df1c-4469-b7c4-14d4160378d4 · outbound

This paper cites Not All Thoughts are Generated Equal: Efficient LLM Reasoning via Multi-Turn Reinforcement Learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Not All Thoughts are Generated Equal: Efficient LLM Reasoning via Multi-Turn Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.503753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.503753Z digest=sha256:2792645c5aef4c58170e188e23455bca61f08857f2cc1eb0c726f77fa0390a32

Observation db03dc27-dd94-410c-aad5-6827c9ad12df · outbound

This paper cites RouteLLM: Learning to Route LLMs with Preference Data.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning RouteLLM: Learning to Route LLMs with Preference Data

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.589061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.589061Z digest=sha256:137892037fea2e472ce879303c9a4cb52c5c7c899c2c49a2e3606a50a432d7ee

Observation ebdb3aa2-b3ba-4ebc-b39e-6308be8b0f6a · outbound

This paper cites InProceedingsofthe2025ConferenceonEmpirical Methods in Natural Language Processing, 8021–8040.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InProceedingsofthe2025ConferenceonEmpirical Methods in Natural Language Processing, 8021–8040

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.673043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.673043Z digest=sha256:b29071bd2622213bdf468bab35335186c461030b361fb62153b7f30793c88c31

Observation 805f982d-50e4-4eb7-acbc-6268eb00bdc9 · outbound

This paper cites The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning The Benefits of a Concise Chain of Thought on Problem-Solving in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.843310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.843310Z digest=sha256:9d06c03e578f5b8df3b0adac5257d18022d72238fdd00f7b6ab7328f4864cfa9

Observation 3d84b4c5-e307-46b6-ab4b-e2182e7b85a1 · outbound

This paper cites Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Towards Efficient Large Language Reasoning Models via Extreme-Ratio Chain-of-Thought Compression

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.912817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.912817Z digest=sha256:a834d52b98009f793767aede1889fddd7fa553e18cb7422f9e47e31ec4e130e0

Observation c3d48be3-8fe0-4a3b-af0c-ffec19f407d7 · outbound

This paper cites D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning D2O: Dynamic Discriminative Operations for Efficient Long-Context Inference of Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.982112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.982112Z digest=sha256:9e87fab500465b533be1479979e56ec12fd30769c73d80085ca6847bb16981ec

Observation 549eccfd-5e7e-4c97-8cf8-65d628511f39 · outbound

This paper cites Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.070191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.070191Z digest=sha256:6b112fcb51e3b2e18e12c25b19c70a3c9cb978000fc7c05fc49994283c072fd9

Observation e0cfc844-8e03-42d8-bbfe-f86823a3b740 · outbound

This paper cites Model Tells You Where to Merge: Adaptive KV Cache Merging for LLMs on Long-Context Tasks.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Model Tells You Where to Merge: Adaptive KV Cache Merging for LLMs on Long-Context Tasks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.153112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.153112Z digest=sha256:d3db381dea428368a5c06110e7f564b6d20cc818110c6a2fc7396f199a80ff64

Observation 93959c88-548e-4514-83e4-ab896f40bf64 · outbound

This paper cites IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning IndexMem: Learned KV-Cache Eviction with Latent Memory for Long-Context LLM Inference

Reference 24

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:07:02.935451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:07:02.216605Z digest=sha256:f277dba49ac635b705cecc87c32c98a453fd90bfc7834414f4c86f0892b1c4a8

Observation 3f881354-a945-4269-a6ce-f1f8beb34ef0 · outbound

This paper cites Demystifying Long Chain-of-Thought Reasoning in LLMs.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Demystifying Long Chain-of-Thought Reasoning in LLMs

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.311998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.311998Z digest=sha256:ee8d2112ba33e437fe52963ef216741ded7bf8a380319828cd8dd593c2250a05

Observation b09b1b72-a8b6-4ee0-b50d-d90f19642436 · outbound

This paper cites TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning TurboQuant: Online Vector Quantization with Near-optimal Distortion Rate

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:02.376649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:02.376649Z digest=sha256:0c0a8cc138ce6f8eebb3c34cee7b8c8d8a0da55f2e1d0530b96c391f3f6ba764

Observation 6f0be876-e02f-4e36-a750-4c5dbe3c5d05 · outbound

This paper cites Done Is Better than Perfect: Unlocking Efficient Reasoning by Structured Multi-Turn Decomposition.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Done Is Better than Perfect: Unlocking Efficient Reasoning by Structured Multi-Turn Decomposition

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T17:07:02.699059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:07:02.442534Z digest=sha256:94847d154a52238f4a9ee69bab83508d9eee8ad0786a40ef8be39194994e743e

Observation 39209269-5347-455c-a632-467254c6e068 · outbound

This paper cites InForty-firstinternationalconferenceonmachine learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning InForty-firstinternationalconferenceonmachine learning

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:03.659023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:07:02.537106Z digest=sha256:fed6f48700cd1ae4869948e8b6a9b9da1359c7703ecab1e6b11502e52b5466bc

Observation c13a16ea-9bbc-4ca1-8f42-ebcf4ab8af63 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.834739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.834739Z digest=sha256:85f2c394ba61a726836aebf4935b5c9c9ec222107811637d5f984d55722354c6

Observation a614b709-e4d9-4874-87f2-c23eaaf79da8 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:01.757657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:01.757657Z digest=sha256:4aa84b4a3b262848b9bab5d0d3b4af741aa6fc9d0aec9b2131a38c89a1593c18

Observation f19a1cce-233f-4218-b01c-f50a35e21906 · outbound

This paper cites PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning PyramidKV: Dynamic KV Cache Compression based on Pyramidal Information Funneling

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.437098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.437098Z digest=sha256:29f5277eae06101fb0e7d83939d22537f4ff7becae72707c6a39f0e948deb96a

Observation 7d523cba-690a-46f3-9833-5a0d50c287b0 · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T17:07:00.309946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:07:00.309946Z digest=sha256:ffc025d6dc795c89129243e1cdbe1590cbf3bdb8f51ddbca59eb60ea36d19d55

Observation 5db40434-4347-4daa-9515-279f3e5efbb0 · outbound

This paper cites Fu,Y.;Chen,J.;Zhuang,Y.;Fu,Z.;Stoica,I.;andZhang,H.

Fewer Tokens, Smaller Cache: Reward-Coordinated Efficient Reasoning Fu,Y.;Chen,J.;Zhuang,Y.;Fu,Z.;Stoica,I.;andZhang,H

Reference 2026

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T17:07:03.996112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-06T17:07:01.005716Z digest=sha256:a224cbef11a284c07bcf831911d0a9f5879b4bf419af9938f2e2e1bcce2eaae8

Pith citing papers

No inbound Pith citation observations are available.