Pith. sign in

Paper Citation Record · LEDGER

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning

As of 20 August 2026, this Paper Citation Record lists 20 of 20 outbound references and 1 inbound Pith citation observation for arXiv:2507.18122.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.18122 v1

Coverage vector

measured 20 of 20 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:43:53.459773Z

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T04:48:31.084050Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-11T04:48:32.388628Z

Reference resolution

20 of 20 outbound references displayed

  • verified exact0
  • verified fuzzy4
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a15ac6c0-6e00-44c6-a210-91eb0c19da9c · outbound

This paper cites The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning The Unreasonable Effectiveness of Entropy Minimization in LLM Reasoning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:52.327498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:52.327498Z digest=sha256:dc89639e72011bb336cee8219105e33d4a78a56eef9ed556078335e2337df22a

Observation cfba05aa-50a3-4f52-af4f-a09d79aacc65 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Training Verifiers to Solve Math Word Problems

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:52.552476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:52.552476Z digest=sha256:8b7a4bb09337c8e4ab6e8820fe797e8b0f3c68c957bff53d091caa376b523970

Observation 7f73b8d6-5b7a-4b5c-b1bf-4b822c840406 · outbound

This paper cites The First Few Tokens Are All You Need: An Efficient and Effective Unsupervised Prefix Fine-Tuning Method for Reasoning Models.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning The First Few Tokens Are All You Need: An Efficient and Effective Unsupervised Prefix Fine-Tuning Method for Reasoning Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:52.762807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:52.762807Z digest=sha256:bbdd9d25752ad11954a24542b5ae0cf764a250dac981e11145979cd49e8a56c5

Observation 97bdacca-4e0c-40f5-aade-522e854e6551 · outbound

This paper cites Scalable best-of-n selection for large language models via self-certainty.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Scalable best-of-n selection for large language models via self-certainty

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:52.767584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:52.767584Z digest=sha256:f43ef1379f31fc53c2a1d97ff692f16ec660b7a2c78a91c5067a3a8623570cbf

Observation 42b667d2-358f-497d-b686-0e1928dc395a · outbound

This paper cites This is likely because we only ever perform few gradient steps on top of the base model, and thus catastrophic forgetting does not occur.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning This is likely because we only ever perform few gradient steps on top of the base model, and thus catastrophic forgetting does not occur

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:43:54.007293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T14:43:53.459773Z digest=sha256:0749dc12ba29f030d8bc6551d380b6f4058cfda5a5eeff89efa4bade40a9be74

Observation bfc426df-37f3-4055-877d-698b692486d3 · outbound

This paper cites Confidence Is All You Need: Few-Shot RL Fine-Tuning of Language Models.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Confidence Is All You Need: Few-Shot RL Fine-Tuning of Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:52.966250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:52.966250Z digest=sha256:f29a3b2ca2775ee11b08eda76579cb802cb289904c274e15fdfd5d5e8838f508

Observation 7674c89e-55f4-45ef-843d-a13a5366a41c · outbound

This paper cites Maximizing Confidence Alone Improves Reasoning.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Maximizing Confidence Alone Improves Reasoning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:53.050703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:53.050703Z digest=sha256:d65cacc4af56fc3f109a335a05910ca6274a2b922941f9af9b81d364fe08832e

Observation c8e74a0b-eb49-4bc3-8944-0735c0ff6bcf · outbound

This paper cites Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Direct Reasoning Optimization: Token-Level Reasoning Reflectivity Meets Rubric Gates for Unverifiable Tasks

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:53.202824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:53.202824Z digest=sha256:91f168329b5d67c635557bc221cb2f28ce5793a825d590d80325ffc9827a0b56

Observation 7b8425a9-9ec3-4a30-bb7f-dd8adec0dc74 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:53.286899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:53.286899Z digest=sha256:4cf103cefcee1726770b03bbf3c80a81f892dcbca7a9a8ac4de8d97b85d53740

Observation 27fc7542-72c9-4dd0-94a4-291aacb9c434 · outbound

This paper cites Absolute Zero: Reinforced Self-play Reasoning with Zero Data.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Absolute Zero: Reinforced Self-play Reasoning with Zero Data

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:53.345503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:53.345503Z digest=sha256:d2a4583116b4d37fa2d02d598a3e75678b3a1e1ed3098e3fc319b06e0c163d0a

Observation acdc95ec-eef2-4b19-8106-6f28a91c8166 · outbound

This paper cites Self-adapting language models.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Self-adapting language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:53.446619Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:53.446619Z digest=sha256:9534a981a268ef855d5bf5b03a372bb57a35887c8718bac9150aceb7aa18e856

Observation 31f970a2-3cf1-41c7-8c0b-f3ef5babd73a · outbound

This paper cites A Additional Confidence Measures • Negative entropy (token-level) − 1 n ∑n i=1 H(π(yi | x, y<i)) =− 1 n ∑n i=1 ∑V j=1 π(yi = j | x, y<i) log π(yi = j | x, y<i).

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning A Additional Confidence Measures • Negative entropy (token-level) − 1 n ∑n i=1 H(π(yi | x, y<i)) =− 1 n ∑n i=1 ∑V j=1 π(yi = j | x, y<i) log π(yi = j | x, y<i)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:43:54.045748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T14:43:53.450939Z digest=sha256:7149c065d7e7b10859abe325ea2f970c890b7ab63197619cb35b3ea04fa4c854

Observation 77d7fea7-9e0b-4014-ba66-8eefbc581c99 · outbound

This paper cites Since the SGD optimizer uses less memory and compute, we continued using SGD for the rest of the experiments.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Since the SGD optimizer uses less memory and compute, we continued using SGD for the rest of the experiments

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:43:54.024139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T14:43:53.455364Z digest=sha256:a96d367b61bd3675d034e18353be20b68ccce2d23c1a4292d06e628aad7fdbcc

Observation d9fb6dec-c749-409a-a738-dfbb35ae1374 · outbound

This paper cites Dynamic evaluation of neural sequence models.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Dynamic evaluation of neural sequence models

Reference 2015

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T14:43:54.062319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T14:43:52.780354Z digest=sha256:5e21eda06fae3eb66d40edad89ee0fcaf8646c5326a746d9745ee1a1f475380d

Observation 8bea7096-5765-490a-b22e-fa3868e39214 · outbound

This paper cites Dynamic Evaluation of Transformer Language Models.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Dynamic Evaluation of Transformer Language Models

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:52.841235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:52.841235Z digest=sha256:bff03e39ac8061e5ef5e074d880e3986ac7dbe9a59fc7735b6694cb276682429

Observation 01bcdfc3-ed00-44d6-a6c4-7d5a244f1fd2 · outbound

This paper cites Self-Improving Transformers Overcome Easy-to-Hard and Length Generalization Challenges.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Self-Improving Transformers Overcome Easy-to-Hard and Length Generalization Challenges

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:52.912793Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:52.912793Z digest=sha256:02c1125451183d06a7b37015ee9aed3e8b07acfd0310f37c8c781773c57a4288

Observation 56935697-518f-45f0-baba-7c650617d06e · outbound

This paper cites Learning to (Learn at Test Time): RNNs with Expressive Hidden States.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Learning to (Learn at Test Time): RNNs with Expressive Hidden States

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:53.124912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:53.124912Z digest=sha256:57422f78f74174b5bafa3520e0ccbb71babe43aa75084192bcb37f4df20905be

Observation 792599a0-d512-40dc-aea2-acf4974531e8 · outbound

This paper cites One-Minute Video Generation with Test-Time Training.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning One-Minute Video Generation with Test-Time Training

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:52.690948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:52.690948Z digest=sha256:10aa6b012af6ef4030d796cb47a567564842ec60ba06ed6163ca5f73f9eb8a7f

Observation f4ab2f1a-28d1-45fc-9b71-6c201380e3f5 · outbound

This paper cites Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Right Question is Already Half the Answer: Fully Unsupervised LLM Reasoning Incentivization

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:53.296475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:53.296475Z digest=sha256:dc4aac91b97d50239e4c0566d11e6903158fb4ef70429a026591cc00344c0560

Observation b2ec1165-c56b-4de8-86c4-693909690b61 · outbound

This paper cites Local Mixtures of Experts: Essentially Free Test-Time Training via Model Merging.

Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning Local Mixtures of Experts: Essentially Free Test-Time Training via Model Merging

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:52.414007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:52.414007Z digest=sha256:a8dedf4d25a3094c68fc62f0837cc196fd634b48a00361064b317f33cb006f9b

Pith citing papers

Observation 8276fbe6-03b1-40f4-85e1-edec7c33838d · inbound

Consilience for Verifier-Free Test-Time Scaling cites this paper.

Consilience for Verifier-Free Test-Time Scaling Maximizing Prefix-Confidence at Test-Time Efficiently Improves Mathematical Reasoning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-11T04:48:32.395054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T04:48:31.084050Z digest=sha256:309913b0e47c064177f367ec250273390f765629127b2435d94e948f0003d6ad