Pith. sign in

Paper Citation Record · LEDGER

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning

As of 10 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2505.17829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17829 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:20.428126Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cbb5158a-e8f9-4d14-893a-453804d038da · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:15.860789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:15.860789Z digest=sha256:b9629174e14bcef6c45aa0ff282df5df81b26e19e7dab5968f01a8fd22e67718

Observation d84d6c05-a592-4a30-9678-7e5de1b18f6f · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:15.928935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:15.928935Z digest=sha256:572ee8ffb35a25e6f3ea4372db7ece00e0bd2dee1bfd08b457dbb9e3227f6193

Observation 8198a83c-174c-47bb-9380-71fc6c8467c9 · outbound

This paper cites Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:15.995858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:15.995858Z digest=sha256:64e39f136bb4b6c7ccb143b21e580a9537fe8f19d5f8c3fc4e92de226edb5d81

Observation bf5dd1f9-059f-4b2f-83db-a84b73358037 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.052650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.052650Z digest=sha256:f9ad4f656a8d36f83a070d3849403133249e70dbb63019ee5921e2538e8e825d

Observation 5408ac83-e892-4b72-8be3-502414775fe8 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.157711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.157711Z digest=sha256:97055aa78b427337e53f2de90b6c2f9f4df38d17f4f70733839458aa3235d4c5

Observation 622a08fe-2874-4516-9be6-0009b06f3410 · outbound

This paper cites An Empirical Study on Eliciting and Improving R1-like Reasoning Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning An Empirical Study on Eliciting and Improving R1-like Reasoning Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.258476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.258476Z digest=sha256:670135b9a700457c173f76cbc30de3ece25f45b2fd35408c7f0ade20acaa53ed

Observation f1958060-a26f-4663-849c-d564bdc09acb · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Training Verifiers to Solve Math Word Problems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.355876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.355876Z digest=sha256:8e53a322ba5a3b8e63f677cdad491f2ced02b2b3b4fb9f3669d6104e43e475b6

Observation cc788a5d-fee3-4fcf-a0f5-5b710919b473 · outbound

This paper cites Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.471846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.471846Z digest=sha256:4e2fc4aeb47afb117e3eb314fc4d300242fa76bf04f0e367df5b9bde2428c491

Observation 94ccc782-0e46-4550-b340-3a0efbd2a360 · outbound

This paper cites The Llama 3 Herd of Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.567297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.567297Z digest=sha256:61ab03bcb2b1aac30b493407583db38e048defa12b2bfb256c9cc85ab90d4438

Observation 547bb6f3-c113-4139-be22-5ddb1d03a8c4 · outbound

This paper cites Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.667323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.667323Z digest=sha256:52e0fe3d65aa62851abde278a65a39364cf3afc488822caf111f3944be849a34

Observation 38b3018e-0145-4e90-a9ac-f5ecea67b504 · outbound

This paper cites rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.786796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.786796Z digest=sha256:f8e3562cc92823b955a1bee4c44d881b11de2a788a9130ba3b07bf027b8acb9e

Observation 91b7c566-b336-4bef-aefc-352dc56b0f58 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.870979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.870979Z digest=sha256:cf046bb2b84a9b465df6d1d3bbdd5915a691a5b1213ccc694cf00b73ce4ccd96

Observation a45e6103-0b37-4c7a-ace9-db5f77a7673e · outbound

This paper cites Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.967173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.967173Z digest=sha256:433677c574025b7ad3d93f338e594309d25725bea8f6d9cd4b9feb6e396fac7d

Observation 69c03bf8-35d6-4f45-a1d2-c4ae7fd881da · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.034269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.034269Z digest=sha256:421b6a0617554abf4357029629ad49300a33d5a3d4d738c716c02d26e79a6b99

Observation 4c3263cb-f17e-41de-bd6a-ca488c2b2211 · outbound

This paper cites ETS: Efficient Tree Search for Inference-Time Scaling.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning ETS: Efficient Tree Search for Inference-Time Scaling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.115091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.115091Z digest=sha256:850809e97b6fc5adf1be6f6ee160965b18607d6aa9d236cff73bc5c47856f278

Observation 895124a2-d20a-4fc3-8433-9edc9e6e7bd9 · outbound

This paper cites Efficient Test-Time Scaling via Self-Calibration.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Efficient Test-Time Scaling via Self-Calibration

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.220177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.220177Z digest=sha256:c7689d4f8a5821c68e10d2a9755b98976493e3cf39aa0316e3b698f86ddf42b8

Observation b0e05e90-6c61-41af-995c-45701c748fc2 · outbound

This paper cites A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.297639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.297639Z digest=sha256:3a256c4bd3b670849694440b1b87d5e6ad3419c305563c4a58538de7ba1e1e89

Observation 9f9ffe20-ea0b-4f37-9e95-257e32a45eae · outbound

This paper cites Enhancing LLM Reasoning with Reward-guided Tree Search.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.406362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.406362Z digest=sha256:e0d7e01c5b7ebe1d822efd84bc122c3bc9e866dfa6e79df761dbaf81dfd3421b

Observation 1a881508-053e-43ab-953e-65baf416e635 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.487613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.487613Z digest=sha256:b4f7b13446f8cfac975d109b6a582b91ce8a02f74f2355aaad0b40cf98139898

Observation b25cc179-2cb9-4684-876d-15fefb5f7993 · outbound

This paper cites Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.591164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.591164Z digest=sha256:92a48776480fd3c53dff3ebabb568d9194b5744dd590b2a659769205c65df706

Observation 21c2f839-5f79-4d08-a540-8d2ff69b315b · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.671386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.671386Z digest=sha256:e39d65b99d04ea8ef0161d1f4b6424f0501925862f7041e15f911b2b6c1dfd0a

Observation cec42c1f-e556-47dc-b89a-fa22ea3aa623 · outbound

This paper cites Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.775098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.775098Z digest=sha256:f2a520cd67b8a5bba7a0c605672722541888354ba2cc0a3442e072d7b7809187

Observation 1249619d-e278-4318-80bc-9206c8487c21 · outbound

This paper cites Improve Mathematical Reasoning in Language Models by Automated Process Supervision.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Improve Mathematical Reasoning in Language Models by Automated Process Supervision

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.861908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.861908Z digest=sha256:8e73efb2396bfa47aa9a0f9f1bffc9ff7180865575198641a35e5602f1356907

Observation 91847d96-bfdd-4c5d-82ce-097a4fdeb616 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:21.696466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:45:17.962098Z digest=sha256:3dbc22c9e527bbfcf063baf3aaa60546db80d1e9ee4e0f17333eefc8766752f6

Observation f3feb280-534b-4aff-be6f-08be76e2c893 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.067650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.067650Z digest=sha256:fe07dacd4d23a2920fa0e6bf902bcb9755f72a2c1198e21f6cff4334cead5641

Observation 8205ddc8-fe73-4af5-90ba-658f20f51cca · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.171836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.171836Z digest=sha256:215608c0d7dfb10ecd851ee0c4b47049a20ac697d26e5d75f42f7d157637b2ab

Observation eb92f12a-5360-4e3d-840b-43ff23e88e91 · outbound

This paper cites Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.293852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.293852Z digest=sha256:ccfb61a993debd7ae82a5dcbe5807de2c71c86340c6cc57dbc0762b40883845a

Observation 8e2d8aad-fc61-4f3d-927e-56495a8c0ea2 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.388703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.388703Z digest=sha256:ffe279aadb15a87f21a913e05af0829dcbaae5fca1339805b0bfc458fb7dd30e

Observation d1c1ddc5-266a-468e-8699-bb896de78f2f · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.474040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.474040Z digest=sha256:45e3544fe06221fa559cec3a2bed234b9a2776600022a5b7c67afa6c95e50fd5

Observation 77f4f98b-17b4-4dde-9671-f1a15288d3a4 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Solving math word problems with process- and outcome-based feedback

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.564018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.564018Z digest=sha256:b3632c983cc4a97da638136a8e18aabe003e3acd461cdfd168dce85406db080e

Observation a15fb25e-acc2-4aea-84cd-154f347a5f9e · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:21.514394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:45:18.664750Z digest=sha256:1491fdbfc3e505fcd2dabeadc04c74f3bfb3e09b8b2390a6b6b73153fc83840b

Observation 3aa87878-559e-499f-a8ea-fbc5ec7b7e47 · outbound

This paper cites OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.769402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.769402Z digest=sha256:0720786ed37f2060d55d930051998d180304eb8f13166d37a000c5a3c2b6f25d

Observation db8b369f-b0ab-4444-9dac-49822359ca93 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.860070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.860070Z digest=sha256:59a47f9027416e2430cea95fa0cef0a8dd5e5d96ca56f37a3d81de4c71d898b0

Observation a085871c-a7ba-4de6-806d-5b693720461d · outbound

This paper cites Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.952147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.952147Z digest=sha256:f977e7a4c6cb2bd423fb02a9792cdc8dc2ddee3f32911991c431d64b4b709847

Observation 86aacc1c-ec73-49e0-b39a-bd72eb478738 · outbound

This paper cites Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.024357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.024357Z digest=sha256:5fc27781c3bdb99193c042fc5511caf6c6bb852637f6a3d8d57dcb2a10a25654

Observation 38e75dee-09bc-464f-9e5b-dbbb4e5cc9e8 · outbound

This paper cites Chain-of-Probe: Examining the Necessity and Accuracy of CoT Step-by-Step.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Chain-of-Probe: Examining the Necessity and Accuracy of CoT Step-by-Step

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:45:20.790255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:45:19.073344Z digest=sha256:dffd8953b40539a9dec05719d761eba840338329ee4c687fa26a1a7000768908

Observation 834c6eb3-761f-43ec-b151-317162dcc95a · outbound

This paper cites Chi, Quoc V.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Chi, Quoc V

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.142078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.142078Z digest=sha256:90a61860c27f5000d042e9bc8e97ef7b524659daa352c17bf3b533b5230fc749

Observation e76e0fb6-1f3d-40d6-b397-352717e997df · outbound

This paper cites Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.217464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.217464Z digest=sha256:331bdd55f5424ed110238df0e7b041e8b0594c431e01f0503f925518ed93cf0a

Observation 984c6ccb-4398-4e84-a8da-5fdbe806bc49 · outbound

This paper cites A Comparative Study on Reasoning Patterns of OpenAI's o1 Model.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning A Comparative Study on Reasoning Patterns of OpenAI's o1 Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.282901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.282901Z digest=sha256:95f003f05bdda1657daa60c3f961558d96873b114d9288c98dcef6749e13f422

Observation 10c14ed8-4a82-4f5e-923e-154a71b61934 · outbound

This paper cites When More is Less: Understanding Chain-of-Thought Length in LLMs.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning When More is Less: Understanding Chain-of-Thought Length in LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.333659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.333659Z digest=sha256:055df7d2577ff006fe01c543cf1097b9be9ad5ef9f7454ceed55812e75efb7ea

Observation e23b3e40-b19e-44cb-9332-01a0262f1503 · outbound

This paper cites Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.417118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.417118Z digest=sha256:0d741765a34ea487aa8a0f1a27e837b483db433614c8531960074ddd19f958c7

Observation 13e43f31-d822-4b72-a7f6-63725eaf9c21 · outbound

This paper cites Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.472575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.472575Z digest=sha256:c34b505702b283c24c52733b574ccc4b8bf7330bf75b6c7f5ef7eb23e4030104

Observation 85cb7881-06b8-4663-b4dc-993a664390ef · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.536322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.536322Z digest=sha256:c28604d3038087e148cac9588df73c921892a94332ca0afd322b08d434c141ac

Observation 0aee6a39-b94b-4531-8ad2-82b8a53dfab7 · outbound

This paper cites Qwen3 Technical Report.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Qwen3 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.631357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.631357Z digest=sha256:85c96bda73c3fc7bd83277e6728e5ee11ac9f5fade2d25fd5a6df49339867bd4

Observation 62608014-6a3b-4aaa-9240-9b01f50e2ffb · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:21.351151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T14:45:19.734663Z digest=sha256:9455a726a0e6cbe484c3dd0a97c01d095f27483fc223fa7729bc6a39cfa9ec23

Observation 74de4af0-8adf-4a44-9744-56519d94e536 · outbound

This paper cites B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.802170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.802170Z digest=sha256:1481c92514012fac05ebb910e4bb6daa0bb6bf08d783bfa857697c2b024ad469

Observation fb9a8d92-a6fb-4231-9b1a-804b31ec998f · outbound

This paper cites Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.883202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.883202Z digest=sha256:346eee485326be08901e3e79d3b7e87b559fae2f633df2b0eb98931dd673cb7e

Observation f03cbe38-d485-4e5e-8883-c58344b01d95 · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.980989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.980989Z digest=sha256:7c31600cbd12fb28e990655f0022a7d39b0d5aa697b0ee19e61c9a7dfeda98a4

Observation 9ac2fd92-d22d-4ebe-a714-758e178760ec · outbound

This paper cites A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.078930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.078930Z digest=sha256:62ca2f2152b4fde96403d2b981874dca3a73f4144c5ccd686dd7e496a6a0fb82

Observation b5839cb0-47d4-4ac5-98dd-348c6f1c80cd · outbound

This paper cites Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.163992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.163992Z digest=sha256:cf7cd52ab96d475c1305584f9f9f4a787c991f7a772626e560c129795bbd9a04

Observation 2e2d339d-9926-4eb1-a780-9707b0305fe9 · outbound

This paper cites ProcessBench: Identifying Process Errors in Mathematical Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning ProcessBench: Identifying Process Errors in Mathematical Reasoning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.276917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.276917Z digest=sha256:58b9af2c58f2e21449622fcae17f03a562ef9eff15f340cbe522783800690e8f

Observation 9cc3677d-eeab-4864-a7b2-214235b9b7de · outbound

This paper cites online" 'onlinestring :=.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning online" 'onlinestring :=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.356452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.356452Z digest=sha256:7356dbb5d83cff3cb484bd65f834afd20f6d2494017d786637f992995fbe289f

Observation 938ded77-216d-48cf-82da-478b2e82f1c6 · outbound

This paper cites write newline.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning write newline

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.428126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.428126Z digest=sha256:d30cc1f59760dd41d5a34486218657278daf02bb3bd23f76b9c2feabf6116943

Pith citing papers

No inbound Pith citation observations are available.