Pith. sign in

Paper Citation Record · LEDGER

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning

As of 10 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 0 inbound Pith citation observations for arXiv:2505.17829.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17829 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:20.428126Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

53 of 53 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation cbb5158a-e8f9-4d14-893a-453804d038da · outbound

This paper cites L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning L1: Controlling How Long A Reasoning Model Thinks With Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:15.860789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:15.860789Z digest=sha256:33b7b0c834d20229cb0aa3b718aeabbceed1e9941231402627efe2ab28dc464f

Observation d84d6c05-a592-4a30-9678-7e5de1b18f6f · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:15.928935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:15.928935Z digest=sha256:17d4ae4cfaa50ce121051b096f7a510a0c1d0d23b656928af4342a266c648e2f

Observation 8198a83c-174c-47bb-9380-71fc6c8467c9 · outbound

This paper cites Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Forest-of-Thought: Scaling Test-Time Compute for Enhancing LLM Reasoning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:15.995858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:15.995858Z digest=sha256:5bdc8e3a434064890cdb9712c8beb1e05b6d36182d4380d4d1ab34c6f7b7c70f

Observation bf5dd1f9-059f-4b2f-83db-a84b73358037 · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.052650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.052650Z digest=sha256:736d62b1c21c24ce661b6874eb9d9db783fe6ab3f062c162ac70d915ae8f7af9

Observation 5408ac83-e892-4b72-8be3-502414775fe8 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.157711Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.157711Z digest=sha256:0c669c7a80c0be889f80b4d332dab64168f8d143978571bdaf23fa4a9bf7beea

Observation 622a08fe-2874-4516-9be6-0009b06f3410 · outbound

This paper cites An Empirical Study on Eliciting and Improving R1-like Reasoning Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning An Empirical Study on Eliciting and Improving R1-like Reasoning Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.258476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.258476Z digest=sha256:974fcb481d1b9d2128d436bafb8425ac453a7dab8c1f315e15ab6f39ee4e74f6

Observation f1958060-a26f-4663-849c-d564bdc09acb · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Training Verifiers to Solve Math Word Problems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.355876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.355876Z digest=sha256:cb2a354da51a8d4c1ba20bcbbd6968b823752c21318a336365197caa7723c268

Observation cc788a5d-fee3-4fcf-a0f5-5b710919b473 · outbound

This paper cites Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Rethinking External Slow-Thinking: From Snowball Errors to Probability of Correct Reasoning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.471846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.471846Z digest=sha256:0805f19d3e446700d8c1032c02f464a67e43b2d15b5ae069e865d0c601e41282

Observation 94ccc782-0e46-4550-b340-3a0efbd2a360 · outbound

This paper cites The Llama 3 Herd of Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning The Llama 3 Herd of Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.567297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.567297Z digest=sha256:9c78f4c416a722454a7794850f66d4288be805cb8ef84beaa863b52efe1aefe2

Observation 547bb6f3-c113-4139-be22-5ddb1d03a8c4 · outbound

This paper cites Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Search, Verify and Feedback: Towards Next Generation Post-training Paradigm of Foundation Models via Verifier Engineering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.667323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.667323Z digest=sha256:59fc79fa2e3bec640012e4d20c3be7ece07fb1de63f8bc645d714f02fa602e1d

Observation 38b3018e-0145-4e90-a9ac-f5ecea67b504 · outbound

This paper cites rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning rStar-Math: Small LLMs Can Master Math Reasoning with Self-Evolved Deep Thinking

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.786796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.786796Z digest=sha256:bfd76f87d9e9f1dc8e346ce458e54f5788d78fcfd4a786e3b5db63afedf6258f

Observation 91b7c566-b336-4bef-aefc-352dc56b0f58 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.870979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.870979Z digest=sha256:4c8a68d5c415abde43111fe0a99c4da057135a203756c51838f85e4826761c0e

Observation a45e6103-0b37-4c7a-ace9-db5f77a7673e · outbound

This paper cites Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:16.967173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:16.967173Z digest=sha256:bd7c75a75077a028bedf437513d05d72194029b0ef4468a400ab984241a92577

Observation 69c03bf8-35d6-4f45-a1d2-c4ae7fd881da · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.034269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.034269Z digest=sha256:a8968787f57c3701e9aa4096c48c10c327c94c0ca39b7513d7a64ca5f325868f

Observation 4c3263cb-f17e-41de-bd6a-ca488c2b2211 · outbound

This paper cites ETS: Efficient Tree Search for Inference-Time Scaling.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning ETS: Efficient Tree Search for Inference-Time Scaling

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.115091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.115091Z digest=sha256:105c8944165c35f9b3042355a84367a6b6305090a65a427fdd3baad1af2a6b8f

Observation 895124a2-d20a-4fc3-8433-9edc9e6e7bd9 · outbound

This paper cites Efficient Test-Time Scaling via Self-Calibration.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Efficient Test-Time Scaling via Self-Calibration

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.220177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.220177Z digest=sha256:04a53967228cc6384f1aeeed3260ab25470a1c57825de32f40facca9d01dc7af

Observation b0e05e90-6c61-41af-995c-45701c748fc2 · outbound

This paper cites A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning A Survey of Test-Time Compute: From Intuitive Inference to Deliberate Reasoning

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.297639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.297639Z digest=sha256:ed256e80767de1c382cba907f97fc9ca5bbf414de5751772659c0593b3d19748

Observation 9f9ffe20-ea0b-4f37-9e95-257e32a45eae · outbound

This paper cites Enhancing LLM Reasoning with Reward-guided Tree Search.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Enhancing LLM Reasoning with Reward-guided Tree Search

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.406362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.406362Z digest=sha256:d7cd51a5b4f04e8da223363304cb53e095c1c5c78ff744ccb525a7e90a004d01

Observation 1a881508-053e-43ab-953e-65baf416e635 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.487613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.487613Z digest=sha256:450838499b52c9766e719d139ca0b58e7aaab9c6fb4bc7aeea282638ab32a76e

Observation b25cc179-2cb9-4684-876d-15fefb5f7993 · outbound

This paper cites Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Escape Sky-high Cost: Early-stopping Self-Consistency for Multi-step Reasoning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.591164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.591164Z digest=sha256:376446267dab6b05ee3ce183870b5b065c45b538bc50f04499d2ec16eb39042d

Observation 21c2f839-5f79-4d08-a540-8d2ff69b315b · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.671386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.671386Z digest=sha256:2a16035cf7c8492096b3393a1f72100e2f66c364afb0614eb05f0ba3e6a9343e

Observation cec42c1f-e556-47dc-b89a-fa22ea3aa623 · outbound

This paper cites Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.775098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.775098Z digest=sha256:cd172fa766156544378c0e65d87b2788cec33a6d1ee3f4f13d365936431e4421

Observation 1249619d-e278-4318-80bc-9206c8487c21 · outbound

This paper cites Improve Mathematical Reasoning in Language Models by Automated Process Supervision.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Improve Mathematical Reasoning in Language Models by Automated Process Supervision

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:17.861908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:17.861908Z digest=sha256:65dee7af89dacd4f9e32327b9f97507e5c6314625184b11d633d4743aee04e2a

Observation 91847d96-bfdd-4c5d-82ce-097a4fdeb616 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:21.696466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:45:17.962098Z digest=sha256:f41236355517a1326f638c604514aa726b4c6956b221899b048473785b414d27

Observation f3feb280-534b-4aff-be6f-08be76e2c893 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.067650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.067650Z digest=sha256:5d5fc518aa203f866572a15a2cb619a978f376c01a2c413804d7b70c5d87b26a

Observation 8205ddc8-fe73-4af5-90ba-658f20f51cca · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.171836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.171836Z digest=sha256:2eed5db4dd9e7da556679ed10d4f451737f30f8bc4654725927734ad792a1fb0

Observation eb92f12a-5360-4e3d-840b-43ff23e88e91 · outbound

This paper cites Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.293852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.293852Z digest=sha256:ee33e3d16d937ba98321074b389dcc7d7794a684ddb55f16201c022bd0b81982

Observation 8e2d8aad-fc61-4f3d-927e-56495a8c0ea2 · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.388703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.388703Z digest=sha256:a04d36a5b2dca10987674babd6ce1c9a9525ee3c255466c1c4e69d3bd7073aaf

Observation d1c1ddc5-266a-468e-8699-bb896de78f2f · outbound

This paper cites Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Stop Overthinking: A Survey on Efficient Reasoning for Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.474040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.474040Z digest=sha256:8145dc1600bed71bd26848cb104db607d7a00399fae93ae115da50c227a3cf1c

Observation 77f4f98b-17b4-4dde-9671-f1a15288d3a4 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Solving math word problems with process- and outcome-based feedback

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.564018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.564018Z digest=sha256:bab7f560798aeb50dc7b2d770759ac0e98eed11f85e166d53dd954458158d015

Observation a15fb25e-acc2-4aea-84cd-154f347a5f9e · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:21.514394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:45:18.664750Z digest=sha256:1d39f6a43d9416bec8786ce739607d682657be07fec9a1a08b8ca70a2d002ec8

Observation 3aa87878-559e-499f-a8ea-fbc5ec7b7e47 · outbound

This paper cites OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.769402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.769402Z digest=sha256:fb80ee62ba4f0a6021a252055e7659478451e8f34bf8c1d75210ddb87c6ac351

Observation db8b369f-b0ab-4444-9dac-49822359ca93 · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.860070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.860070Z digest=sha256:6729c8a6bc4b8dc25f0b816e363856404c752f2416696cb38b85980c34ddb7de

Observation a085871c-a7ba-4de6-806d-5b693720461d · outbound

This paper cites Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Towards Self-Improvement of LLMs via MCTS: Leveraging Stepwise Knowledge with Curriculum Preference Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:18.952147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:18.952147Z digest=sha256:46aaa94fb51e734261c0a106f88c6cd41af73f7b362c265443acf8464e217c14

Observation 86aacc1c-ec73-49e0-b39a-bd72eb478738 · outbound

This paper cites Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Chi, Sharan Narang, Aakanksha Chowdhery, and Denny Zhou

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.024357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.024357Z digest=sha256:1442455cea4b9557822c2a321844938b3cec9ed502cd955b40996d3f6cd4b456

Observation 38e75dee-09bc-464f-9e5b-dbbb4e5cc9e8 · outbound

This paper cites Chain-of-Probe: Examining the Necessity and Accuracy of CoT Step-by-Step.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Chain-of-Probe: Examining the Necessity and Accuracy of CoT Step-by-Step

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:45:20.790255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:45:19.073344Z digest=sha256:0a76b3e927f0f9b92c6202aa4fd730fb8a81b88a87fd6cb86d1614883faa6bdb

Observation 834c6eb3-761f-43ec-b151-317162dcc95a · outbound

This paper cites Chi, Quoc V.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Chi, Quoc V

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.142078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.142078Z digest=sha256:ca1ce98869c0e21ef4e58decbe6ea0e0cab305ad5ae01b8e245453f6b686fdff

Observation e76e0fb6-1f3d-40d6-b397-352717e997df · outbound

This paper cites Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Beyond Examples: High-level Automated Reasoning Paradigm in In-Context Learning via MCTS

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.217464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.217464Z digest=sha256:6897fb01aa1ddac4559238e97fb3a7cfe0e94ba3d3d4a13c34e9e841039ff21c

Observation 984c6ccb-4398-4e84-a8da-5fdbe806bc49 · outbound

This paper cites A Comparative Study on Reasoning Patterns of OpenAI's o1 Model.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning A Comparative Study on Reasoning Patterns of OpenAI's o1 Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.282901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.282901Z digest=sha256:20d8f605f5b207c65f3efb89efaa25b12d41d472d45dad58be510d27cc3e5ba0

Observation 10c14ed8-4a82-4f5e-923e-154a71b61934 · outbound

This paper cites When More is Less: Understanding Chain-of-Thought Length in LLMs.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning When More is Less: Understanding Chain-of-Thought Length in LLMs

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.333659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.333659Z digest=sha256:6bc19631645e7d43699f72ac20d72b9a9e6d8c6a449f168c5448a6fcf6c8251a

Observation e23b3e40-b19e-44cb-9332-01a0262f1503 · outbound

This paper cites Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Enhancing LLM Reasoning via Critique Models with Test-Time and Training-Time Supervision

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.417118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.417118Z digest=sha256:5136232a489b203c16cbef519879fa28730f8ef6533d257a440d41b5b89fbaab

Observation 13e43f31-d822-4b72-a7f6-63725eaf9c21 · outbound

This paper cites Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Towards System 2 Reasoning in LLMs: Learning How to Think With Meta Chain-of-Thought

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.472575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.472575Z digest=sha256:946ead569194a0e95e026a8db5adb4ab662ed2b34b8f59b3837ca8b133e146ed

Observation 85cb7881-06b8-4663-b4dc-993a664390ef · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.536322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.536322Z digest=sha256:ea19f5b67b7d179ea2d5fc6377ab71c9a73373bedc852fe2182b3aac75a1eacd

Observation 0aee6a39-b94b-4531-8ad2-82b8a53dfab7 · outbound

This paper cites Qwen3 Technical Report.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Qwen3 Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.631357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.631357Z digest=sha256:11f9484589ac91783fe0bc927097be7014c76f9f780308c87330170c8b83fdc0

Observation 62608014-6a3b-4aaa-9240-9b01f50e2ffb · outbound

This paper cites an unresolved cited work.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:21.351151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:45:19.734663Z digest=sha256:4082d5d5a1163e668b5818215128eb942ae458e61641101b6ac3769114a6f2e3

Observation 74de4af0-8adf-4a44-9744-56519d94e536 · outbound

This paper cites B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning B-STaR: Monitoring and Balancing Exploration and Exploitation in Self-Taught Reasoners

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.802170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.802170Z digest=sha256:59f1e672861d356bef4097685079f61523c06578a9e3441f42023e40f265d806

Observation fb9a8d92-a6fb-4231-9b1a-804b31ec998f · outbound

This paper cites Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Accessing GPT-4 level Mathematical Olympiad Solutions via Monte Carlo Tree Self-refine with LLaMa-3 8B

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.883202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.883202Z digest=sha256:757686eab4758fbe964c9ec0cb3e2a364b0b7f5a0b5eb784bb5bb0fc18198c4b

Observation f03cbe38-d485-4e5e-8883-c58344b01d95 · outbound

This paper cites Generative Verifiers: Reward Modeling as Next-Token Prediction.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Generative Verifiers: Reward Modeling as Next-Token Prediction

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:19.980989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:19.980989Z digest=sha256:b3baae0a0cd37c0412a6aeafa5bc92b236e64ff247ca31272c140f06fc4bb385

Observation 9ac2fd92-d22d-4ebe-a714-758e178760ec · outbound

This paper cites A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning A Survey on Test-Time Scaling in Large Language Models: What, How, Where, and How Well?

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.078930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.078930Z digest=sha256:7715d912b781059ae69b41ea0bae5a487d576efe209ab0f0f28e315ec59734c8

Observation b5839cb0-47d4-4ac5-98dd-348c6f1c80cd · outbound

This paper cites Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.163992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.163992Z digest=sha256:473e252dae4e31e6101ef99f251be6cb4aadaa8ab8f3544d8dcdf36e7b768126

Observation 2e2d339d-9926-4eb1-a780-9707b0305fe9 · outbound

This paper cites ProcessBench: Identifying Process Errors in Mathematical Reasoning.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning ProcessBench: Identifying Process Errors in Mathematical Reasoning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.276917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.276917Z digest=sha256:35a89bda1bd95684840e5c42664df0fa255b2028e649d4477dcedc674add1379

Observation 9cc3677d-eeab-4864-a7b2-214235b9b7de · outbound

This paper cites online" 'onlinestring :=.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning online" 'onlinestring :=

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.356452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.356452Z digest=sha256:d14a0e5b78e551a060c2c26ee93ad9067b0ce997d72ef63d4834f2f3016e4122

Observation 938ded77-216d-48cf-82da-478b2e82f1c6 · outbound

This paper cites write newline.

Stepwise Reasoning Checkpoint Analysis: A Test Time Scaling Method to Enhance LLMs' Reasoning write newline

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:20.428126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:45:20.428126Z digest=sha256:5ba8e6043ea0c8cc5b47ba65927fa6a3739007eafcc8fc3c97498ff91c820d60

Pith citing papers

No inbound Pith citation observations are available.