Pith. sign in

Paper Citation Record · LEDGER

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning

As of 11 August 2026, this Paper Citation Record lists 21 of 21 outbound references and 2 inbound Pith citation observations for arXiv:2502.00271.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.00271 v1

Coverage vector

measured 21 of 21 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T19:40:23.148446Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T17:29:36.978463Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T10:37:01.586875Z

Reference resolution

21 of 21 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 744e3e67-edb8-492d-bfa8-7df699e64bd7 · outbound

This paper cites write newline.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.069698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.069698Z digest=sha256:5e761487ac26baaafe6e08ad0985c450404cd9175db749c70c1a94aed0694626

Observation a5849c0e-2c78-44af-a983-d54180a1beff · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.074525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.074525Z digest=sha256:e9de839999a729f9adbcb8549550ea6edba2cdb1f9e3712ac976bfaaeebabcb8

Observation 85ff0d55-9264-4da6-917e-1d6b3c42b054 · outbound

This paper cites AlphaMath Almost Zero: Process Supervision without Process.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning AlphaMath Almost Zero: Process Supervision without Process

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.078834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.078834Z digest=sha256:6f9889630a7bc122df225602577882d0bf5eb49a512274b2ab78c4bddb3749bf

Observation 3e9559b4-f024-43c5-a1dc-e18742691358 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Evaluating Large Language Models Trained on Code

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.084259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.084259Z digest=sha256:00d04f98395d114eb284e3b1c5e172f810343351db5be72cbc9dd44e567f291e

Observation 883e0dc8-ed61-4fd2-935b-44b01fd060ce · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Training Verifiers to Solve Math Word Problems

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.088985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.088985Z digest=sha256:4f78ee1951fd80db83be9a0fbc8c250717f38b98a2dbbd119463f9d86d575ce8

Observation 93709917-bd33-4624-9e93-69c7db90de06 · outbound

This paper cites J., Wang, Z., Wang, D.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning J., Wang, Z., Wang, D

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.092881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.092881Z digest=sha256:a0a19254b3834ea091ee5a90d61c444ed838aeadf2aee96e8fb85ef4826d18d7

Observation ba09c96f-2afa-4c7d-a301-a4757b22fd5d · outbound

This paper cites Measuring mathematical problem solving with the MATH dataset.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Measuring mathematical problem solving with the MATH dataset

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:40:23.557217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:40:23.096843Z digest=sha256:b778c256ee8834f1f7877d89f466e3b1accd543d5fa178919a85c638be334498

Observation 0e88601e-a161-439b-bb66-ee0815d695ed · outbound

This paper cites Mistral 7B.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Mistral 7B

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.100597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.100597Z digest=sha256:7e6d5940778dd4c411c6c84401106a9938e0fdbfd6f07105427968c0a95fb24c

Observation ff44c0ef-ab54-4673-8528-658213c76270 · outbound

This paper cites H., Gonzalez, J., Zhang, H., and Stoica, I.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning H., Gonzalez, J., Zhang, H., and Stoica, I

Reference 9

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-09T19:40:23.489714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:40:23.104679Z digest=sha256:530308dccadd0723ef5a151145b2502281cc0e5dae55361bc80123b79f0f41f3

Observation 8e48135d-87d9-4bec-af31-c4db065f570c · outbound

This paper cites Let's verify step by step.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Let's verify step by step

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:40:23.546210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:40:23.108121Z digest=sha256:c1474dfe84a85c76e8a71254276c651b5fa04bcf4dbfba3d59aeee8bcd7ca75e

Observation aadf6fb9-04b0-4ee2-8b39-30aeef6e620f · outbound

This paper cites and Hutter, F.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning and Hutter, F

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.112056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.112056Z digest=sha256:dcdaadbc7d39d4e4da364427aac8637dcb7289b8a593a56506ec723f08c4c23c

Observation 549205b8-de25-4e71-ae92-1083ba1de86b · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.115838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.115838Z digest=sha256:e9196785b74fa5b2aba055235ec184f1e628a75feabf74ec33e1ac5e7ac16168

Observation 8a7779d7-5f82-4fbb-8b6c-2f3340ab25bb · outbound

This paper cites Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.119992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.119992Z digest=sha256:259adec830d77eba7f55e523fe58db447b62981f73337f123a453b60a8a4cb48

Observation c5f33746-c561-443d-820b-d1dac2d91379 · outbound

This paper cites Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Toward Self-Improvement of LLMs via Imagination, Searching, and Criticizing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.123889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.123889Z digest=sha256:f55d400a077b90ed37bfad69f547b9cfa183fa41b4ef243362f640dc54d0797b

Observation 0a65fbfb-aec7-4eb8-8dd8-78ba994b1e88 · outbound

This paper cites M., Wen, Y., Zhang, W., and Wang, J.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning M., Wen, Y., Zhang, W., and Wang, J

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:40:23.529149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:40:23.127555Z digest=sha256:5279ce80510954ebd7ff26f287c051cba20e619e3d32c01171241b07bb62b25c

Observation 0addff3a-e015-46a8-bace-23efde81ee20 · outbound

This paper cites Math-shepherd: Verify and reinforce llms step-by-step without human annotations.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Math-shepherd: Verify and reinforce llms step-by-step without human annotations

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.130971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.130971Z digest=sha256:45a1143f0c7236f7f701700272b49e3eafc5035b32b6f58d70b79be628f79fa7

Observation 033fa8bd-02a4-4f7c-aaeb-a549158c2163 · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.134519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.134519Z digest=sha256:6dc8dffd47c02cdab630aec0e56b40c2673f4dffcfaca5e7cc986d3b010e9c7f

Observation 108c777e-5545-419b-ad46-2e8ae8f06ef1 · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.138069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.138069Z digest=sha256:115b767f16bde8135a7e5f23471a4c6cf757174d50ff1db511d23dac34be17e2

Observation 27f6e7b4-c7cd-4c8a-99a2-6dc6050eb82b · outbound

This paper cites DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning DeepSeek-Prover-V1.5: Harnessing Proof Assistant Feedback for Reinforcement Learning and Monte-Carlo Tree Search

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.141252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.141252Z digest=sha256:42a500fbeea3d3edcd3a7560d338f615b9b1a8a6d2cad483faeb2e48b2f54960

Observation 6ca27096-e68d-43a5-86dc-72956e4c92ee · outbound

This paper cites Ovm, outcome-supervised value models for planning in mathematical reasoning.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning Ovm, outcome-supervised value models for planning in mathematical reasoning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T19:40:23.145024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T19:40:23.145024Z digest=sha256:37da0cc36e6d0d2aad8aa4d35ecd9d088f53e3a9a9dd927c0ae3cf31d5a362cc

Observation eda9567d-7538-4fd0-a6f4-4fd1e9d90385 · outbound

This paper cites M., and Polu, S.

Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning M., and Polu, S

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T19:40:23.518634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-09T19:40:23.148446Z digest=sha256:460fd1fed2c4a8a74d08cab44dcc940700fc3155fae6b77452bd6f7481af2a6b

Pith citing papers

Observation aa2d8285-e5e1-4178-9e02-f0c1cab4405f · inbound

Best-of-$N$ TTS Evaluation is Confounded by ASR Family Alignment cites this paper.

Best-of-$N$ TTS Evaluation is Confounded by ASR Family Alignment Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-10T10:37:01.588242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-10T10:31:23.170749Z digest=sha256:8d06dcd6e4742338d3c419e94cc2686094b3d9f2c57be8e2a9e7c34bf54e18ba

Observation 7de6432c-46c4-449b-aaab-3c62782a9a9c · inbound

Verify, Repair, Repeat, or Stop? Robust Stopping for Noisy Verify-Repair Loops in LLM Agents cites this paper.

Verify, Repair, Repeat, or Stop? Robust Stopping for Noisy Verify-Repair Loops in LLM Agents Scaling Flaws of Verifier-Guided Search in Mathematical Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T17:29:36.978463Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:29:36.978463Z digest=sha256:84407a0f27b85c2a844e1fe1df6b1824b9ecfc4ff4c281779243fce3f6733962