Pith. sign in

Paper Citation Record · LEDGER

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning

As of 11 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 1 inbound Pith citation observation for arXiv:2501.13622.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.13622 v4

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:49:44.048289Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T13:15:43.851865Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved16
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 30225360-2779-46b4-b106-f34e737d8793 · outbound

This paper cites URL: " 'urlintro :=.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning URL: " 'urlintro :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:43.968146Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:43.968146Z digest=sha256:3f93d60647857fb915943a5668dbbfc2eb0aea54edc37ff793b33cd2a7fcb131

Observation 34fad8ee-bea0-4340-818f-fcf21558aaf5 · outbound

This paper cites write newline.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:43.973589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:43.973589Z digest=sha256:c3f2704ccc516acad91d3eeab07cd76300b6a8d0d9128a2fe2b609e7f7a4506e

Observation f8461951-039b-4288-8846-adbdee47537f · outbound

This paper cites GPT-4 Technical Report.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:43.978622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:43.978622Z digest=sha256:100db7abdfffafed3271b668669b92efb0625281e7dc5aa2305aa5cdf8842589

Observation 60c2bd3b-5c17-4f89-b76e-d5d59204aa6d · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Training Verifiers to Solve Math Word Problems

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:43.984230Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:43.984230Z digest=sha256:ffa734861caaf045bb630f82fc1b1cb0cb610a78111a4073a9cc1aaf73e811e0

Observation ed242d56-0d17-457c-838f-be991575e9ad · outbound

This paper cites The Llama 3 Herd of Models.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning The Llama 3 Herd of Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:43.989758Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:43.989758Z digest=sha256:b1d4605fb4df7e859385d3fe4a97a2539a819ee102beaeda41b5f11f2f33e4de

Observation 934a5e0f-c222-4324-86dd-443c7b3077c5 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:43.994609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:43.994609Z digest=sha256:52c8786f4c9a3b1a1735eb56b027c31925f5799b99b4bb38f7ff9ea8556e9b7a

Observation 9ff8ff86-e8e8-4441-b94f-6aa4fc42a896 · outbound

This paper cites Large Language Models Cannot Self-Correct Reasoning Yet.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Large Language Models Cannot Self-Correct Reasoning Yet

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.000861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.000861Z digest=sha256:cbf7bf677682a9f9ab2c71a5228083192573b48378c0b8875150b476dc8d7c37

Observation adfd210a-640d-43a1-9e64-37347db2a49e · outbound

This paper cites Challenges and Applications of Large Language Models.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Challenges and Applications of Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.005994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.005994Z digest=sha256:13c23fe30571a4db45bab6bffcaa46fbf9d27f825b0eaf21a0761a0ecdc21d42

Observation 2c82e3f7-d0a7-4282-9afb-3a36ecf8df7f · outbound

This paper cites GSM-Plus: A Comprehensive Benchmark for Evaluating the Robustness of LLMs as Mathematical Problem Solvers.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning GSM-Plus: A Comprehensive Benchmark for Evaluating the Robustness of LLMs as Mathematical Problem Solvers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.011913Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.011913Z digest=sha256:494f8d21384dcb8b0ec25c19f9b1cf9ff29ebd0889620b71fa46976734382aba

Observation 71dbc219-e964-4d7b-b9f2-442416720d72 · outbound

This paper cites Process Reward Model with Q-Value Rankings.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Process Reward Model with Q-Value Rankings

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.016691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.016691Z digest=sha256:204d0ab337653f01118f5f2e342b5ec790e441952d21dc3c3c08d3d3cfe29741

Observation 0ee91a7a-9306-4f5b-8098-5870b2a83679 · outbound

This paper cites Let's Verify Step by Step.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Let's Verify Step by Step

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.021340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.021340Z digest=sha256:0bb4f9fcb922aacf8ee5b95eeb2a97d03a97d6fc20fb18fe87bba9b4f8144e91

Observation e32b0a2f-4069-4f64-a2d1-05e0f94038df · outbound

This paper cites Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.026345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.026345Z digest=sha256:5f6c6147565087427645b39671f061b8a9650f9ecf709f43b27ba20e6ed45a8e

Observation 32143455-1347-4f23-999b-6242152982c5 · outbound

This paper cites Solving math word problems with process- and outcome-based feedback.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Solving math word problems with process- and outcome-based feedback

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.033590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.033590Z digest=sha256:2cd95921f251f15445d84d0f27ebe426a666b8e2e37a7a7e30779520c26b47dc

Observation 1f472304-3412-4936-9631-e7a91b1b3ba7 · outbound

This paper cites an unresolved cited work.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Unresolved cited work

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.039314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.039314Z digest=sha256:c98276eb5575374e1d6e94a264cea74f1f413ef7446a75313f315031a6df4c11

Observation 2c2921d5-4880-4ef2-b8af-9d34493337ae · outbound

This paper cites Qwen2.5 Technical Report.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning Qwen2.5 Technical Report

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.043784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.043784Z digest=sha256:616421264d19aaeeaddba56624924461698fe23d2499f02aa6353bb62f023b67

Observation 5044d472-b652-4884-99f1-01dd711cd724 · outbound

This paper cites ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search.

Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning ReST-MCTS*: LLM Self-Training via Process Reward Guided Tree Search

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T15:49:44.048289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:49:44.048289Z digest=sha256:1f1d737704799bbc8253395ec112e885634d93edfc6da6f9ce7917594bad3d48

Pith citing papers

Observation e405932f-9075-4701-adf2-ceb39301c7c2 · inbound

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards cites this paper.

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T13:15:43.851865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T13:15:43.851865Z digest=sha256:6363108ef6489f1c423da83274199db72ade34f99d3e3b3d22cc32df7dafde40