Pith. sign in

Paper Citation Record · LEDGER

OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2311.09724.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.09724 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T20:42:44.448540Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T01:42:19.143537Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9956b270-e8c7-4a25-bc1f-1de8cec547a5 · inbound

Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations cites this paper.

Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:34:15.839774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-14T22:34:15.638114Z digest=sha256:5e8872c43bb124d8ed58581324a6a4350205df7dd9b1f6f15d452d3966b9fac1

Observation dda4c43d-4eda-4702-894c-c170c1b47d5e · inbound

Improve Mathematical Reasoning in Language Models by Automated Process Supervision cites this paper.

Improve Mathematical Reasoning in Language Models by Automated Process Supervision OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:53:45.980330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:53:45.878221Z digest=sha256:d1d68cd91af7b9ca4906dbe825a4744f6620579eb3574ca5394af376415057d0

Observation 36265d67-bc8c-45c6-a7fa-360777269ff2 · inbound

Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning cites this paper.

Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-21T01:42:19.147154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T01:42:19.004468Z digest=sha256:5505a600fa5b4e02c6cfa363b63259b218a4c7fca40b0654f0ceeb373750f7bf

Observation deb6154c-f8e9-43c0-b3a3-4d8d2e812720 · inbound

Reward-Guided Speculative Decoding for Efficient LLM Reasoning cites this paper.

Reward-Guided Speculative Decoding for Efficient LLM Reasoning OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-09T20:42:44.448540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:42:44.448540Z digest=sha256:59cdafd204f3af903d853943dbe84254db232a5faf4f6b2dd43e2507b9a74c81

Observation 3a378e7c-2876-4756-ab69-67ca96a89215 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 183

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.251438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:ef30085a45ed391bec5da1b8fa08ad9344d14b13846d747bd08874617aced773

Observation c1e2bd89-95e1-4960-a808-d574518db6c9 · inbound

SCOPE: Compress Mathematical Reasoning Steps for Efficient Automated Process Annotation cites this paper.

SCOPE: Compress Mathematical Reasoning Steps for Efficient Automated Process Annotation OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:39:19.866278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:39:19.866278Z digest=sha256:da44ec4137f4bde2226601a8db4dba69a503c85a3de140052969ebfb9f78745b

Observation bd0674c3-a3c1-4f74-af92-f26678f2827c · inbound

ProgRM: Build Better GUI Agents with Progress Rewards cites this paper.

ProgRM: Build Better GUI Agents with Progress Rewards OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:37:48.213972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:37:48.213972Z digest=sha256:efff264c4e1c4e387adce3be9e5a6336dc6a8ce8889d86f27a233cdd8e64f9d6

Observation c048bd41-ef28-4f99-888c-ac0ddaf58a90 · inbound

UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents cites this paper.

UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T13:34:16.142382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:34:16.142382Z digest=sha256:50b111daf0a7ee00dc3ad3b07c9d923b1093686f39f5ef0c7ff447f4a61e8436

Observation d6a07ceb-aa06-48d0-9f29-7c11db37d6ad · inbound

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset cites this paper.

BMMR: A Large-Scale Bilingual Multimodal Multi-Discipline Reasoning Dataset OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T20:15:52.784527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:15:52.784527Z digest=sha256:791013b1803fbc485efb66e92680a7322f9bee8b30dee4cb14a23b187e613343

Observation f3bb7a7e-4fed-436d-96e9-e780a1c9cff7 · inbound

Goldilocks RL: Tuning Task Difficulty to Escape Sparse Rewards for Reasoning cites this paper.

Goldilocks RL: Tuning Task Difficulty to Escape Sparse Rewards for Reasoning OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:40:20.797484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T21:38:56.393217Z digest=sha256:1bcec38b2461de621b7b067ce0bc238d74952f09254bf1d36940c4a2194eb2d8

Observation 96475b4b-1752-4fe7-a134-21bb3150d19c · inbound

Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents cites this paper.

Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:25:35.763541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T12:22:13.551709Z digest=sha256:9fb1e67d143a938462914f1657a585bd89f94fb88049a6bedd20bfbc16cce3a7

Observation 4bcc19fd-c171-4037-ac66-9ee1108858cb · inbound

Process Supervision of Confidence Margin for Calibrated LLM Reasoning cites this paper.

Process Supervision of Confidence Margin for Calibrated LLM Reasoning OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:41:12.177273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-08T08:19:09.437464Z digest=sha256:6f01e03b6a7c11221c3e898a8333150387e0c0c6faec7023a0bac56a2ec0da79

Observation daf0c399-4c67-45a7-b326-e87d8f4b3b52 · inbound

Reducing Credit Assignment Variance via Counterfactual Reasoning Paths cites this paper.

Reducing Credit Assignment Variance via Counterfactual Reasoning Paths OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-21T00:49:18.793365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T00:45:41.280228Z digest=sha256:ad67504ef592cd802b3861cd9866e8028f13ff5c7f385ba86e7b3058f429948b