Pith. sign in

Paper Citation Record · LEDGER

AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2412.15084.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.15084 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 16 of 16 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:33:37.634222Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T00:04:22.516001Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a560f3f3-4675-482a-a41a-9db7deaa3965 · inbound

Process Reinforcement through Implicit Rewards cites this paper.

Process Reinforcement through Implicit Rewards AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 87

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T20:23:30.950535Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-11T20:23:30.763794Z digest=sha256:cfc63b792030185842d54932fe8cf2268fad3eeee77a5b00bac757dace886a7e

Observation d74ba07c-0741-4af1-9acd-5d3812101dba · inbound

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning cites this paper.

Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-16T12:47:10.237446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T12:47:10.146795Z digest=sha256:28e455cf58d6935eff42c3850089a52661ac683580039de03b1142dbb5272156

Observation df7555d8-7f06-4fdf-a615-6036640f0082 · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:32:01.241571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:01025ea0d7751b0e30f5bd0eb74e0398db473a9c87ef0e78bb4e346820ecd9e2

Observation a15bbb19-5bf1-4a84-b846-63bc33d71447 · inbound

EasyMath: A 0-shot Math Benchmark for SLMs cites this paper.

EasyMath: A 0-shot Math Benchmark for SLMs AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:33:37.634222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:33:37.634222Z digest=sha256:7f35b92cdc36af001303fd4b8518e2707db75ccd58ec1e87f01b128f95e747fe

Observation 943f7d5d-3417-4c90-a906-f22ca245594a · inbound

AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning cites this paper.

AceReason-Nemotron: Advancing Math and Code Reasoning through Reinforcement Learning AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:06:35.406551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:06:35.406551Z digest=sha256:38dd91360ba8e8390394616e7d5a53a6999914262f2994a35cec078f8560a9dd

Observation 0be86bb8-dca2-4b9f-9e6b-4ee186768982 · inbound

Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem cites this paper.

Unleashing the Reasoning Potential of Pre-trained LLMs by Critique Fine-Tuning on One Problem AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:13:53.714891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T11:13:53.714891Z digest=sha256:2eb88a9e24a098e999b58eb82ff58158dd4bcb92f5e710392e1eb902c0e5ff87

Observation 30b95365-5af2-40c3-b92b-c5556d29a205 · inbound

Improving Large Language Models with Concept-Aware Fine-Tuning cites this paper.

Improving Large Language Models with Concept-Aware Fine-Tuning AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:28:23.918291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:28:23.918291Z digest=sha256:668165971710c1b1127099d1e65338e2ec464b9ced8faa7d2d92f1edf1371262

Observation d6db3524-6140-4563-bca6-9ab0e975239c · inbound

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy cites this paper.

AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T00:41:35.455218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:41:35.455218Z digest=sha256:42f64ac897698a0d0604169a1c0ec4d94f0a5ed8030821bbdaf1308f756e919d

Observation 5b3d8c42-f4b6-4829-a6fe-a0399f6c3e87 · inbound

OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique cites this paper.

OpenCodeReasoning-II: A Simple Test Time Scaling Approach via Self-Critique AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:11:17.943363Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:11:17.943363Z digest=sha256:20e195c6f249d69a8c50b3b583c143b4d0e42fe6dfc5f4d723f80aaa20f3b420

Observation 0f27a9ea-c53b-4e9c-b200-791d9693c070 · inbound

JT-Math: A Multi-Stage Framework for Advanced Mathematical Reasoning in Large Language Models cites this paper.

JT-Math: A Multi-Stage Framework for Advanced Mathematical Reasoning in Large Language Models AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T14:11:52.807336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:11:52.807336Z digest=sha256:3ffc6e472732e2605d7cc35f8a05ffc9d882546c66a5b2f2d253374c406dc8d9

Observation 2831f5a5-a1e4-45f5-803c-555773a4587d · inbound

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance cites this paper.

Falcon-H1: A Family of Hybrid-Head Language Models Redefining Efficiency and Performance AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T11:44:05.055041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:44:05.055041Z digest=sha256:f3b5ff783489be6ef53bc190e8525fed5d2cfa3fbf0e5ff09e887eff0112db64

Observation 760542d6-a34c-4a94-a573-b9607593e220 · inbound

The Majority is not always right: RL training for solution aggregation cites this paper.

The Majority is not always right: RL training for solution aggregation AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T23:02:32.090666Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T23:02:32.090666Z digest=sha256:7bf2552262296193d55c78c64dd56c08dc4d7224b14c2234d1d4d610364c209b

Observation 7ad69638-11ba-4189-969a-275c8af045aa · inbound

Fine-Tuning Small Reasoning Models for Quantum Field Theory cites this paper.

Fine-Tuning Small Reasoning Models for Quantum Field Theory AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-10T03:24:14.972224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T03:23:18.770963Z digest=sha256:274685e25b6f1864ccbc5e1589b9897c469005a7a0a61f29fbb5dc404459b324

Observation ade2d6c3-aeec-4041-9d28-85dd780f31e9 · inbound

ZAYA1-VL-8B Technical Report cites this paper.

ZAYA1-VL-8B Technical Report AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 172

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:21:23.382878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T01:15:16.607346Z digest=sha256:bcc186832d14a6af21226901ba6f70e7c213dcf779b01244ddcf3607a9760746

Observation bb3f1db2-5ab2-42bc-8764-8d2ea171608a · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 154

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T00:04:22.517246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-07T23:59:38.702609Z digest=sha256:7184b60d117029aa3967b17827b4ffd8c3c452e5e1ac6f6848b02b81b4683bf3

Observation f6725e72-865c-4171-8287-ee9018f9808e · inbound

Unified Audio Intelligence Without Regressing on Text Intelligence cites this paper.

Unified Audio Intelligence Without Regressing on Text Intelligence AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling

Reference 154

Resolution
unresolved
no resolver link, observed 2026-07-11T07:46:49.059192Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T07:46:49.059192Z digest=sha256:ac934532a3436f652bbb3c45ce59445144b216dbdf09508d2b8632a303e78c46