Pith. sign in

Paper Citation Record · LEDGER

MathPile: A Billion-Token-Scale Pretraining Corpus for Math

As of 15 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2312.17120.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.17120 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:39:36.932384Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 46132507-3758-4f0b-abf3-937c8fd2b8a6 · inbound

DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models cites this paper.

DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models MathPile: A Billion-Token-Scale Pretraining Corpus for Math

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-24T03:23:49.429902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-05-24T03:23:18.827351Z digest=sha256:f87697c44d36d74921bc83552a1128680f1ac78e76d4c519ce2352b04903c3e5

Observation e5b6f45e-af7c-439c-9150-1561a4ece9ba · inbound

Mars-PO: Multi-Agent Reasoning System Preference Optimization cites this paper.

Mars-PO: Multi-Agent Reasoning System Preference Optimization MathPile: A Billion-Token-Scale Pretraining Corpus for Math

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T10:39:36.932384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T10:39:36.932384Z digest=sha256:3d758c6a78e41d3adb8eceb70c7c0cd02c4b12de544cf331dd7f3e3e4da07449

Observation b4ae5a33-8175-4d6d-96f5-c889bb51f00a · inbound

AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling cites this paper.

AceMath: Advancing Frontier Math Reasoning with Post-Training and Reward Modeling MathPile: A Billion-Token-Scale Pretraining Corpus for Math

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T11:43:24.586940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T11:43:24.586940Z digest=sha256:31e3a0a06007f8cff6202bf20e5a990afcba808837e625cfb6117fd5f1492648

Observation 5c89650b-33bd-4992-a534-de617d6b196d · inbound

Dynamic Skill Adaptation for Large Language Models cites this paper.

Dynamic Skill Adaptation for Large Language Models MathPile: A Billion-Token-Scale Pretraining Corpus for Math

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T00:46:46.300740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T00:46:46.300740Z digest=sha256:c188e0888f4cba46a5d9cd96169f714c0bafd94f63f96cf74a6028914244a575

Observation a1faefd3-5a9d-4ca9-9eff-a40f8258ebac · inbound

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems cites this paper.

MathFlow: Enhancing the Perceptual Flow of MLLMs for Visual Mathematical Problems MathPile: A Billion-Token-Scale Pretraining Corpus for Math

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-22T22:57:13.299730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-22T22:55:34.238427Z digest=sha256:b4849bfefb0ae97dc8f77f867b362719cf97d738a466a9c254b122e8d9078140

Observation 3b4612a2-3425-4484-9dae-1f571c030c6a · inbound

MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning cites this paper.

MDPO: Multi-Granularity Direct Preference Optimization for Mathematical Reasoning MathPile: A Billion-Token-Scale Pretraining Corpus for Math

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:29:35.680196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:29:35.680196Z digest=sha256:99169caac95be0e91e17dd199ef5699a683d93a90ef09c2dab54bc210c3449f5

Observation 91e9def5-e837-48b2-b972-09dd171b7349 · inbound

StepHint: Multi-level Stepwise Hints Enhance Reinforcement Learning to Reason cites this paper.

StepHint: Multi-level Stepwise Hints Enhance Reinforcement Learning to Reason MathPile: A Billion-Token-Scale Pretraining Corpus for Math

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-06T20:26:47.613196Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:26:47.613196Z digest=sha256:fe39cbc2e2c54bb929818fdb6baeacfcd980feece1035cf9c5aa09a4e0d544b4