Pith. sign in

Paper Citation Record · LEDGER

Can A Gamer Train A Mathematical Reasoning Model?

As of 18 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2506.08935.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08935 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:03:22.166335Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3cbc36e8-0509-45b7-a539-35431a402049 · outbound

This paper cites Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?.

Can A Gamer Train A Mathematical Reasoning Model? Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.393043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.393043Z digest=sha256:3b39c4bdb6a4051db9816f9ce41219c01b8ce339270e57e1f69331190822032a

Observation 544a754f-e876-4d6b-994c-6ceef92d32e3 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Can A Gamer Train A Mathematical Reasoning Model? LoRA: Low-Rank Adaptation of Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.566380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.566380Z digest=sha256:6f4a10083a00dbbdd26a0d170aa60fb4f11dc149dbf8ccda0af50f47fa1ce0c2

Observation 04986e11-49e8-4dd6-8ace-19c6b744973f · outbound

This paper cites s1: Simple test-time scaling.

Can A Gamer Train A Mathematical Reasoning Model? s1: Simple test-time scaling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.780247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.780247Z digest=sha256:60c021d607db13ffcd4b3ad889729d59391749c3c26cc0cf56fa948e02b3a738

Observation d2866b02-8b67-4a79-a955-e77869c69d5c · outbound

This paper cites https://github.com/Jiayi-Pan/TinyZero.

Can A Gamer Train A Mathematical Reasoning Model? https://github.com/Jiayi-Pan/TinyZero

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:03:22.834765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:03:21.813208Z digest=sha256:22a4d3186f9b1c9500d3ca452b7add0f5224a2b95f8f9fa7b804f1bf3df6d49b

Observation 53c2a64f-45d1-4f5a-922b-6db29596c5a8 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Can A Gamer Train A Mathematical Reasoning Model? DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.885652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.885652Z digest=sha256:b1caf7ab1e478998e899b58ae7848cee1225d702f179f01aaeef7e3c40cf40e7

Observation 75a29c77-a704-4f57-a59c-4acba6a27bd1 · outbound

This paper cites https://novasky- ai.github.io/posts/sky-t1.

Can A Gamer Train A Mathematical Reasoning Model? https://novasky- ai.github.io/posts/sky-t1

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:03:22.629238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T05:03:21.978742Z digest=sha256:0e4695936e730a608e1c6a82cc123afc7d8b450a2a8d6b38a337675d1b1554cf

Observation d9a622f5-a6be-4aaf-8554-b7ede7922f75 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Can A Gamer Train A Mathematical Reasoning Model? Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:22.072238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:22.072238Z digest=sha256:02de802e0091aefe9f429e3ca2fa3159aeb487f18cda66a180ceb84c260dcf10

Observation 3f308fce-c9de-428b-9bb8-328833fb15f1 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Can A Gamer Train A Mathematical Reasoning Model? Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:22.166335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:22.166335Z digest=sha256:05810f22e24679d4b2b055efb62e0e0d3542c42b8d2707a6b75235b1273de253

Observation a06b9743-567f-473e-a005-628e5a6b21be · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Can A Gamer Train A Mathematical Reasoning Model? Measuring Massive Multitask Language Understanding

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.491475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.491475Z digest=sha256:506014caa7b80dd0dc65d26b576ac28eb21a53d9e62b0422e4581626609d0487

Observation d8486f86-9c96-4722-b765-98df0273a797 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Can A Gamer Train A Mathematical Reasoning Model? Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.225062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.225062Z digest=sha256:ab9060e516baa38c923b7522610ceb426aec34e4e822928732567391fa7be237

Observation 50b2e901-02a1-4f37-904a-fa46d9a881be · outbound

This paper cites Solving Quantitative Reasoning Problems with Language Models.

Can A Gamer Train A Mathematical Reasoning Model? Solving Quantitative Reasoning Problems with Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.738356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.738356Z digest=sha256:1415f596094998a1b59e8e41743900d701979cd2fb531778f6c52adcd67221f7

Observation 79ab7925-0d95-4377-8d10-0c77e92cd292 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

Can A Gamer Train A Mathematical Reasoning Model? FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.279182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.279182Z digest=sha256:994e21c9956b0bc43d4d428297a8d0ff2ed4ee07b4a61cf10cac9de2ed1e1fab

Observation 1fa0c303-c60d-4cd1-9227-57027454b319 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Can A Gamer Train A Mathematical Reasoning Model? Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.654953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.654953Z digest=sha256:f771cc247be541c61744a604c3b9515dd877651e08a07e4580d16c78e00944e8

Observation d67c9a9d-fc43-45a7-a8bc-3fa3bd0ceb49 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Can A Gamer Train A Mathematical Reasoning Model? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.349189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.349189Z digest=sha256:92c0ee15708c5a0fe2c7b8696d9e53f5547a4f7bfa8adabbcfdfd8170df35631

Pith citing papers

No inbound Pith citation observations are available.