Pith. sign in

Paper Citation Record · LEDGER

Can A Gamer Train A Mathematical Reasoning Model?

As of 9 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 0 inbound Pith citation observations for arXiv:2506.08935.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08935 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:03:22.166335Z

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

14 of 14 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3cbc36e8-0509-45b7-a539-35431a402049 · outbound

This paper cites Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?.

Can A Gamer Train A Mathematical Reasoning Model? Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.393043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.393043Z digest=sha256:ba0746b75cbc23ba37e67794deebb7969a0f05c4edc37b0cc8018a171d0d3f03

Observation 544a754f-e876-4d6b-994c-6ceef92d32e3 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Can A Gamer Train A Mathematical Reasoning Model? LoRA: Low-Rank Adaptation of Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.566380Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.566380Z digest=sha256:cb15563ec9d8296843d0c30443a04915fbae32c6bb46547b8be4e9e97cf0fe94

Observation 04986e11-49e8-4dd6-8ace-19c6b744973f · outbound

This paper cites s1: Simple test-time scaling.

Can A Gamer Train A Mathematical Reasoning Model? s1: Simple test-time scaling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.780247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.780247Z digest=sha256:241bd71c160039abd01b23005bde210c7b624f0ff8a7f3ecd7db5cb3f67828fa

Observation d2866b02-8b67-4a79-a955-e77869c69d5c · outbound

This paper cites https://github.com/Jiayi-Pan/TinyZero.

Can A Gamer Train A Mathematical Reasoning Model? https://github.com/Jiayi-Pan/TinyZero

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:03:22.834765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:03:21.813208Z digest=sha256:871ff989db0bcec4665f07d43d517c07ad562438784a2eaa34e1aec7d0b4b309

Observation 53c2a64f-45d1-4f5a-922b-6db29596c5a8 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Can A Gamer Train A Mathematical Reasoning Model? DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.885652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.885652Z digest=sha256:4e9a62f748ff4ce7de99c20dacfcdb19fecd6ce3e2ce01ab0aa00fad3d5d401f

Observation 75a29c77-a704-4f57-a59c-4acba6a27bd1 · outbound

This paper cites https://novasky- ai.github.io/posts/sky-t1.

Can A Gamer Train A Mathematical Reasoning Model? https://novasky- ai.github.io/posts/sky-t1

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:03:22.629238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T05:03:21.978742Z digest=sha256:d80f45b35c0c37d4bbfeff9d880a6f88f64cdaeafcc8d4b470688befa5f244e3

Observation d9a622f5-a6be-4aaf-8554-b7ede7922f75 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Can A Gamer Train A Mathematical Reasoning Model? Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:22.072238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:22.072238Z digest=sha256:4a8800df277aebf4276200fc38ed3c270815bd9e64ac686c1a5bfdd5cd2bf0ad

Observation 3f308fce-c9de-428b-9bb8-328833fb15f1 · outbound

This paper cites Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement.

Can A Gamer Train A Mathematical Reasoning Model? Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:22.166335Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:22.166335Z digest=sha256:743b503297dbe8d64e59c6997bf8db93d857716df8a1dcc493dff356e6a1e19a

Observation a06b9743-567f-473e-a005-628e5a6b21be · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Can A Gamer Train A Mathematical Reasoning Model? Measuring Massive Multitask Language Understanding

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.491475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.491475Z digest=sha256:d33b5a13015fb6a9c969f446e96e29a92004a28088087bdf939ccd9f6fc8092e

Observation d8486f86-9c96-4722-b765-98df0273a797 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Can A Gamer Train A Mathematical Reasoning Model? Training Verifiers to Solve Math Word Problems

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.225062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.225062Z digest=sha256:eed6153e67abf9aeae427ff8df8aa310ae3a2d7af3c1e0fbf775b11258eb9b63

Observation 50b2e901-02a1-4f37-904a-fa46d9a881be · outbound

This paper cites Solving Quantitative Reasoning Problems with Language Models.

Can A Gamer Train A Mathematical Reasoning Model? Solving Quantitative Reasoning Problems with Language Models

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.738356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.738356Z digest=sha256:ddbb2fb50f92f0691a768b2ec1c332685aa66af3a2e445325f723d31dc5805bb

Observation 79ab7925-0d95-4377-8d10-0c77e92cd292 · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

Can A Gamer Train A Mathematical Reasoning Model? FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.279182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.279182Z digest=sha256:ba66d6a2ce09c28c69ad55e5a3d137fc0bdfbf39dc129c6e655e9d9ed5e3bc14

Observation 1fa0c303-c60d-4cd1-9227-57027454b319 · outbound

This paper cites Tulu 3: Pushing Frontiers in Open Language Model Post-Training.

Can A Gamer Train A Mathematical Reasoning Model? Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.654953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.654953Z digest=sha256:4bbc6059711b32f07472795ed97e882d1199921f6a6e0c08491b31c763bb4e1b

Observation d67c9a9d-fc43-45a7-a8bc-3fa3bd0ceb49 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Can A Gamer Train A Mathematical Reasoning Model? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-07T05:03:21.349189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:03:21.349189Z digest=sha256:56b2f1e2fd0be368b4bf84f383c409fe23038ab842da6fde6244b01cb1a96755

Pith citing papers

No inbound Pith citation observations are available.