Pith. sign in

Paper Citation Record · LEDGER

CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2401.06961.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.06961 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:33:53.478015Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T01:12:20.889417Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dffef66e-f120-4612-9347-99c579b7279a · inbound

Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models cites this paper.

Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:09:15.024469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-15T09:09:14.884516Z digest=sha256:1c55d8cd6475cb123bba27d3dfc4edee96b8478b8e1b75573c0437207deec695

Observation ac0fbba7-2344-495a-9eee-f697900bc646 · inbound

Division-of-Thoughts: Harnessing Hybrid Language Model Synergy for Efficient On-Device Agents cites this paper.

Division-of-Thoughts: Harnessing Hybrid Language Model Synergy for Efficient On-Device Agents CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T01:02:19.347716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T01:02:19.347716Z digest=sha256:d62e3728191ab7b42584f95331dee176434c822a9df495390b1d5ad354d97a77

Observation bab767fe-f319-4009-a8d9-849885fb46e4 · inbound

Bridging Language Models and Financial Analysis cites this paper.

Bridging Language Models and Financial Analysis CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-23T01:12:20.891928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-23T01:08:58.528533Z digest=sha256:83e9d3f5d5e87d5c49e8c571ec5c948f9a05d7a5770d96a9d5eff658518cd657

Observation 977c2e83-8fa3-429a-a89e-285fad4c092e · inbound

Evaluating Judges as Evaluators: The JETTS Benchmark of LLM-as-Judges as Test-Time Scaling Evaluators cites this paper.

Evaluating Judges as Evaluators: The JETTS Benchmark of LLM-as-Judges as Test-Time Scaling Evaluators CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T11:33:53.478015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:33:53.478015Z digest=sha256:afd6cb7afd349cdf477f104d1d49884baa9b2c6b141bb8c4309dd1466eddf517

Observation 041b8145-2141-4cbd-842f-054c0ce7e256 · inbound

ASyMOB: Algebraic Symbolic Mathematical Operations Benchmark cites this paper.

ASyMOB: Algebraic Symbolic Mathematical Operations Benchmark CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T13:05:07.985794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:05:07.985794Z digest=sha256:b73a5afb2156bc9e88b9cf5bc7af36ce9ffdd42a79d3c70dcc6ed617e61dfe0b

Observation 2f6bf573-832f-44a8-8a96-c25d0699a9ca · inbound

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning cites this paper.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.200382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.200382Z digest=sha256:c50b42aafc65e4cebc107736b4931eb82a8a01ebf83a34d56ea96f41cf932c04

Observation f8c269fe-e868-4948-bbb6-f42d051ab0e9 · inbound

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability cites this paper.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T15:30:26.808039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:30:26.808039Z digest=sha256:8a7de19945c67f60f0a36ce8ab331254b02ccfdca1f8e2ec5cc287898ae2f9fe

Observation c09727e4-5cb0-4e9e-b901-e4c578b27683 · inbound

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability cites this paper.

TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:18:57.408778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:18:57.408778Z digest=sha256:52b3a95976e83f65f3db6a2be75aac6ded97bb8ed9122506628f4564af0a3231