Pith. sign in

Paper Citation Record · LEDGER

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning

As of 9 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 2 inbound Pith citation observations for arXiv:2509.07711.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.07711 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T21:54:48.207513Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T02:43:21.225324Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T17:03:40.960916Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 91980745-a815-4497-aa47-6645066c68b4 · outbound

This paper cites write newline.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.163592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.163592Z digest=sha256:b35c755095f56d34e54383826dde29b660abea8540791f9340615cab6f97e175

Observation 8ec6397f-0a48-44fe-9296-537bc56b8ec4 · outbound

This paper cites Have LLMs Advanced Enough? A Challenging Problem Solving Benchmark For Large Language Models.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning Have LLMs Advanced Enough? A Challenging Problem Solving Benchmark For Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.168151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.168151Z digest=sha256:ad7e3662da60c6877439554a52d1883973a495b009140a4fb20c1e7ac3b96464

Observation c02c81a9-9915-4f04-a03a-c4afc77e0d54 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning Training Verifiers to Solve Math Word Problems

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.172194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.172194Z digest=sha256:882fa852cf2f95974975540d6593fabdc3feb74b96f767ef7b9c7d7a1e5943dd

Observation 9f338992-0680-4ea8-bdcc-575fc8494ce0 · outbound

This paper cites MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning MathOdyssey: Benchmarking Mathematical Problem-Solving Skills in Large Language Models Using Odyssey Math Data

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.175953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.175953Z digest=sha256:046ddc7778ae27669549c0faf59fa4c13846f46a6c8824ee610b357301893a3a

Observation 2fa9f505-165b-462c-b939-4e3402f737cf · outbound

This paper cites C., Buzzard, K., Gowers, T., Liu, P.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning C., Buzzard, K., Gowers, T., Liu, P

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T21:54:48.346205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-04T21:54:48.179897Z digest=sha256:b452b556b9e2173dc81b8a88aef768321534588b7cf892ab208cad9d922a0b03

Observation 4aca38e9-be89-4ee7-a789-2e9140c49eef · outbound

This paper cites Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning Omni-MATH: A Universal Olympiad Level Mathematic Benchmark For Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.183408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.183408Z digest=sha256:6590ef3352a7aed6226aa2f38d6201e5d2f25e534bccc87bb277eedd926aa90c

Observation 8fdbb723-3940-4b1c-8063-473fe749f53e · outbound

This paper cites OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.186924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.186924Z digest=sha256:5f3d16ca5e2cfca6f5c048d8978c41820364d7ed4575373308efd9312d97bcd6

Observation 9a3c3ed3-3b79-4a8e-a1e1-a46f8a41fe13 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning Measuring Mathematical Problem Solving With the MATH Dataset

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.190554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.190554Z digest=sha256:9be4225d3f6656746f5feba1eb2ebc7cb69c7f7a2c61f5ddea0447b105e4d15a

Observation 474d1bb9-b514-48ba-8b01-f66d93506437 · outbound

This paper cites OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.193761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.193761Z digest=sha256:c655ca74cd72cd4ac804e623be2639adc34121234ca535212d6d912dbb425ee1

Observation cc3b578e-4cf2-4520-aef8-505414412de6 · outbound

This paper cites Solving Quantitative Reasoning Problems with Language Models.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning Solving Quantitative Reasoning Problems with Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.197117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.197117Z digest=sha256:25eaa22dd570bd23a67b3220e2e6b31e39af505634d7b06384a83443f93f1da2

Observation 2f6bf573-832f-44a8-8a96-c25d0699a9ca · outbound

This paper cites CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning CHAMP: A Competition-level Dataset for Fine-Grained Analyses of LLMs' Mathematical Reasoning Capabilities

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.200382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.200382Z digest=sha256:77121754049a4a55a004664ab112875d5296e85fdbac15b34eeee20e1c6e9f55

Observation d14ab61b-d3ac-4f1f-9f82-52e38198063a · outbound

This paper cites Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.204024Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.204024Z digest=sha256:5bad7e2659a3cf94513f341cf9337e825cea46f3ae1acfef0b907cbb282b71e3

Observation 1be36bea-8b80-4796-9a91-01506996d97d · outbound

This paper cites Solving olympiad geometry without human demonstrations.

RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning Solving olympiad geometry without human demonstrations

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T21:54:48.207513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:54:48.207513Z digest=sha256:939f4f4980a34045f5b37f0818d91427a9dce0a57428653f81124461d1c0b166

Pith citing papers

Observation b6ba982a-cd0d-407f-8a8e-f0e3206bb462 · inbound

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking cites this paper.

Prefix-Safe Bayesian Belief Tracking for LLM Reasoning Reliability:Separating Calibration from Ranking RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:03:40.962251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T16:59:25.127619Z digest=sha256:3e7a48aed9728efebb917a46059e405379918800393b04c74e75cfb78996c299

Observation 3e529a14-449a-45c3-aa37-861e484015e8 · inbound

AdvancedMathBench: A Benchmark Suite for Advanced Mathematical Proof Generation and Verification cites this paper.

AdvancedMathBench: A Benchmark Suite for Advanced Mathematical Proof Generation and Verification RIMO: An Easy-to-Evaluate, Hard-to-Solve Olympiad Benchmark for Advanced Mathematical Reasoning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-14T02:43:21.225324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T02:43:21.225324Z digest=sha256:dc86b8746a09ced9476b32157c500518dd5d07f60d5106de59e9eae263061d68