Pith. sign in

Paper Citation Record · LEDGER

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests

As of 8 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 2 inbound Pith citation observations for arXiv:2506.04894.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.04894 v1

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:35:02.149902Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T09:22:06.285118Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T09:23:10.509262Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved18
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e77ef952-976b-4ba7-94ab-d1b4ac2d1122 · outbound

This paper cites A Survey of Large Language Models.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests A Survey of Large Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.064673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.064673Z digest=sha256:f59a831748ed66cfc9449e9bc884b6189d74db040210993f90a2ab5cf74c4539

Observation 26129e3c-be2f-4187-b8b5-68cf2eba32cc · outbound

This paper cites https://openai.com/o1/, 2024.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests https://openai.com/o1/, 2024

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:35:02.400755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.068921Z digest=sha256:1d2f2f2985e7a64cc92a3655ace248bb0a4d181782242044ef12fa3e96d2de88

Observation ee4e85c7-7869-4b11-a110-3191341deefd · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.072166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.072166Z digest=sha256:82869977bc26550c4e7ad7f94712467e85004effbea332a29e3745b8fda9c58a

Observation 8297e588-252c-468d-9eac-ec1d5cc78065 · outbound

This paper cites an unresolved cited work.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:35:02.390764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.075898Z digest=sha256:d8a8618a2653ee540e118c507d775f1150e553bd34c3cb59b1cb96bea21c5c2c

Observation 3f1773dd-1287-4e9a-be38-143068c21443 · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.078965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.078965Z digest=sha256:b35869cce6088b93c393dd981f54154f741e5bf75df03fc18231e9cbeb61665a

Observation 0153c955-1265-4112-80d0-b32548f7555f · outbound

This paper cites Can Language Models Solve Olympiad Programming?.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Can Language Models Solve Olympiad Programming?

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.083166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.083166Z digest=sha256:dc55453211238d6b9d456b39d523a81e4f1c4442b963452bd28d9a7675c349a9

Observation 7d8cdb37-b2b4-48b3-9ac1-14879d128c64 · outbound

This paper cites CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.087340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.087340Z digest=sha256:6a6ec0cb860bc9a044064968fed0dee93138d420fc5bbdfc479b236d55efe635

Observation 88feea31-4c79-4398-9151-0cb66d8085c7 · outbound

This paper cites Reflexion: Language Agents with Verbal Reinforcement Learning.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Reflexion: Language Agents with Verbal Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.091598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.091598Z digest=sha256:5c86bfd098e5d940b049e77df9eaf413dff5f428b53869d3099e236332a65dc2

Observation ff232cf5-5847-4411-a246-d0c2a69f316f · outbound

This paper cites an unresolved cited work.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:35:02.379835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.095896Z digest=sha256:23dc409171760bae4155159c19f523a7f66cd0954dd5e943fb1187588e6ef670

Observation b71aa269-6c29-43ee-8021-8794b2ccca8a · outbound

This paper cites Program Synthesis with Large Language Models.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Program Synthesis with Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.099581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.099581Z digest=sha256:43c97ac88a2bf51a9146e96e346e679a74a9269ed5b8235bc7da1450b2427ee5

Observation 6cf597c2-5c45-42d9-bec7-0a2af91442fe · outbound

This paper cites Measuring Coding Challenge Competence With APPS.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Measuring Coding Challenge Competence With APPS

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.102815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.102815Z digest=sha256:b6346ac4d7ca72e41c654f8c998a6abf4704f9e0c550c051460782b635bea379

Observation 857463fc-6d53-45ae-bf67-fbd019d0a77c · outbound

This paper cites Mankowitz, Esme Sutherland Robson, Pushmeet Kohli, Nando de Freitas, Koray Kavukcuoglu, and Oriol Vinyals.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Mankowitz, Esme Sutherland Robson, Pushmeet Kohli, Nando de Freitas, Koray Kavukcuoglu, and Oriol Vinyals

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:35:02.370461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.106739Z digest=sha256:4b0527a353c90553b0df6315d3cd8f093f2832272437629cb7521df58e7829c5

Observation 313fd987-6613-4c39-8bd6-c6425ed2de53 · outbound

This paper cites xCodeEval: A Large Scale Multilingual Multitask Benchmark for Code Understanding, Generation, Translation and Retrieval.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests xCodeEval: A Large Scale Multilingual Multitask Benchmark for Code Understanding, Generation, Translation and Retrieval

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.109846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.109846Z digest=sha256:c175c5130a720c994e5a30cfeb6a38cee382949ccb7aa13b3af08cadb93596b7

Observation 20962425-cee4-4456-8526-4482cde1277e · outbound

This paper cites Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Self-Play Fine-Tuning Converts Weak Language Models to Strong Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.113373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.113373Z digest=sha256:4c09fdb4eb662ef92f551f5a0179c1e1dddea09f4304f47edde26be00ad3bc6f

Observation 64625227-ca3c-4f83-aa9e-2cf775e2a563 · outbound

This paper cites Multi-turn rl training for cuda kernel generation.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Multi-turn rl training for cuda kernel generation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:35:02.360452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.117383Z digest=sha256:7913c8846b4763e0b5cc6569922187de989ac51006c49ba9de5fe05189b0febc

Observation ac9e268a-8594-43a4-8285-48859606379a · outbound

This paper cites Is Self-Repair a Silver Bullet for Code Generation?.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Is Self-Repair a Silver Bullet for Code Generation?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.120455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.120455Z digest=sha256:9709a7c271b1fbb59ba26f4400f38e2852524ee4c833c880b9a42559f342b9a8

Observation b1e36e96-768f-4d86-a294-6eddbee9cf53 · outbound

This paper cites OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests OpenCodeInterpreter: Integrating Code Generation with Execution and Refinement

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.124225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.124225Z digest=sha256:3247c7ce22293e8dfaa9a39d76d71fd0a7d0606f833c8ac79c0f6f0ff0aab686

Observation 003e6e37-0338-4fdb-989d-6c79e5794682 · outbound

This paper cites SPoC: Search-based Pseudocode to Code.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests SPoC: Search-based Pseudocode to Code

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:35:02.350049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.127317Z digest=sha256:56207722a1285851ffb596cf5da303c532e2006a049970d9de4b02422b867571

Observation 141f6732-d71d-414f-a849-015d7191659a · outbound

This paper cites Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Kimina-Prover Preview: Towards Large Formal Reasoning Models with Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.130547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.130547Z digest=sha256:d76033c70ce927bd1bde449400bf404420c2c669ca65942c717f34f616a6a621

Observation bcb97050-bb2e-477c-914b-4236db4aa6d2 · outbound

This paper cites an unresolved cited work.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:35:02.339445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.134377Z digest=sha256:f6e9956f66d84b2ec126244fdbfb8f90bf9265b53d3b783a73d0e978f327a76f

Observation 80de60c8-3f66-40a9-913c-4d607dae410c · outbound

This paper cites QwQ-32b: Embracing the power of reinforcement learning, March 2025.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests QwQ-32b: Embracing the power of reinforcement learning, March 2025

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:35:02.328705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.138061Z digest=sha256:0c5b6c2ddccabbfb4020d1970f6a8710e3a9367fa520a2f5a75586ed585743c7

Observation 91706d69-cc3c-4652-96bf-4e1f55e08da2 · outbound

This paper cites An Empirical Study on Eliciting and Improving R1-like Reasoning Models.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests An Empirical Study on Eliciting and Improving R1-like Reasoning Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T10:35:02.142063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:35:02.142063Z digest=sha256:b27d64d02f8203e792a7c663562f8d13ebaa7beeb1faa3ed9a7ada6af63f1da7

Observation fa1eb5f7-65cb-4f62-883a-f2706f29a21f · outbound

This paper cites an unresolved cited work.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:35:02.317935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.146268Z digest=sha256:9401c66c5f3977c451f7fed127bc808f81f7611d9257e0b5e7abda240bd58827

Observation 49a4dd87-d001-41d3-b3c7-cd905c9f2adf · outbound

This paper cites Category1EnglishName.

ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests Category1EnglishName

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:35:02.307798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T10:35:02.149902Z digest=sha256:18691ea260f4ede7a798b960501d8faad236908056c6f29d13017a08466dfdf7

Pith citing papers

Observation e4539fd5-b99d-4d9f-a2a4-64a894d074c0 · inbound

Forge: Quality-Aware Reinforcement Learning for NP-Hard Optimization in LLMs cites this paper.

Forge: Quality-Aware Reinforcement Learning for NP-Hard Optimization in LLMs ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T02:46:18.912715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T02:44:33.143247Z digest=sha256:9d25baf2a3af535434c383b91faaa9e80db8e52cdaa518843b688a75f0f9627a

Observation cd3847b9-467e-4bdf-bfa5-13c1fb21692a · inbound

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback cites this paper.

A-ProS: Towards Reliable Autonomous Programming Through Multi-Model Feedback ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-20T09:23:10.510877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T09:22:06.285118Z digest=sha256:66132ec32dc6082abd0dde8d261aa8527029a57327cf30870dba7c66ea639170