Pith. sign in

Paper Citation Record · LEDGER

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

As of 9 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 17 inbound Pith citation observations for arXiv:2506.11928.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11928 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T01:08:02.659078Z

measured 83 of 83 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:03:28.866165Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:49:30.015976Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy46
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7e61df85-12b3-4408-8dd3-3d2b2ecd8704 · outbound

This paper cites URLhttps://github.com/openai/human-eval.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://github.com/openai/human-eval

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.912248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:57.650121Z digest=sha256:f26e4ec889aacfb80ade479495fcec134ede65f82d33e8089b76f9849449d2a0

Observation fbc2a48f-8ec1-4527-bd64-2946931f1197 · outbound

This paper cites URLhttps://icpc.foundation/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://icpc.foundation/

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.710948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:57.704459Z digest=sha256:4267939e3006d8d27d0d7a81e54dfc3858a1751dd6266417787bf0c67173a79e

Observation a95ef9c6-6065-413f-8639-c5ba3ee1626f · outbound

This paper cites URLhttps://icpc.global/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://icpc.global/

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.550852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:57.729387Z digest=sha256:c1a06cea079bc4e7b13d67b72d8ea88aea5f5c0171c10c599dabe2a3e29fb329

Observation bb6c5693-05d5-4e4c-957a-a4859a81bbea · outbound

This paper cites URLhttps://ioinformatics.org/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://ioinformatics.org/

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.397852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:57.772854Z digest=sha256:dd0df43815fd3cd4f47afd82825b42fbf7a3e4b2f8552450583ef095ab8d5e17

Observation 601c7cb8-a1f1-4da4-bdc6-20c95ffc6422 · outbound

This paper cites URLhttps://mitit.org/About.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://mitit.org/About

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.208163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:57.863687Z digest=sha256:60c9ffc155847a0efd7b049d023eca38d05998105ea3094662a023629d3c4fdb

Observation 0a51eef2-ed57-464b-b5c0-71273b4bf5c6 · outbound

This paper cites URLhttps://noi.cn/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://noi.cn/

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.982844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:57.954851Z digest=sha256:e307ced467aece9df5a9bc2da7c7e597b0d696c80015edd6975e192a1e050a46

Observation fefce80b-9ae5-4ed1-be0f-59f1f83870dd · outbound

This paper cites URLhttps://thusaac.com/public.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://thusaac.com/public

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.795735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.039732Z digest=sha256:2858e9984f7aadaec4ba1004dac239609ad95a7a103dd205186b5a8a0f8378cc

Observation 1f30c23d-5f85-48b5-8115-5c728e016eeb · outbound

This paper cites URLhttps://usaco.org/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://usaco.org/

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.647479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.125534Z digest=sha256:40b169ee3ff1e4b7d25acaacfaadfd1320622997a294768682c6ddfc2568c245

Observation 9279a984-dee4-4fe5-bbe0-aef6d947a445 · outbound

This paper cites URL https://icpc.global/worldfinals/fact-sheet/ ICPC-Fact-Sheet.pdf.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URL https://icpc.global/worldfinals/fact-sheet/ ICPC-Fact-Sheet.pdf

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.455653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.172537Z digest=sha256:de08ddfe487d6304cbf26a5c4e946d4c1d473f0cbf3b50279914f1eb96829c79

Observation e19e66de-0cf3-47cb-a0d1-49421139d572 · outbound

This paper cites Model card addendum: Claude 3.5 sonnet.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card addendum: Claude 3.5 sonnet

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.241010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.275845Z digest=sha256:3717d6e49b8ef6cf2996f7d4af74d0249cd6ae7f7286e27444b9bcf672f11627

Observation 534e24e2-3845-4559-b68e-eb40de527b50 · outbound

This paper cites Claude 3.5 Sonnet, 2024.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Claude 3.5 Sonnet, 2024

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.054698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.346965Z digest=sha256:cd65911bd79e7b602a5f84e42b8bc9142af5fe8045a5906f29e1e621da24c0b5

Observation 890b873e-e208-449b-8b1a-ecdc8bf3bae4 · outbound

This paper cites System card: Claude 3.7 sonnet (max reasoning).

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Claude 3.7 sonnet (max reasoning)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.891714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.417539Z digest=sha256:72722002209883b82af40ee80e9bbe79743915c7e8f806cd3faa179bd2a50612

Observation bc025a9b-ae8b-4298-b6f4-d1ed636e0cc4 · outbound

This paper cites System card: Claude 3.7 sonnet (no reasoning variant).

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Claude 3.7 sonnet (no reasoning variant)

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.713673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.505902Z digest=sha256:5527da9e36ea6fdc8cf4e6c20c0963d9abda27d0bc59800fc91663c58db435a2

Observation 7cc2a4d9-b654-4841-8151-b7ec4c9f343c · outbound

This paper cites Program Synthesis with Large Language Models.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Program Synthesis with Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:58.590186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:58.590186Z digest=sha256:d38112c601f41d38b6d551b5d9145aeaf48e7e8880e59a766763fd2dd06cf4e6

Observation 49e3dee7-7ebe-4c34-80a3-9705146f1833 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Evaluating Large Language Models Trained on Code

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:58.655057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:58.655057Z digest=sha256:d8663847392285cf2cef8e86a3786968affb56606d17dbd67eb2fcda26ed0811

Observation aa5e5b86-f8b7-4ffd-9135-4b704838e165 · outbound

This paper cites Official website of the china collegiate programming contest (ccpc), 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Official website of the china collegiate programming contest (ccpc), 2025

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.563332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.763693Z digest=sha256:b1565812ec427ddd451ab0c0f7e14d7bd2be5d3e314383d56c8220bd3db73c2c

Observation b9d895ed-2165-48cd-983f-822bc8f7ee18 · outbound

This paper cites Model card: Qwen-max (qwen 2.5 max).https://huggingface.co/Qwen, 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Qwen-max (qwen 2.5 max).https://huggingface.co/Qwen, 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.403057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.807679Z digest=sha256:3d55367229cbe9d8916d0fc7403dd4052c160a5fdf29dbf0aac264b7981c000a

Observation 98b590cb-eb67-49ab-a45a-f735e7573275 · outbound

This paper cites Interactive Problems: Guide for Participants, 2015.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Interactive Problems: Guide for Participants, 2015

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.273753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.870319Z digest=sha256:957cb83b4faa610f9f31c681a5a407af90f1cc7127d9091ef1f37835657a65b3

Observation 44140a0d-feac-4d1e-9f8f-0e06318b4424 · outbound

This paper cites MHPP: Exploring the Capabilities and Limitations of Language Models Beyond Basic Code Generation.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? MHPP: Exploring the Capabilities and Limitations of Language Models Beyond Basic Code Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:58.930241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:58.930241Z digest=sha256:d6fc340464ead1d236498e6ef04125fc28ffbb79a6cee85857348dc772a22fe6

Observation da2e4af5-431b-4d31-96ad-d58262e1a9f2 · outbound

This paper cites Model card: Gemini 2.0 flash reasoning.https://blog.google/technology/ google-deepmind/gemini-model-updates-february-2025/, 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemini 2.0 flash reasoning.https://blog.google/technology/ google-deepmind/gemini-model-updates-february-2025/, 2025

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.105664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:58.983607Z digest=sha256:13b5e06b397ed6b7c91d018e313717064f612cc72afa92fc51d92339a5785590

Observation d9d37b68-7746-4afc-b2f1-5cd7fcbbc93e · outbound

This paper cites Model card: Gemini 2.5 flash.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemini 2.5 flash

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.931038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:59.042706Z digest=sha256:d8a6d409171249f954d9b3e8534fbceda940bd5fb0351f8f521e8f9771dfb54d

Observation ea9583b1-adee-44f9-bcea-6384fbfe084e · outbound

This paper cites Model card: Gemini 2.5 pro.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemini 2.5 pro

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.761427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:59.113155Z digest=sha256:aa9b3a707a7b17923acbcca458fb0602caaf3c1e32f1991e61f9a008ce027615

Observation b71899ab-f791-46a4-8e37-76cfc634e293 · outbound

This paper cites Model card: Gemma 3 27b.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemma 3 27b

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.620939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:59.170629Z digest=sha256:a585040fcbd6be74a205fa435b9c1022e435a3e2ded0d293c395e01b37456a43

Observation fb141544-2751-47ee-9bae-772200c57b33 · outbound

This paper cites Model card: Deepseek v3.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek v3

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.487487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:59.221958Z digest=sha256:786fda525e28bea245440f9709ab672fa15c1e014adbb8ca55e7cca919e3c819

Observation ab040e01-ad0d-481a-8131-0bc4fd6f0d76 · outbound

This paper cites Model card: Deepseek r1.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek r1

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.343556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:59.307600Z digest=sha256:ba30b592dffcc4ab2fb2e0a850379b73861264217c7edb40ec9b3f52ecbb82b4

Observation 0341b531-4233-4697-8669-0999687fa032 · outbound

This paper cites Model card: Deepseek -r1-distill-llama-70b.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek -r1-distill-llama-70b

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.187694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:59.369728Z digest=sha256:b3e46277d092c69fd47d2adc0fa4105985cbc6bda30fd01310829f4707304987

Observation f5330032-db56-4de9-980d-cedff58496b3 · outbound

This paper cites Model card: Deepseek -v3-0324.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek -v3-0324

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.031594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:59.460378Z digest=sha256:4ecead03f3f526c2821e9755ee2b4772a81d6ed7c215e393c9ce278450632a2c

Observation c3457360-fb3c-4697-85d8-6269f99c6dbd · outbound

This paper cites Crosscodeeval: A diverse and multilingual benchmark for cross-file code completion.Advances in Neural Information Processing Systems, 36:46701–46723, 2023.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Crosscodeeval: A diverse and multilingual benchmark for cross-file code completion.Advances in Neural Information Processing Systems, 36:46701–46723, 2023

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.898533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:59.541691Z digest=sha256:b01ab8dbb4082ae1c0c4809d017bc01cd5332c817a83d0dd013a3c8e2ef430ea

Observation 0a4544c9-699f-4bc4-baff-eb138c2dd571 · outbound

This paper cites Evaluating the performance of large language models in competitive programming: A multi-year, multi-grade analysis.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Evaluating the performance of large language models in competitive programming: A multi-year, multi-grade analysis

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.766010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:07:59.591769Z digest=sha256:b93752be99aa6e4e1b8599f9857e634a26c9fbc5bffcee77878855575b8e183b

Observation b8acc336-1679-4fcb-9886-f24733f45d75 · outbound

This paper cites Mathematics and games.Eureka, 2:6–8, 1939.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Mathematics and games.Eureka, 2:6–8, 1939

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.639019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.639019Z digest=sha256:01ac87493c3e487caeb1976d3cd8890ec6256667eabd2547a2487373459f4348

Observation 48e7c5d4-46a5-4548-a2da-b95fd9ebe06d · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.728279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.728279Z digest=sha256:9e33676e6bf6418953f00cb81c31812eda9bde575d7b609b1237453ea4c57697

Observation 2c2e3929-fc0a-4716-a136-adb14b1f5c78 · outbound

This paper cites Llm-pros: Analyzing large language models’ performance in competitive problem solving.arXiv preprint arXiv:2502.04355, 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Llm-pros: Analyzing large language models’ performance in competitive problem solving.arXiv preprint arXiv:2502.04355, 2025

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.778246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.778246Z digest=sha256:1fd4d03a0375b6c23ca252c301c7290128919c69193c048eb3a71b133ce60e2f

Observation 5637c6f3-e880-4a07-a26c-5372d634466d · outbound

This paper cites Competition-Level Problems are Effective LLM Evaluators.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Competition-Level Problems are Effective LLM Evaluators

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.838876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.838876Z digest=sha256:52f8a1f6770559935eb470ed5a1aae8759f1793dec4799eceff833f739d37b30

Observation 64f8c4ad-35fb-4f20-ba1f-50281ce84772 · outbound

This paper cites OpenAI o1 System Card.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? OpenAI o1 System Card

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.903622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.903622Z digest=sha256:15517fee27aa9bc148bbe3ab9cccdda1b3f1d8aeaf5324b333841f1679af3ffe

Observation 09e0b0da-7573-4402-a07b-62a162f0725c · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.995412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.995412Z digest=sha256:591ec645e21f91da4320fc8ab32fe3614e3b1d61f8f9e12c4221d22d8f8b48d0

Observation eacc619b-c547-4daa-9675-869670242216 · outbound

This paper cites Swe-bench: Can language models resolve real-world github issues? InThe Twelfth International Conference on Learning Representations, 2024.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Swe-bench: Can language models resolve real-world github issues? InThe Twelfth International Conference on Learning Representations, 2024

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.585390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.043224Z digest=sha256:9eea1b0c9cb296ab1a13e073209f7cc35b461392b3ae903e21adc7ac847ce22a

Observation b8f52cd7-2864-4d56-ba98-dfc61aabea0a · outbound

This paper cites Competition-level code generation with alphacode.Science, 378(6624):1092–1097, 2022.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Competition-level code generation with alphacode.Science, 378(6624):1092–1097, 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.400024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.106253Z digest=sha256:a7059667503a194993f1af9a578ee9f15a8987e4812c1285e983955552dc1ae2

Observation 3a57fd1f-cf44-4052-92ac-fdb0e13bb8b9 · outbound

This paper cites Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:00.136265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:00.136265Z digest=sha256:bffec160463e9492d322b8021d12892df246884744250a4703cc33425b08eb11

Observation 1c7c66ac-080f-48c8-99d7-7c4238722359 · outbound

This paper cites RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:00.224402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:00.224402Z digest=sha256:024942237c8926a50dcfd3904b71efc15414807b555a0b18070a547fa5e067ba

Observation 2a5816f7-5113-4be4-93f0-6dd0635d1c5f · outbound

This paper cites Model card: Meta llama 3.1 405b instruct.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Meta llama 3.1 405b instruct

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.243026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.298428Z digest=sha256:d14755cb2acd8f2187e9798d31b88546e53ddcf49ffbf8e33eed322eab6c2f16

Observation bf0414bd-6abb-4287-abb1-2a91bb278ba2 · outbound

This paper cites Codeforces, 2010.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Codeforces, 2010

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.092759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.390010Z digest=sha256:1ef5ecfab1d4c834c22cd0a22575ac840750d2494c70d81977e48d7cfd9f2221

Observation a109ece0-d067-4063-80b1-2c9b04adcf97 · outbound

This paper cites System card: Gpt-4o.https://openai.com/index/gpt-4o-system-card/, 2024.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Gpt-4o.https://openai.com/index/gpt-4o-system-card/, 2024

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.929728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.496630Z digest=sha256:a95903940b4eafeed877bff9b2afd267f21f25a0e60faf6261784388a1868a7e

Observation 4bdc3cca-17ab-44b2-8444-704b6360f6e1 · outbound

This paper cites Release note: Gpt -4.1.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Release note: Gpt -4.1

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.742707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.581780Z digest=sha256:0f08eb9ebeda3dd27d3fefba1749a8859606cd64ed52d78990846488a0227c70

Observation 1ec49e91-dc12-4df7-bc73-6dcfa7d1caa9 · outbound

This paper cites Release note: Gpt -4.1 mini.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Release note: Gpt -4.1 mini

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.585966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.645729Z digest=sha256:92372b0b2cf4fad425817ef2af01a560f45b405a28f75a1dcbc2ff4248d99a4c

Observation 3bbe85c7-ba89-445e-b49b-a841293f34cb · outbound

This paper cites System card: Gpt-4.5.https://openai.com/index/gpt-4-5-system-card/, 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Gpt-4.5.https://openai.com/index/gpt-4-5-system-card/, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.430459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.738149Z digest=sha256:7713b02d29d0b50d1645703cb9d3e7c73bb53ccd62a7cfdea9b7a0ed62923ba9

Observation 82afb4a0-dc7b-4d0c-878a-207275dc207b · outbound

This paper cites System card: Openai o3 -mini.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Openai o3 -mini

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.235377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.762122Z digest=sha256:e02f138d046b9096373239d31d3444f2c08b86ad52dcc79d1efc6ca160f28ed8

Observation 72585f97-7f07-47fd-ab89-890f50257977 · outbound

This paper cites System card: Openai o4 -mini (including the o4-mini-high variant).

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Openai o4 -mini (including the o4-mini-high variant)

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.021170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.784875Z digest=sha256:13fa43a3be8e30b2999ae87a310ab3f0a18aa01c57172368f334e4852cae92f7

Observation e4266e79-9914-4b7f-ac7b-5b9adc9071d3 · outbound

This paper cites Introducing openai o3 and o4-mini.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Introducing openai o3 and o4-mini

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:05.763343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:00.804479Z digest=sha256:f5fca2734c7c160f2b0dfefe9b9cc7c01227c88bdc7128d35fe979b1deceb82f

Observation 90499ba3-6223-499a-a37c-87532b727c53 · outbound

This paper cites CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:00.872714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:00.872714Z digest=sha256:ee0ae38eb94f2510e0031cb5c3fea188beaded9a00b8bbcebead797a834004af

Observation f12630cb-c19d-4557-aeda-5b50fccd89e1 · outbound

This paper cites Model card: Llama 4 maverick 17b instruct.https://huggingface.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Llama 4 maverick 17b instruct.https://huggingface

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:05.486831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:01.002796Z digest=sha256:3c20f594f67e544a936179bc16c18898569d26e32806fac757d9469f021792dd

Observation 37011065-6ddf-4089-9c7f-356ad251e219 · outbound

This paper cites Can Language Models Solve Olympiad Programming?.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Can Language Models Solve Olympiad Programming?

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:01.161542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:01.161542Z digest=sha256:ec88c9a35bf96d910ee453ae72e83dc438144eca1ce7b1e417bcdd232ac0286d

Observation 4cd6fca4-66f2-4931-948f-c2a9a702ef2e · outbound

This paper cites New features: friends, tags and more, 2011.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? New features: friends, tags and more, 2011

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:05.241322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:01.269488Z digest=sha256:1c7d5f04d92c24a7a86b0f83308737c54e8a670694a0b85840ac17e016538333

Observation 96575616-7b23-4d58-bd45-682e4b209017 · outbound

This paper cites Learning task decomposition to assist humans in competitive programming.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Learning task decomposition to assist humans in competitive programming

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:05.139259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:01.345958Z digest=sha256:3e819f1e66a74826fc7af1beb97e68532ef5626652323f4096f79e5f28d0eba4

Observation 88adc4bf-2f71-4e37-9529-14dff9aec9bb · outbound

This paper cites LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:01.421727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:01.421727Z digest=sha256:5c0e5421797e3e427584c85461f1e55fc14eb0d9de88a65d5186e7583e92a0d8

Observation 43990623-9f95-4e6f-b41c-66e9fdb93613 · outbound

This paper cites Evaluating the smooth control of attribute intensity in text generation with llms.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Evaluating the smooth control of attribute intensity in text generation with llms

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.979832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:01.521759Z digest=sha256:f4588331abbf906564a52adf21d112b176aee86064012f46b286cd9a42e6f41e

Observation 4f4b5262-6124-48f8-88d6-2e78c2aa831f · outbound

This paper cites BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:01.638982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:01.638982Z digest=sha256:7e727712e7ae6d5e849fa5d8442a55013e4b776df3430c924c37f14f69934920

Observation 3636a231-c72e-41fa-a983-710c247cf86a · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:04.845244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:01.709186Z digest=sha256:7a55a2d1d5bd47179bcfa2470c2aba883e59af82448de1181c34c4a28a54bd86

Observation f5f19399-1ea9-4735-b99b-161cb8d89466 · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:04.707083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:01.791116Z digest=sha256:95aeb6aa8c4e86cb1fa1ff7faf2e65572dc45904a190463693fa226d31029d37

Observation 2d1a3438-f692-4830-9cd2-5358542e9963 · outbound

This paper cites Fails Sample.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Fails Sample

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.555144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:01.885531Z digest=sha256:ecf3822562c5239822be649fbbeba29cf67656ad8142d096d29ff2915733bd52

Observation d4c6fd24-4d4e-4a71-928e-7810bcaaa26e · outbound

This paper cites Therefore every element that appears before that minimum in the original array must be moved to the back (and thus increased by 1).

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Therefore every element that appears before that minimum in the original array must be moved to the back (and thus increased by 1)

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.394086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:01.957114Z digest=sha256:f686c37d41ad49360b876278ccbe54be90e07aa2781b5ee9e67ddbe613f5663d

Observation 8935f76a-d61b-4745-b55b-1cf526007798 · outbound

This paper cites 56 LiveCodeBench Pro.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? 56 LiveCodeBench Pro

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.211737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:02.042300Z digest=sha256:8b164afca0744b88ce36c5fab46196ff21e361f54089dde443f06c71763415ff

Observation a3a469b2-1b30-45e8-9ebc-d28b2990eb5b · outbound

This paper cites suffix minima sequence.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? suffix minima sequence

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.061712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:02.169936Z digest=sha256:727cbfc48a6cdfadf6a766c06c9565cbe9183d1e8ba93234e287ae671c928368

Observation fd2468dc-9169-443c-b8ec-bb0d943be1fd · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:03.869637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:02.235507Z digest=sha256:2aa19abbc53b51e7de1d2d0ca8dfbc714d8d44b8aa544cd1a6902065d662bd00

Observation ba0ca66c-6d6d-4804-9d39-58eef4254009 · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:03.689735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:02.411334Z digest=sha256:00c588ea61d824ed771086eca6915ad919af6a45871d571e96a93b35767e6dd9

Observation 93f38720-770a-4968-a3f6-5ce6f7a481e2 · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:03.504941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:02.527033Z digest=sha256:bac88f43f9042d7bffd6bc9abf8d1cfce8875eba9b6da15d9f8e0e4ff4a4b99b

Observation 1c0aa292-0487-41dd-833a-a6abbee63e8f · outbound

This paper cites op",i, "moved id.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? op",i, "moved id

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:03.317359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T01:08:02.659078Z digest=sha256:2bb0c0a3a4ec1b6cec8b8b49ed2c25fde0597b6d02388b8e28a2551b451ebed5

Pith citing papers

Observation c7195200-d475-4c66-a2ce-d0b42a41c470 · inbound

Evaluating and Improving Large Language Models for Competitive Program Generation cites this paper.

Evaluating and Improving Large Language Models for Competitive Program Generation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:03:28.866165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:03:28.866165Z digest=sha256:eadf4f2dc5b94b5e140af747703576509cb75c1d5995991c8f0426676cee1d90

Observation 71430e1e-2df9-4de9-8f9f-1a75c545da80 · inbound

When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions cites this paper.

When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T13:38:53.452180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:38:53.452180Z digest=sha256:22018c0ce0e281cef9e32c3384bd2d530e9a9d711569d6499b623f96e8640e57

Observation b367fcbe-3f42-4b12-972c-c2789e8c36da · inbound

AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators cites this paper.

AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:28.722411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:18:28.722411Z digest=sha256:38fa49cf40d175f0f59878672aa868d34bd4ba96767965923306205a5e6edd0b

Observation 2d323982-ee0c-4726-8350-623cdeec47c1 · inbound

SWE-IF: Aligning Code Evaluation with Human Preference cites this paper.

SWE-IF: Aligning Code Evaluation with Human Preference LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T11:02:11.745455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:02:11.745455Z digest=sha256:cbf581e6f8dfe6fb7571d004e45aa8173921eda230bdea4f514a7bea27937b24

Observation 5097b6ee-bf1c-41cf-9363-d406b73278e7 · inbound

Seed1.8 Model Card: Towards Generalized Real-World Agency cites this paper.

Seed1.8 Model Card: Towards Generalized Real-World Agency LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.505943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T07:44:02.827006Z digest=sha256:617c8ddc1aad028130a37c2b34e983357167b7b5fdfc8c260cdb67d183b0fef9

Observation 730970d6-5bdf-4565-9ba9-86f803a259eb · inbound

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation cites this paper.

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:26:03.637624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T14:54:51.051959Z digest=sha256:48c87ff4d2191ecd78520e6ed60fa8a088bac0062c296df2df9287cf92d79ec1

Observation df670645-a671-4e74-a829-470279966bf0 · inbound

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data cites this paper.

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 50

Resolution
malformed identifier
arxiv_id, observed 2026-05-15T00:18:21.825093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T00:18:03.185565Z digest=sha256:b405f3645d046e7674764c5a5d8b5c60b01c75f2ebb3da66f361251cf4bccab3

Observation 7ae241cc-42ad-4e3a-9c25-9e2ac774aacd · inbound

When Independent Sampling Outperforms Agentic Reasoning cites this paper.

When Independent Sampling Outperforms Agentic Reasoning LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:24.650752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T02:46:48.375901Z digest=sha256:aca098fcd202db3411a2543b50affaeef727a66833f80b6f148ed3feb2f64538

Observation 2418976e-ec38-418d-b92d-db719289bc63 · inbound

OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation cites this paper.

OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:08:58.496946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T03:06:54.507594Z digest=sha256:5eec9d9e718ebcb24a0fdb7aa9d1a2bf3e52e4e1ecf4de99813c4cc6ba2cd0e7

Observation d47c6f4b-b4ff-4808-b7fb-4be2ddcef824 · inbound

OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation cites this paper.

OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:53:43.632151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T20:51:47.393589Z digest=sha256:83d20b3d84041331e62a8f1095a565ab357706e2db32edc296fa21038571337c

Observation 5f6ae091-65ae-4ec5-ba5d-d3daf5545fc7 · inbound

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language cites this paper.

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:13:40.360955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T19:12:42.655490Z digest=sha256:a7e23ff9613afe094497e4a5aec63732a101f6cfcc22e33c6c35afec61fb1bd4

Observation ddf3f5fc-6ef1-4643-ac82-0bf69b4b469d · inbound

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language cites this paper.

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T05:10:07.337213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:10:07.337213Z digest=sha256:461520d0fd20ed04fe67f90c8b233ddc352610bd55a5238b78f729cdd59bf366

Observation 22285609-f64f-414a-8922-67919d027659 · inbound

Multi-LCB: Extending LiveCodeBench to Multiple Programming Languages cites this paper.

Multi-LCB: Extending LiveCodeBench to Multiple Programming Languages LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:30.017890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T17:39:27.723253Z digest=sha256:0c997eb68652a18bcfd891739fc7a88d341b066416a8a69cd943f20ddb722c79

Observation 27459532-5311-4797-bfbd-02367b28ed9a · inbound

AxDafny: Agentic Verified Code Generation in Dafny cites this paper.

AxDafny: Agentic Verified Code Generation in Dafny LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:55:41.400062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-01T05:01:14.526987Z digest=sha256:b8f8082a413b82fd4facf27088bc61530cdda85555c927e7d7b44bcd31c8cae8

Observation 81c56263-0e36-44a4-8479-47a99fab39fb · inbound

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity cites this paper.

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 143

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:17.391406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T18:57:46.841456Z digest=sha256:d3bd33ed2cc1c12deeeb2fa5bb3fa88a71c0b5e0922fb1dd993f85f7ab777cce

Observation 5ec4a4ed-614e-45ce-b287-45738d798b37 · inbound

PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization cites this paper.

PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T12:12:58.486710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:12:58.486710Z digest=sha256:e37d7b8e9343b8517076fc3decd15a0297f4cd2dc35e0debaaa89dc4ea6b32ac

Observation 9a3dfbe5-1a45-452a-9e72-65be6d87cc1d · inbound

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks cites this paper.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.058356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.058356Z digest=sha256:b7946b2d2fe4992388f5f6a6eb3b0050a12e968601d2914e7917a5d7c2402cc0