Pith. sign in

Paper Citation Record · LEDGER

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

As of 18 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 17 inbound Pith citation observations for arXiv:2506.11928.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11928 v1

Coverage vector

measured 66 of 66 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T01:08:02.659078Z

measured 83 of 83 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 17 of 17 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:03:28.866165Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T03:49:30.015976Z

Reference resolution

66 of 66 outbound references displayed

  • verified exact0
  • verified fuzzy46
  • unresolved20
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7e61df85-12b3-4408-8dd3-3d2b2ecd8704 · outbound

This paper cites URLhttps://github.com/openai/human-eval.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://github.com/openai/human-eval

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.912248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:57.650121Z digest=sha256:1fe3448e430e9e82335f3dfd141d826b5dbb8c4c38bc74312b44911895364d33

Observation fbc2a48f-8ec1-4527-bd64-2946931f1197 · outbound

This paper cites URLhttps://icpc.foundation/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://icpc.foundation/

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.710948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:57.704459Z digest=sha256:a9d400e9e430d9b5104717609c202a444eca3cb98954c3d731d59a090b10547a

Observation a95ef9c6-6065-413f-8639-c5ba3ee1626f · outbound

This paper cites URLhttps://icpc.global/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://icpc.global/

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.550852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:57.729387Z digest=sha256:429eb5340559fd6e181a189ede6f1fbf8ab731461e8676ebbf705944f96c5bf3

Observation bb6c5693-05d5-4e4c-957a-a4859a81bbea · outbound

This paper cites URLhttps://ioinformatics.org/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://ioinformatics.org/

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.397852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:57.772854Z digest=sha256:9414c120b8804ee56744e80327eb367bf10c92d38d33db59fc6724e16f167acf

Observation 601c7cb8-a1f1-4da4-bdc6-20c95ffc6422 · outbound

This paper cites URLhttps://mitit.org/About.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://mitit.org/About

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:11.208163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:57.863687Z digest=sha256:5c925664ce0234cbb92c4830472568b7c26ad6edc332ee59f90bd8753437470a

Observation 0a51eef2-ed57-464b-b5c0-71273b4bf5c6 · outbound

This paper cites URLhttps://noi.cn/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://noi.cn/

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.982844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:57.954851Z digest=sha256:caee8ea37567f0589821f694d53d9abde4d4903224df8c9f14e626241626cd96

Observation fefce80b-9ae5-4ed1-be0f-59f1f83870dd · outbound

This paper cites URLhttps://thusaac.com/public.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://thusaac.com/public

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.795735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.039732Z digest=sha256:2996a396b1cf0f5e75bdc58721e47144515834873aadfff71b7b4a7daae436b2

Observation 1f30c23d-5f85-48b5-8115-5c728e016eeb · outbound

This paper cites URLhttps://usaco.org/.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://usaco.org/

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.647479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.125534Z digest=sha256:4f5bc638b93007f590cd6a568afc79673e7220f547a0c8c5877a5c49b43783fc

Observation 9279a984-dee4-4fe5-bbe0-aef6d947a445 · outbound

This paper cites URL https://icpc.global/worldfinals/fact-sheet/ ICPC-Fact-Sheet.pdf.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URL https://icpc.global/worldfinals/fact-sheet/ ICPC-Fact-Sheet.pdf

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.455653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.172537Z digest=sha256:66ca623d3d8fbb86e7f7cd7b1a53040e274e5aa2814c1272830a80d9bbb7572c

Observation e19e66de-0cf3-47cb-a0d1-49421139d572 · outbound

This paper cites Model card addendum: Claude 3.5 sonnet.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card addendum: Claude 3.5 sonnet

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.241010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.275845Z digest=sha256:3d939680e42f23899ba6148d0840ce6d828857b36d4bd7648c0708852d340a1d

Observation 534e24e2-3845-4559-b68e-eb40de527b50 · outbound

This paper cites Claude 3.5 Sonnet, 2024.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Claude 3.5 Sonnet, 2024

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:10.054698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.346965Z digest=sha256:72744d2c09533a3b64e2902c80feffe8c69648d7cb7ebd420ca69d884a178b0a

Observation 890b873e-e208-449b-8b1a-ecdc8bf3bae4 · outbound

This paper cites System card: Claude 3.7 sonnet (max reasoning).

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Claude 3.7 sonnet (max reasoning)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.891714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.417539Z digest=sha256:7f4a18dbc6e24b6682a974590c4eb9bd4fd5da9dfddf6c63d743319f261885b2

Observation bc025a9b-ae8b-4298-b6f4-d1ed636e0cc4 · outbound

This paper cites System card: Claude 3.7 sonnet (no reasoning variant).

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Claude 3.7 sonnet (no reasoning variant)

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.713673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.505902Z digest=sha256:94004e9a7c4cba2826a6473f68fb8bfa4a34ba5f7373083ab6f7dac652c4346a

Observation 7cc2a4d9-b654-4841-8151-b7ec4c9f343c · outbound

This paper cites Program Synthesis with Large Language Models.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Program Synthesis with Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:58.590186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:58.590186Z digest=sha256:78a0bc9a70a3bc707564ac6d766210a29a8f6bc0c4bfc004894c8ce8f64c8efd

Observation 49e3dee7-7ebe-4c34-80a3-9705146f1833 · outbound

This paper cites Evaluating Large Language Models Trained on Code.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Evaluating Large Language Models Trained on Code

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:58.655057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:58.655057Z digest=sha256:9a28c9bc5aec825fed39657e583d6a684ec7b14a0cc5f3b4e165edcea118e81c

Observation aa5e5b86-f8b7-4ffd-9135-4b704838e165 · outbound

This paper cites Official website of the china collegiate programming contest (ccpc), 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Official website of the china collegiate programming contest (ccpc), 2025

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.563332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.763693Z digest=sha256:988d7697f82ff2fbec379ec8c804fe96d0de2b410aaf2612002eb3f30b99e9e2

Observation b9d895ed-2165-48cd-983f-822bc8f7ee18 · outbound

This paper cites Model card: Qwen-max (qwen 2.5 max).https://huggingface.co/Qwen, 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Qwen-max (qwen 2.5 max).https://huggingface.co/Qwen, 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.403057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.807679Z digest=sha256:8682430b5a2c74c99728862ab634db176e128b828a3ff59ddabc245192632f1c

Observation 98b590cb-eb67-49ab-a45a-f735e7573275 · outbound

This paper cites Interactive Problems: Guide for Participants, 2015.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Interactive Problems: Guide for Participants, 2015

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.273753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.870319Z digest=sha256:1aafb7e7b25b6a555cc8aa035cbd151ccc8116145fc70279ed2a30fb279d523d

Observation 44140a0d-feac-4d1e-9f8f-0e06318b4424 · outbound

This paper cites MHPP: Exploring the Capabilities and Limitations of Language Models Beyond Basic Code Generation.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? MHPP: Exploring the Capabilities and Limitations of Language Models Beyond Basic Code Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:58.930241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:58.930241Z digest=sha256:70c343a031a07d5099b45f259ef48c9131ef0babc10fa203e12a26d2d8fe166d

Observation da2e4af5-431b-4d31-96ad-d58262e1a9f2 · outbound

This paper cites Model card: Gemini 2.0 flash reasoning.https://blog.google/technology/ google-deepmind/gemini-model-updates-february-2025/, 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemini 2.0 flash reasoning.https://blog.google/technology/ google-deepmind/gemini-model-updates-february-2025/, 2025

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:09.105664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:58.983607Z digest=sha256:b2cfc9d005b52bbc23ffe7f1f07730707e043457dd37f87ba9da60f1d02e9b44

Observation d9d37b68-7746-4afc-b2f1-5cd7fcbbc93e · outbound

This paper cites Model card: Gemini 2.5 flash.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemini 2.5 flash

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.931038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:59.042706Z digest=sha256:1aa1f123e5128ef5d778f092bc0a2aa9c0ce19fc2206801779f07c986d9afbba

Observation ea9583b1-adee-44f9-bcea-6384fbfe084e · outbound

This paper cites Model card: Gemini 2.5 pro.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemini 2.5 pro

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.761427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:59.113155Z digest=sha256:e520f77ce3761c3fbb638eb26a1c3b52ad8d42c5a22191810397b3512e352d87

Observation b71899ab-f791-46a4-8e37-76cfc634e293 · outbound

This paper cites Model card: Gemma 3 27b.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemma 3 27b

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.620939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:59.170629Z digest=sha256:4c5c6af321ddd982be27fb22b8adfd38096608a54f45f4a073da44301069684a

Observation fb141544-2751-47ee-9bae-772200c57b33 · outbound

This paper cites Model card: Deepseek v3.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek v3

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.487487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:59.221958Z digest=sha256:ac69e150a1d9659b7bfa122b67ac8be380216914f318bbbc988a564d97fc232d

Observation ab040e01-ad0d-481a-8131-0bc4fd6f0d76 · outbound

This paper cites Model card: Deepseek r1.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek r1

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.343556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:59.307600Z digest=sha256:0dec92c2f34f637cd28eda1548f89dc897ab5ead605306565b84e88ff7262b5e

Observation 0341b531-4233-4697-8669-0999687fa032 · outbound

This paper cites Model card: Deepseek -r1-distill-llama-70b.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek -r1-distill-llama-70b

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.187694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:59.369728Z digest=sha256:3df053ebb11f48cfab72f34b4815c033c2b579815c317ddec0e310c0175406b5

Observation f5330032-db56-4de9-980d-cedff58496b3 · outbound

This paper cites Model card: Deepseek -v3-0324.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek -v3-0324

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:08.031594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:59.460378Z digest=sha256:270a17be19b1ae10fdfbb647f71bd08a374af285153092fbc6040a4918c42112

Observation c3457360-fb3c-4697-85d8-6269f99c6dbd · outbound

This paper cites Crosscodeeval: A diverse and multilingual benchmark for cross-file code completion.Advances in Neural Information Processing Systems, 36:46701–46723, 2023.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Crosscodeeval: A diverse and multilingual benchmark for cross-file code completion.Advances in Neural Information Processing Systems, 36:46701–46723, 2023

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.898533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:59.541691Z digest=sha256:d28f43996e29ea6d3e2e719a6b0717f4c5cc7344ff729c03b9c7195f1921ea9a

Observation 0a4544c9-699f-4bc4-baff-eb138c2dd571 · outbound

This paper cites Evaluating the performance of large language models in competitive programming: A multi-year, multi-grade analysis.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Evaluating the performance of large language models in competitive programming: A multi-year, multi-grade analysis

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.766010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:07:59.591769Z digest=sha256:336591c246c1b07d183846644d9ec675ad048c4b2a5e0c406aeb44e1ceb234c0

Observation b8acc336-1679-4fcb-9886-f24733f45d75 · outbound

This paper cites Mathematics and games.Eureka, 2:6–8, 1939.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Mathematics and games.Eureka, 2:6–8, 1939

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.639019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.639019Z digest=sha256:9ee3f39369104e3d8487feb33c113b5477c387b2694714ba5cb418bdb4a7c557

Observation 48e7c5d4-46a5-4548-a2da-b95fd9ebe06d · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.728279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.728279Z digest=sha256:1a4f022c457f9736d7f3d8c5558a4c4355052475987c32fe7c022b76e40d6db2

Observation 2c2e3929-fc0a-4716-a136-adb14b1f5c78 · outbound

This paper cites Llm-pros: Analyzing large language models’ performance in competitive problem solving.arXiv preprint arXiv:2502.04355, 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Llm-pros: Analyzing large language models’ performance in competitive problem solving.arXiv preprint arXiv:2502.04355, 2025

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.778246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.778246Z digest=sha256:e5102afca81a7ce1519d935fa143a101eb752ce247ba6c01d660ece40ec4e565

Observation 5637c6f3-e880-4a07-a26c-5372d634466d · outbound

This paper cites Competition-Level Problems are Effective LLM Evaluators.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Competition-Level Problems are Effective LLM Evaluators

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.838876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.838876Z digest=sha256:c6a88b572bb35b0d2a26760450bb5743e55d2e0244e98f134bf7f44c67168b5f

Observation 64f8c4ad-35fb-4f20-ba1f-50281ce84772 · outbound

This paper cites OpenAI o1 System Card.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? OpenAI o1 System Card

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.903622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.903622Z digest=sha256:63ffb056a5ca8bbbb283f2dfbd4243daed7ca4ee4bde23b029e728cdc2e88adb

Observation 09e0b0da-7573-4402-a07b-62a162f0725c · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:59.995412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:59.995412Z digest=sha256:3133703bcd3af9a7904a3d2033b941d06c63e8c01abd6fce806b2e29e08b4499

Observation eacc619b-c547-4daa-9675-869670242216 · outbound

This paper cites Swe-bench: Can language models resolve real-world github issues? InThe Twelfth International Conference on Learning Representations, 2024.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Swe-bench: Can language models resolve real-world github issues? InThe Twelfth International Conference on Learning Representations, 2024

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.585390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.043224Z digest=sha256:19803f618debfbd6f035d6e898017b1d419ad33655149e627209b679cd473aab

Observation b8f52cd7-2864-4d56-ba98-dfc61aabea0a · outbound

This paper cites Competition-level code generation with alphacode.Science, 378(6624):1092–1097, 2022.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Competition-level code generation with alphacode.Science, 378(6624):1092–1097, 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.400024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.106253Z digest=sha256:e3ef1241f5e04521884e7a61c44a81f2dedf79c29cc2f4784f6a4c28d3e7e997

Observation 3a57fd1f-cf44-4052-92ac-fdb0e13bb8b9 · outbound

This paper cites Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:00.136265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:00.136265Z digest=sha256:584c6367bda1b6cba9b639ec39c7e2487d1d2422544fbc8d3aa018476d4d1627

Observation 1c7c66ac-080f-48c8-99d7-7c4238722359 · outbound

This paper cites RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:00.224402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:00.224402Z digest=sha256:0415e0a8e9101f87f24b4f69a9893b1ca10544bc5d3889b8d9ad2ac6050e3095

Observation 2a5816f7-5113-4be4-93f0-6dd0635d1c5f · outbound

This paper cites Model card: Meta llama 3.1 405b instruct.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Meta llama 3.1 405b instruct

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.243026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.298428Z digest=sha256:f9669e536d80fc26189857bb793cd039e496dcd40180cb6c4a734b383ffba562

Observation bf0414bd-6abb-4287-abb1-2a91bb278ba2 · outbound

This paper cites Codeforces, 2010.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Codeforces, 2010

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:07.092759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.390010Z digest=sha256:7fb4881446dffa5d45e99e060f8c40142b477290a01d30a8d91cf1c3b0cf6d6c

Observation a109ece0-d067-4063-80b1-2c9b04adcf97 · outbound

This paper cites System card: Gpt-4o.https://openai.com/index/gpt-4o-system-card/, 2024.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Gpt-4o.https://openai.com/index/gpt-4o-system-card/, 2024

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.929728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.496630Z digest=sha256:b7d6d7e7b3bfcd893cd12f6bd6faa89346499dd7258485f96ee7754e04a5559e

Observation 4bdc3cca-17ab-44b2-8444-704b6360f6e1 · outbound

This paper cites Release note: Gpt -4.1.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Release note: Gpt -4.1

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.742707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.581780Z digest=sha256:860caf53776438af4e3c43d4eaa905ee5a5470a8c4c4e28015bdf34979a37d0f

Observation 1ec49e91-dc12-4df7-bc73-6dcfa7d1caa9 · outbound

This paper cites Release note: Gpt -4.1 mini.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Release note: Gpt -4.1 mini

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.585966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.645729Z digest=sha256:bf6cddb5ef348799b871ea8e932a0cd83f1090b70a5836f09b0a9771ce3a763a

Observation 3bbe85c7-ba89-445e-b49b-a841293f34cb · outbound

This paper cites System card: Gpt-4.5.https://openai.com/index/gpt-4-5-system-card/, 2025.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Gpt-4.5.https://openai.com/index/gpt-4-5-system-card/, 2025

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.430459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.738149Z digest=sha256:ecd0d0101a4b20befdc13b4486cf67491e0dce12f949ccdc3359b91ea4e2b6f6

Observation 82afb4a0-dc7b-4d0c-878a-207275dc207b · outbound

This paper cites System card: Openai o3 -mini.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Openai o3 -mini

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.235377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.762122Z digest=sha256:748e49a0f075a33d7a4d9d4da942084d314301ecf2f70bced215ae69ac2f1dae

Observation 72585f97-7f07-47fd-ab89-890f50257977 · outbound

This paper cites System card: Openai o4 -mini (including the o4-mini-high variant).

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Openai o4 -mini (including the o4-mini-high variant)

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:06.021170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.784875Z digest=sha256:d07b6e3a58d318f566a8b1c702e65d94dbf74f283114baa30967b82da1957d2b

Observation e4266e79-9914-4b7f-ac7b-5b9adc9071d3 · outbound

This paper cites Introducing openai o3 and o4-mini.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Introducing openai o3 and o4-mini

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:05.763343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:00.804479Z digest=sha256:1fcc749517e8cadfb02dd34b77ec7491beb0ff7c8b1d0f5a4c4bb0dd23c89113

Observation 90499ba3-6223-499a-a37c-87532b727c53 · outbound

This paper cites CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:00.872714Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:00.872714Z digest=sha256:f9a53ccc81e047858bc7e8a25a4c6f01e9583f43358a847b3389897a1543755b

Observation f12630cb-c19d-4557-aeda-5b50fccd89e1 · outbound

This paper cites Model card: Llama 4 maverick 17b instruct.https://huggingface.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Llama 4 maverick 17b instruct.https://huggingface

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:05.486831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:01.002796Z digest=sha256:58fa4ab33917af42b439dd6f97c91778f79aec422873130c22a61be4e5107f24

Observation 37011065-6ddf-4089-9c7f-356ad251e219 · outbound

This paper cites Can Language Models Solve Olympiad Programming?.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Can Language Models Solve Olympiad Programming?

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:01.161542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:01.161542Z digest=sha256:d85fdeb9415c75323f180a10f9cb5fca1ffbe5f6d4ca60a149fe9950be21c349

Observation 4cd6fca4-66f2-4931-948f-c2a9a702ef2e · outbound

This paper cites New features: friends, tags and more, 2011.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? New features: friends, tags and more, 2011

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:05.241322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:01.269488Z digest=sha256:2eaf4577226eb92641b92c1926c68ec1613b17393e1c88c990bcb6f8fee8c576

Observation 96575616-7b23-4d58-bd45-682e4b209017 · outbound

This paper cites Learning task decomposition to assist humans in competitive programming.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Learning task decomposition to assist humans in competitive programming

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:05.139259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:01.345958Z digest=sha256:409697fd70d937adc610d4da273c69a4062f21c24d44c09ffaf255d7ba77a9e2

Observation 88adc4bf-2f71-4e37-9529-14dff9aec9bb · outbound

This paper cites LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:01.421727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:01.421727Z digest=sha256:88ebec483cdc0a652dd3d06480cded60262579cffc9082a508f12a4cfea55240

Observation 43990623-9f95-4e6f-b41c-66e9fdb93613 · outbound

This paper cites Evaluating the smooth control of attribute intensity in text generation with llms.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Evaluating the smooth control of attribute intensity in text generation with llms

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.979832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:01.521759Z digest=sha256:b7dac05f66efa6b8680354fea504af70c37a7a58476f75da09c0610b578a28ba

Observation 4f4b5262-6124-48f8-88d6-2e78c2aa831f · outbound

This paper cites BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:01.638982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:08:01.638982Z digest=sha256:496cd6ee9ec48f72f67d44693ebd251d8ab34e38f5f53830507f236fc70b45f3

Observation 3636a231-c72e-41fa-a983-710c247cf86a · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:04.845244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:01.709186Z digest=sha256:1aad2dd71e8c856f274f40661b2c97ffd7ac454d508dc203526badc53c9ea62c

Observation f5f19399-1ea9-4735-b99b-161cb8d89466 · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:04.707083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:01.791116Z digest=sha256:5fcbf014ea8bb46f329e71faf54423d7b9437b66acb42ff8bc5bbcf6301923ee

Observation 2d1a3438-f692-4830-9cd2-5358542e9963 · outbound

This paper cites Fails Sample.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Fails Sample

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.555144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:01.885531Z digest=sha256:f78114f1235ff59141a941bc3defd5b941ecbf7939c2f3d523b4c3eb02fdcbda

Observation d4c6fd24-4d4e-4a71-928e-7810bcaaa26e · outbound

This paper cites Therefore every element that appears before that minimum in the original array must be moved to the back (and thus increased by 1).

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Therefore every element that appears before that minimum in the original array must be moved to the back (and thus increased by 1)

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.394086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:01.957114Z digest=sha256:e61b335610e9eecf2dc5aa99eddac629a699b67e179ea49e2ea9d16db32939e0

Observation 8935f76a-d61b-4745-b55b-1cf526007798 · outbound

This paper cites 56 LiveCodeBench Pro.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? 56 LiveCodeBench Pro

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.211737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:02.042300Z digest=sha256:9bf33830bcdb168aea49ea9f74221082ec3c87c5481ebef1a6f9fd086b094787

Observation a3a469b2-1b30-45e8-9ebc-d28b2990eb5b · outbound

This paper cites suffix minima sequence.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? suffix minima sequence

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:04.061712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:02.169936Z digest=sha256:7cab297d64e110aca7183921806a2164862a5f54f9cb7761c841d73cc8128c2d

Observation fd2468dc-9169-443c-b8ec-bb0d943be1fd · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 63

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:03.869637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:02.235507Z digest=sha256:8dd1968480da8e62858564eabab53ec3dc0e559d36bb791622192f1a55c69434

Observation ba0ca66c-6d6d-4804-9d39-58eef4254009 · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:03.689735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:02.411334Z digest=sha256:79ceb2d702ddf950ae09669fb842af68cdbb2148dfce4431fdf24fcb86ac059e

Observation 93f38720-770a-4968-a3f6-5ce6f7a481e2 · outbound

This paper cites an unresolved cited work.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T01:08:03.504941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:02.527033Z digest=sha256:8b9f570001bdb7d5541a0f1f1f1a0618af0b0d2a4793fb73c6405e1577855922

Observation 1c0aa292-0487-41dd-833a-a6abbee63e8f · outbound

This paper cites op",i, "moved id.

LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? op",i, "moved id

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T01:08:03.317359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T01:08:02.659078Z digest=sha256:c91b39bfefdad3a0bdad534e0fd64d81845f8d680b858ea3661d20bb896e7af2

Pith citing papers

Observation c7195200-d475-4c66-a2ce-d0b42a41c470 · inbound

Evaluating and Improving Large Language Models for Competitive Program Generation cites this paper.

Evaluating and Improving Large Language Models for Competitive Program Generation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T22:03:28.866165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:03:28.866165Z digest=sha256:1d26a5ff2951eca0f141cdbcafb941a972038bed6a5583e0b4ca84b4f8f514a8

Observation 71430e1e-2df9-4de9-8f9f-1a75c545da80 · inbound

When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions cites this paper.

When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T13:38:53.452180Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:38:53.452180Z digest=sha256:9a523b2cb0eddcc1fdea980695f3a4dfa39bdf48fca3b5c51d6d606eed01ce6b

Observation b367fcbe-3f42-4b12-972c-c2789e8c36da · inbound

AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators cites this paper.

AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:28.722411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:18:28.722411Z digest=sha256:1e0522776ec43c1a265ff643d3db1860713af9dc93bac59423f4253e8e9deda8

Observation 2d323982-ee0c-4726-8350-623cdeec47c1 · inbound

SWE-IF: Aligning Code Evaluation with Human Preference cites this paper.

SWE-IF: Aligning Code Evaluation with Human Preference LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T11:02:11.745455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:02:11.745455Z digest=sha256:37f1887299804140ec666320c7d6f3f04d4b8a1074b08926aa3bac337e659e17

Observation 5097b6ee-bf1c-41cf-9363-d406b73278e7 · inbound

Seed1.8 Model Card: Towards Generalized Real-World Agency cites this paper.

Seed1.8 Model Card: Towards Generalized Real-World Agency LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 93

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:45:14.505943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T07:44:02.827006Z digest=sha256:2ed0cadd792a4936de69cbdb49a3ecb6983d2681105f6eabbf529fda7b609b65

Observation 730970d6-5bdf-4565-9ba9-86f803a259eb · inbound

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation cites this paper.

CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:26:03.637624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-10T14:54:51.051959Z digest=sha256:ee4f0cbf5b529ba5e4f9673715af09c9491ddcbd66beb28df2b934c8c1a0256a

Observation df670645-a671-4e74-a829-470279966bf0 · inbound

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data cites this paper.

How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 50

Resolution
malformed identifier
arxiv_id, observed 2026-05-15T00:18:21.825093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T00:18:03.185565Z digest=sha256:487c3eb86cb5e84828c2eb5010d669005cedff845ccb7a994da2378369c3dff0

Observation 7ae241cc-42ad-4e3a-9c25-9e2ac774aacd · inbound

When Independent Sampling Outperforms Agentic Reasoning cites this paper.

When Independent Sampling Outperforms Agentic Reasoning LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:31:24.650752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-12T02:46:48.375901Z digest=sha256:3204d24d6ec930a95cdc1f74e13cd6ccfcad01446ef7369c9196597ff5f938bb

Observation 2418976e-ec38-418d-b92d-db719289bc63 · inbound

OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation cites this paper.

OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:08:58.496946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-15T03:06:54.507594Z digest=sha256:ddd5e6440d485e4280892e79c72925220e5392da051dbdc680056b330afd23e4

Observation d47c6f4b-b4ff-4808-b7fb-4be2ddcef824 · inbound

OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation cites this paper.

OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:53:43.632151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T20:51:47.393589Z digest=sha256:f6e946a9d7b4883453d05da488aa6b306cd79bf42b46ebc71772ae7b29c1d81f

Observation 5f6ae091-65ae-4ec5-ba5d-d3daf5545fc7 · inbound

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language cites this paper.

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:13:40.360955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-20T19:12:42.655490Z digest=sha256:20286c6ebf662945cac6e416cf0d6e312a4343a28a21274395fd5f89bcb49ab0

Observation ddf3f5fc-6ef1-4643-ac82-0bf69b4b469d · inbound

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language cites this paper.

Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T05:10:07.337213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:10:07.337213Z digest=sha256:7475d7be8e48804b2715c64eb4b4977fe7d64e4b053c9c982b644cf2db8a60c7

Observation 22285609-f64f-414a-8922-67919d027659 · inbound

Multi-LCB: Extending LiveCodeBench to Multiple Programming Languages cites this paper.

Multi-LCB: Extending LiveCodeBench to Multiple Programming Languages LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:30.017890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-06-26T17:39:27.723253Z digest=sha256:a9c7ccaf9e012e5e660a31777554d4e727eefd288a4fb3f4954f522edd74ece9

Observation 27459532-5311-4797-bfbd-02367b28ed9a · inbound

AxDafny: Agentic Verified Code Generation in Dafny cites this paper.

AxDafny: Agentic Verified Code Generation in Dafny LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:55:41.400062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-07-01T05:01:14.526987Z digest=sha256:9db36c5147c0194da0ebc329379e832604179c938679e20cc695616da808f764

Observation 81c56263-0e36-44a4-8479-47a99fab39fb · inbound

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity cites this paper.

Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 143

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:07:17.391406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-02T18:57:46.841456Z digest=sha256:6addad68dc7695639c77ca1eec8328ede08f6e2615dc33f8f9c16e9f4bb490b0

Observation 5ec4a4ed-614e-45ce-b287-45738d798b37 · inbound

PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization cites this paper.

PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-01T12:12:58.486710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:12:58.486710Z digest=sha256:592703c7a923b32ade2e7ae5469f53cc4dfa55266310be5dfd6d564bf4349b10

Observation 9a3dfbe5-1a45-452a-9e72-65be6d87cc1d · inbound

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks cites this paper.

GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-04T22:55:34.058356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:55:34.058356Z digest=sha256:f8378bb8f7478f11b981efa25155323024c4b06e6dea5fa1119eeb9266091e1c