Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T01:08:02.659078Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 17 inbound Pith citation observations for arXiv:2506.11928.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T01:08:02.659078Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:03:28.866165Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T03:49:30.015976Z
66 of 66 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7e61df85-12b3-4408-8dd3-3d2b2ecd8704 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://github.com/openai/human-eval
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fbc2a48f-8ec1-4527-bd64-2946931f1197 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://icpc.foundation/
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a95ef9c6-6065-413f-8639-c5ba3ee1626f · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://icpc.global/
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bb6c5693-05d5-4e4c-957a-a4859a81bbea · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://ioinformatics.org/
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 601c7cb8-a1f1-4da4-bdc6-20c95ffc6422 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://mitit.org/About
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0a51eef2-ed57-464b-b5c0-71273b4bf5c6 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://noi.cn/
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fefce80b-9ae5-4ed1-be0f-59f1f83870dd · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://thusaac.com/public
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1f30c23d-5f85-48b5-8115-5c728e016eeb · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URLhttps://usaco.org/
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9279a984-dee4-4fe5-bbe0-aef6d947a445 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? URL https://icpc.global/worldfinals/fact-sheet/ ICPC-Fact-Sheet.pdf
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e19e66de-0cf3-47cb-a0d1-49421139d572 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card addendum: Claude 3.5 sonnet
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 534e24e2-3845-4559-b68e-eb40de527b50 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Claude 3.5 Sonnet, 2024
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 890b873e-e208-449b-8b1a-ecdc8bf3bae4 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Claude 3.7 sonnet (max reasoning)
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bc025a9b-ae8b-4298-b6f4-d1ed636e0cc4 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Claude 3.7 sonnet (no reasoning variant)
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7cc2a4d9-b654-4841-8151-b7ec4c9f343c · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Program Synthesis with Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 49e3dee7-7ebe-4c34-80a3-9705146f1833 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Evaluating Large Language Models Trained on Code
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa5e5b86-f8b7-4ffd-9135-4b704838e165 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Official website of the china collegiate programming contest (ccpc), 2025
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b9d895ed-2165-48cd-983f-822bc8f7ee18 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Qwen-max (qwen 2.5 max).https://huggingface.co/Qwen, 2025
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 98b590cb-eb67-49ab-a45a-f735e7573275 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Interactive Problems: Guide for Participants, 2015
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 44140a0d-feac-4d1e-9f8f-0e06318b4424 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? MHPP: Exploring the Capabilities and Limitations of Language Models Beyond Basic Code Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da2e4af5-431b-4d31-96ad-d58262e1a9f2 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemini 2.0 flash reasoning.https://blog.google/technology/ google-deepmind/gemini-model-updates-february-2025/, 2025
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d9d37b68-7746-4afc-b2f1-5cd7fcbbc93e · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemini 2.5 flash
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ea9583b1-adee-44f9-bcea-6384fbfe084e · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemini 2.5 pro
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b71899ab-f791-46a4-8e37-76cfc634e293 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Gemma 3 27b
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fb141544-2751-47ee-9bae-772200c57b33 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek v3
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ab040e01-ad0d-481a-8131-0bc4fd6f0d76 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek r1
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0341b531-4233-4697-8669-0999687fa032 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek -r1-distill-llama-70b
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5330032-db56-4de9-980d-cedff58496b3 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Deepseek -v3-0324
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c3457360-fb3c-4697-85d8-6269f99c6dbd · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Crosscodeeval: A diverse and multilingual benchmark for cross-file code completion.Advances in Neural Information Processing Systems, 36:46701–46723, 2023
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0a4544c9-699f-4bc4-baff-eb138c2dd571 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Evaluating the performance of large language models in competitive programming: A multi-year, multi-grade analysis
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b8acc336-1679-4fcb-9886-f24733f45d75 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Mathematics and games.Eureka, 2:6–8, 1939
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48e7c5d4-46a5-4548-a2da-b95fd9ebe06d · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c2e3929-fc0a-4716-a136-adb14b1f5c78 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Llm-pros: Analyzing large language models’ performance in competitive problem solving.arXiv preprint arXiv:2502.04355, 2025
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5637c6f3-e880-4a07-a26c-5372d634466d · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Competition-Level Problems are Effective LLM Evaluators
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64f8c4ad-35fb-4f20-ba1f-50281ce84772 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? OpenAI o1 System Card
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09e0b0da-7573-4402-a07b-62a162f0725c · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eacc619b-c547-4daa-9675-869670242216 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Swe-bench: Can language models resolve real-world github issues? InThe Twelfth International Conference on Learning Representations, 2024
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b8f52cd7-2864-4d56-ba98-dfc61aabea0a · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Competition-level code generation with alphacode.Science, 378(6624):1092–1097, 2022
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3a57fd1f-cf44-4052-92ac-fdb0e13bb8b9 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Is your code generated by chatgpt really correct? rigorous evaluation of large language models for code generation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c7c66ac-080f-48c8-99d7-7c4238722359 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a5816f7-5113-4be4-93f0-6dd0635d1c5f · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Meta llama 3.1 405b instruct
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bf0414bd-6abb-4287-abb1-2a91bb278ba2 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Codeforces, 2010
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a109ece0-d067-4063-80b1-2c9b04adcf97 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Gpt-4o.https://openai.com/index/gpt-4o-system-card/, 2024
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4bdc3cca-17ab-44b2-8444-704b6360f6e1 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Release note: Gpt -4.1
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1ec49e91-dc12-4df7-bc73-6dcfa7d1caa9 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Release note: Gpt -4.1 mini
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3bbe85c7-ba89-445e-b49b-a841293f34cb · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Gpt-4.5.https://openai.com/index/gpt-4-5-system-card/, 2025
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 82afb4a0-dc7b-4d0c-878a-207275dc207b · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Openai o3 -mini
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 72585f97-7f07-47fd-ab89-890f50257977 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? System card: Openai o4 -mini (including the o4-mini-high variant)
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e4266e79-9914-4b7f-ac7b-5b9adc9071d3 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Introducing openai o3 and o4-mini
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 90499ba3-6223-499a-a37c-87532b727c53 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? CodeElo: Benchmarking Competition-level Code Generation of LLMs with Human-comparable Elo Ratings
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f12630cb-c19d-4557-aeda-5b50fccd89e1 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Model card: Llama 4 maverick 17b instruct.https://huggingface
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 37011065-6ddf-4089-9c7f-356ad251e219 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Can Language Models Solve Olympiad Programming?
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4cd6fca4-66f2-4931-948f-c2a9a702ef2e · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? New features: friends, tags and more, 2011
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 96575616-7b23-4d58-bd45-682e4b209017 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Learning task decomposition to assist humans in competitive programming
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 88adc4bf-2f71-4e37-9529-14dff9aec9bb · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? LeetCodeDataset: A Temporal Dataset for Robust Evaluation and Efficient Training of Code LLMs
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43990623-9f95-4e6f-b41c-66e9fdb93613 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Evaluating the smooth control of attribute intensity in text generation with llms
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4f4b5262-6124-48f8-88d6-2e78c2aa831f · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3636a231-c72e-41fa-a983-710c247cf86a · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f5f19399-1ea9-4735-b99b-161cb8d89466 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2d1a3438-f692-4830-9cd2-5358542e9963 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Fails Sample
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d4c6fd24-4d4e-4a71-928e-7810bcaaa26e · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Therefore every element that appears before that minimum in the original array must be moved to the back (and thus increased by 1)
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8935f76a-d61b-4745-b55b-1cf526007798 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? 56 LiveCodeBench Pro
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a3a469b2-1b30-45e8-9ebc-d28b2990eb5b · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? suffix minima sequence
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fd2468dc-9169-443c-b8ec-bb0d943be1fd · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ba0ca66c-6d6d-4804-9d39-58eef4254009 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 93f38720-770a-4968-a3f6-5ce6f7a481e2 · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? Unresolved cited work
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1c0aa292-0487-41dd-833a-a6abbee63e8f · outbound
LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming? op",i, "moved id
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c7195200-d475-4c66-a2ce-d0b42a41c470 · inbound
Evaluating and Improving Large Language Models for Competitive Program Generation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71430e1e-2df9-4de9-8f9f-1a75c545da80 · inbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b367fcbe-3f42-4b12-972c-c2789e8c36da · inbound
AutoCodeBench: Large Language Models are Automatic Code Benchmark Generators LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d323982-ee0c-4726-8350-623cdeec47c1 · inbound
SWE-IF: Aligning Code Evaluation with Human Preference LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5097b6ee-bf1c-41cf-9363-d406b73278e7 · inbound
Seed1.8 Model Card: Towards Generalized Real-World Agency LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 730970d6-5bdf-4565-9ba9-86f803a259eb · inbound
CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation df670645-a671-4e74-a829-470279966bf0 · inbound
How to Fine-Tune a Reasoning Model? A Teacher-Student Cooperation Framework to Synthesize Student-Consistent SFT Data LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7ae241cc-42ad-4e3a-9c25-9e2ac774aacd · inbound
When Independent Sampling Outperforms Agentic Reasoning LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2418976e-ec38-418d-b92d-db719289bc63 · inbound
OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d47c6f4b-b4ff-4808-b7fb-4be2ddcef824 · inbound
OpenDeepThink: Parallel Reasoning via Bradley-Terry Aggregation LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5f6ae091-65ae-4ec5-ba5d-d3daf5545fc7 · inbound
Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ddf3f5fc-6ef1-4643-ac82-0bf69b4b469d · inbound
Syntax Without Semantics: Teaching Large Language Models to Code in an Unseen Language LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22285609-f64f-414a-8922-67919d027659 · inbound
Multi-LCB: Extending LiveCodeBench to Multiple Programming Languages LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 27459532-5311-4797-bfbd-02367b28ed9a · inbound
AxDafny: Agentic Verified Code Generation in Dafny LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 81c56263-0e36-44a4-8479-47a99fab39fb · inbound
Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 143
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5ec4a4ed-614e-45ce-b287-45738d798b37 · inbound
PerfAgent: Profiler-Guided Iterative Refinement for Repository-Level Code Optimization LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a3dfbe5-1a45-452a-9e72-65be6d87cc1d · inbound
GABench: A Comprehensive Benchmark for Evaluating LLM Agents on Graph Analysis Tasks LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.