Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:38:53.742442Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 11 inbound Pith citation observations for arXiv:2507.20439.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:38:53.742442Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T14:46:30.803463Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
43 of 43 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 9df98d93-dda8-4a14-b4d0-f058be2027d3 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions IEEE Recommended Practice for Software Requirements Specifications
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a27cb455-ca01-4845-932d-4b37bcdd76f2 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23de4ca7-f7e8-4054-8996-c4466765f738 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c6a564b6-62bf-444e-bc17-df086f55d3ee · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Program Synthesis with Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fac6cd1-51e4-41fa-9b8b-0bb56846416b · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 87fd0afb-5e94-4171-9253-850d8306a7ec · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eff87302-791c-471f-ba3b-c2a38e176e51 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Evaluating Large Language Models Trained on Code
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b13455d5-6388-4554-8959-e2e884814991 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6a20980-cb7d-4729-a40e-20159448581b · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 365d18e0-4cd8-49f8-9f39-dde997ede50a · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Systematic Evaluation of GPT-3 for Zero-Shot Personality Estimation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 161d3724-ede3-4f83-94cc-e2353ac70e94 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Measuring Coding Challenge Competence With APPS
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 725e18f3-d124-42a1-99bc-91ff4d39a635 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c056224-3794-45f4-8d08-3fad5c1eee98 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca011c1e-cbf9-4062-8c5e-770820f1c436 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2fed75e1-c953-40cd-b241-b8c68755a4c1 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f85e91c9-e344-4da7-833b-a4e2fd9a7e11 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ee94a9d0-b71b-42cf-8498-dc7de55822a9 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5358355d-f85f-46c2-ab56-11a97fdb9406 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20e2a215-9354-4920-b9b5-6e3907489b99 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a18c74fa-41c2-4b64-9e60-c3939ed13364 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aface219-411c-4fbe-9a10-cfbfea79d036 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 982fa63d-cedf-4c65-9275-22cf2d04f7b1 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 25
Source-reported events for the cited work
correction dated 2022-04-04. Source: crossref record 10.1007/s00766-022-00378-4->10.1007/s00766-021-00367-z:correction, observed 2026-07-11T02:57:50.434999+00:00. This notice travels one citation hop only.
Observation fa8f5d74-5fa1-4823-8f1e-083426e837c7 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39c13da2-363a-4edd-8061-d9b3ceb02ba4 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 45e8c2d1-c011-4765-8ee6-54bf5031f8ad · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 74c4d79d-8833-4016-9175-2a3328b65748 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e34977af-ddde-4c41-b37f-5331bc09f3cb · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc21c2dd-ad0a-4abe-bfc5-21ccb202e113 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7f9f76a7-6b47-4e1d-8575-5cde5af1b13c · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7362a76e-f3f8-4a6d-b97a-5d3d60c1cc65 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Code Llama: Open Foundation Models for Code
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61318ea6-a0f4-41c4-8f1b-7d14faacdcbc · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2542d470-b572-492f-803d-c90b569f308b · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 353dce04-6766-4e73-870b-f95226b2934c · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e37dd0df-15a1-48b8-85a3-2fda37ee3973 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Recent Advances in Software Effort Estimation using Machine Learning
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0cfa7f88-6cea-4a69-8188-34069b47c5d3 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Is Your AI-Generated Code Really Safe? Evaluating Large Language Models on Secure Code Generation with CodeSecEval
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d00f9a7-7715-4037-9396-90903b31b54f · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation cbc77a68-d9ca-4d9d-a33d-29c36136f676 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions DeceptPrompt: Exploiting LLM-driven Code Generation via Adversarial Natural Language Instructions
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a3ba20e-0ce5-4251-8cc3-30bc68fb2477 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions ReCode: Robustness Evaluation of Code Generation Models
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71430e1e-2df9-4de9-8f9f-1a75c545da80 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions LiveCodeBench Pro: How Do Olympiad Medalists Judge LLMs in Competitive Programming?
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f86c9d45-f912-4324-9e17-cf9f78ff88d5 · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a1570c58-f6a5-43b8-b3bc-45a685e623fc · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 365c216c-8b0a-48cd-a867-06d1e6cf07da · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Unresolved cited work
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f0cc826d-8542-49ee-b570-66ae34b1b1de · outbound
When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions Software: Practice and experience 52, 1 (2022), 39–65
Reference 2022
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 198ab81f-943e-426d-98b9-64b4c293c359 · inbound
Defective Task Descriptions in LLM-Based Code Generation: Detection and Analysis When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 70807c01-e5e1-4017-aeca-80197e2fbaef · inbound
When Prompt Under-Specification Improves Code Correctness: An Exploratory Study of Prompt Wording and Structure Effects on LLM-Based Code Generation When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 77d8fa42-aa49-4696-b4b2-df3eb3d40d82 · inbound
ClarifyCodeBench: Evaluating LLMs on Clarifying Ambiguous Requirements for Code Generation When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fe9cb4ff-ad6e-4b67-b6d8-139a902c8018 · inbound
Underspecification does not imply Incoherence: The Risks of Semantic Collapse in Coding Models When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b9f16e6d-8e35-411a-8ed3-6728eba63bfd · inbound
Guiding Human Validation of LLM-Generated Code via Verifiable Literate Programming When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation bb9764c9-2676-429e-8063-7ebddb3de03f · inbound
Obey, Diverge, Collapse: Blind Obedience to Incorrect Instructions Drives Code LLMs to Irrecoverable Code Semantic Collapse When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 158d4e41-4d94-43b7-a6fc-b6a5bb766ec5 · inbound
From Failing to Passing: Evolving Natural Language Prompt Optimization Rules for LLM Code Generation When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bd8e22d-9906-4da1-90b9-62a267294cc7 · inbound
On the risk of coding before testing: An empirical study on LLM-based test generation workflow When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d2e8f86-3a11-4e4b-bc0e-b731365ec346 · inbound
Automatically Evolving Prompt Guidelines for Task-Specific Optimization When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c853259f-c273-4ff6-b6a2-3bac8e2e1e11 · inbound
AssumptionMiner: Extracting, Tracing, and Revising Implicit Assumptions in LLM Code Generation When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b7984f4-f449-4cc1-9e8d-bd32b8930620 · inbound
VClare: Resolving Imperfect Specifications in LLM-Based Verilog Generation When Prompts Go Wrong: Evaluating Code Model Robustness to Ambiguous, Contradictory, and Incomplete Task Descriptions
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.