Pith. sign in

Paper Citation Record · LEDGER

An Empirical Study of the Non-determinism of ChatGPT in Code Generation

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2308.02828.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.02828 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:20:36.409945Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T03:55:59.649374Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f03f1396-e185-44d4-9960-42fa563fa4db · inbound

CodeMind: Evaluating Large Language Models for Code Reasoning cites this paper.

CodeMind: Evaluating Large Language Models for Code Reasoning An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-24T03:55:59.653750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-24T03:53:55.964755Z digest=sha256:b4c5f41a2f6e4b5b6695c84b4aca5e6f506afdea87cf80300fd4a845b24e86fb

Observation e0fd58f2-ce37-4ae2-8996-b20245ca42ac · inbound

Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation cites this paper.

Prompting and Fine-tuning Large Language Models for Automated Code Review Comment Generation An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T20:00:37.467468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:00:37.467468Z digest=sha256:ab1e4079698a385ddeb919560343519a7ad56fcd40ee474574c60b4f98aa9ee0

Observation 23d8a6fa-a00a-4337-8c41-2717e31af7d3 · inbound

Curse of Attention: A Kernel-Based Perspective for Why Transformers Fail to Generalize on Time Series Forecasting and Beyond cites this paper.

Curse of Attention: A Kernel-Based Perspective for Why Transformers Fail to Generalize on Time Series Forecasting and Beyond An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T20:09:35.201442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:09:35.201442Z digest=sha256:d3fc5cc7b76a899f46b59fca2fb84a4c9f8dc22c986b89315dc1ef4ca342ecfa

Observation c506b427-5351-411f-8261-9cfecc8d43f6 · inbound

Can You Trust LLM Judgments? Reliability of LLM-as-a-Judge cites this paper.

Can You Trust LLM Judgments? Reliability of LLM-as-a-Judge An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T14:04:26.296299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:04:26.296299Z digest=sha256:1436ed689daaecb54a08ef0280047585f401abcafbe2747e1baff146eb432a4d

Observation 3ff8a30a-60a4-4747-87e0-1468c8e951a0 · inbound

Improving the Readability of Automatically Generated Tests using Large Language Models cites this paper.

Improving the Readability of Automatically Generated Tests using Large Language Models An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T04:28:10.232745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:28:10.232745Z digest=sha256:eaaf6e133983031196246c08232e649c1d6e3f1b8510d4e8799b144b8183f681

Observation caeb8e5e-e543-4074-8081-e6e45573e391 · inbound

Linear Feedback Control Systems for Iterative Prompt Optimization in Large Language Models cites this paper.

Linear Feedback Control Systems for Iterative Prompt Optimization in Large Language Models An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T17:42:11.126560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:42:11.126560Z digest=sha256:bbbb0f1ee14e630e1ed087727ba63d9122eced8b92b4d8c925a5fbb5fe1f1b8a

Observation 92acd925-dc8f-46ff-bafe-35e8ffd2b8f2 · inbound

Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature Review cites this paper.

Understanding the LLM-ification of CHI: Unpacking the Impact of LLMs at CHI through a Systematic Literature Review An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 127

Resolution
unresolved
no resolver link, observed 2026-08-10T17:06:55.547924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:06:55.547924Z digest=sha256:18504dbf4514f87714f07c479a8590dfb1b4f86653eaaac75a9ce4f8178bfcd1

Observation 36398d83-49b3-4d53-bc28-9315c526c932 · inbound

LCTG Bench: LLM Controlled Text Generation Benchmark cites this paper.

LCTG Bench: LLM Controlled Text Generation Benchmark An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-10T13:53:45.092902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T13:53:45.092902Z digest=sha256:0e64c793f24245099440067a76a0c392e22e7d5afed5afb7ba7c46d76a761b3a

Observation 0ea245dc-5117-438b-8ac5-6644e675528e · inbound

Trustworthiness in Stochastic Systems: Towards Opening the Black Box cites this paper.

Trustworthiness in Stochastic Systems: Towards Opening the Black Box An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-10T13:10:07.861278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T13:10:07.861278Z digest=sha256:43d727e525644d4852f17e840bb1113f70d52e7dadc49f716d271494967acef7

Observation 77948597-9ab7-4152-991c-acbb4d0f6b72 · inbound

Multiple Abstraction Level Retrieve Augment Generation cites this paper.

Multiple Abstraction Level Retrieve Augment Generation An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T05:33:52.428134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T05:33:52.428134Z digest=sha256:2463b72acc4fb4461e74459887fb2440cf54d409c40646add4174a15d44f03c4

Observation 062d7830-b111-4850-9aa7-cd55e6fb166f · inbound

On Iterative Evaluation and Enhancement of Code Quality Using GPT-4o cites this paper.

On Iterative Evaluation and Enhancement of Code Quality Using GPT-4o An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T12:58:05.623259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:58:05.623259Z digest=sha256:e9153fae2fb90568b85425d9a40256298c28e8e0c7ff349baf4e6a9820510755

Observation f9521dc3-06bb-487f-ad55-40d14889c3dc · inbound

Theoretical Benefit and Limitation of Diffusion Language Model cites this paper.

Theoretical Benefit and Limitation of Diffusion Language Model An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T20:56:23.111096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:56:23.111096Z digest=sha256:c207d0e50f5b9e8874f225b61a7c7e2f70acf868a0442fe0ba2abea09b04f96e

Observation fbffb1b7-0c94-4fd5-960f-8daa18b15ed9 · inbound

Knowledge-Enhanced Program Repair for Data Science Code cites this paper.

Knowledge-Enhanced Program Repair for Data Science Code An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T20:37:30.787061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T20:37:30.787061Z digest=sha256:3781c7d2ebe4f112ca689dd2bbfbe640766be32355bb252df53081a105467d63

Observation b39a50ce-aefd-4a09-ab10-0b6da73ab2c3 · inbound

Inducing Vulnerable Code Generation in LLM Coding Assistants cites this paper.

Inducing Vulnerable Code Generation in LLM Coding Assistants An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T11:20:36.409945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:20:36.409945Z digest=sha256:84f76b86628a5010c16ea02cedf83e49920b63ac7e887091cedfe952025ed539

Observation b519e547-cd5c-44ce-95b6-446d076111b8 · inbound

Do Automatic Comment Generation Techniques Fall Short? Exploring the Influence of Method Dependencies on Code Understanding cites this paper.

Do Automatic Comment Generation Techniques Fall Short? Exploring the Influence of Method Dependencies on Code Understanding An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T05:56:23.666478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:56:23.666478Z digest=sha256:9d7ccce01e3d2e7ba6fd3dcc5d6fa9c47a088cf7fc2593cae6cc79c0dd9f0f89

Observation 4aca4d47-e4b5-4494-8911-32ae20c3b70b · inbound

Leveraging Large Language Models for Command Injection Vulnerability Analysis in Python: An Empirical Study on Popular Open-Source Projects cites this paper.

Leveraging Large Language Models for Command Injection Vulnerability Analysis in Python: An Empirical Study on Popular Open-Source Projects An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:29:06.864738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:29:06.864738Z digest=sha256:87ffad51df581749162e43fb5554631f1664e6bbe1faac79fabb7e19c32a978b

Observation 5cce1f86-dc76-4c07-9735-5c24990f59cb · inbound

Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs cites this paper.

Fault Localisation and Repair for DL Systems: An Empirical Study with LLMs An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T11:10:32.049198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:10:32.049198Z digest=sha256:650056104dbfaace471ad7ff4594ecc085e819c1b8b6a6d706e49fb7692ce8ef

Observation 974da0fe-cf1f-4843-85a9-b7f554ac70dc · inbound

Synthetic Heuristic Evaluation: A Comparison between AI- and Human-Powered Usability Evaluation cites this paper.

Synthetic Heuristic Evaluation: A Comparison between AI- and Human-Powered Usability Evaluation An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T20:38:56.675833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:38:56.675833Z digest=sha256:39d12ff8fb94f6c40a3d4410b8e40fef75a3c34847b93f21898f5f6af90b61e1

Observation 3c032d82-af34-4d65-8be7-5c181400e7f9 · inbound

From Prompt to Pipeline: Large Language Models for Scientific Workflow Development in Bioinformatics cites this paper.

From Prompt to Pipeline: Large Language Models for Scientific Workflow Development in Bioinformatics An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T17:52:52.443803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:52:52.443803Z digest=sha256:17124297257506cd6884faf339dc4d0619ca50946a06d17e0ea5eccc928f9885

Observation edcfbdb7-c434-4a4d-a75a-b37ee0f7422a · inbound

Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation cites this paper.

Compiled AI: Deterministic Code Generation for LLM-Based Workflow Automation An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:40:51.582602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T18:57:34.386632Z digest=sha256:138057bee09c96d1c5f1a23acd5392ff3225bc7bc0ab339972de968fcf9be34c

Observation 1d1cc71e-c609-41cd-be06-c078ef862029 · inbound

Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation cites this paper.

Co-Located Tests, Better AI Code: How Test Syntax Structure Affects Foundation Model Code Generation An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:01:05.205896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-10T04:21:39.637962Z digest=sha256:0754407a1b21a364d20423175b495c8ec68eef1633d7ea069a38dd767337ad01

Observation 8605fe87-f9ce-485e-a020-e9a7974deccc · inbound

Introducing Background Temperature to Characterise Hidden Randomness in Large Language Models cites this paper.

Introducing Background Temperature to Characterise Hidden Randomness in Large Language Models An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:21:07.514412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-08T12:14:47.518940Z digest=sha256:b7de3a74bc7e986a19d0420bb426c0f4d0c47aee5827377623a9b282b1e034be

Observation bf75eb89-5d89-4956-8e2c-226473f03b92 · inbound

Where Does the Noise Come From? A Variance-Components Decomposition of Non-Determinism in LLM Brand Answers cites this paper.

Where Does the Noise Come From? A Variance-Components Decomposition of Non-Determinism in LLM Brand Answers An Empirical Study of the Non-determinism of ChatGPT in Code Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T05:40:05.494011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:40:05.494011Z digest=sha256:17e55b6bd109c9be15cbd8f4207447d734be368a1a1326c059e05395ec2a9acf