Pith. sign in

Paper Citation Record · LEDGER

Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2402.16906.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2402.16906 v6

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:49:35.726058Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

3
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 04a7b415-3d41-4651-9b90-e8370852942a · inbound

Specification-Driven Code Translation Powered by Large Language Models: How Far Are We? cites this paper.

Specification-Driven Code Translation Powered by Large Language Models: How Far Are We? Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-23T07:35:28.768708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T07:35:19.736921Z digest=sha256:a36dd0a103166ca0662da1793595a8d170a0c9820f64cd225946ca4a9f6b15c1

Observation 6439f9ff-47aa-407f-b472-50c46a5dbb26 · inbound

MARCO: Meta-Reflection with Cross-Referencing for Code Reasoning cites this paper.

MARCO: Meta-Reflection with Cross-Referencing for Code Reasoning Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:49:35.726058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:49:35.726058Z digest=sha256:ca9d429127668a3a633c53c4f7350d526f133f5e9a52cbffbb9f86999fb8e467

Observation e48c0abb-38b3-46e7-a054-709324835ec9 · inbound

Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach cites this paper.

Enhancing LLM-Based Code Generation with Complexity Metrics: A Feedback-Driven Approach Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:42:48.385369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:42:48.385369Z digest=sha256:7fcfb85537a45bd3c0e699e6104fb29e7e68a8595cf92f91e883fd2ecea5bbd0

Observation f1a9f071-1258-4a48-8b0f-b0dc9ec1bbdf · inbound

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team cites this paper.

Xolver: Multi-Agent Reasoning with Holistic Experience Learning Just Like an Olympiad Team Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T00:24:14.998105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:24:14.998105Z digest=sha256:f9802cb7650abef19f53112b432981888fd8cc199e2b050eeba3d234bf10bfcd

Observation 234bcb10-d3a6-47f2-bc4d-fe64c78c9cb2 · inbound

Measuring and Augmenting Large Language Models for Solving Capture-the-Flag Challenges cites this paper.

Measuring and Augmenting Large Language Models for Solving Capture-the-Flag Challenges Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:06.211001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:35:06.211001Z digest=sha256:8dd5e3e7e26f7f06863cb5c996dc4e2b80143e7b9f1455a199a32fcf7a4583b2

Observation 880d12be-9dc4-422d-9236-c69685b58a62 · inbound

CodeReasoner: Enhancing the Code Reasoning Ability with Reinforcement Learning cites this paper.

CodeReasoner: Enhancing the Code Reasoning Ability with Reinforcement Learning Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T14:50:28.075565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:50:28.075565Z digest=sha256:22846fbd8a1297e69a40b54c6ed4d1ba3d4ac25b98402746b8d7256b275335c9

Observation cb38483e-6e4a-4f33-898d-fb433cae4d56 · inbound

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security cites this paper.

Running in CIRCLE? A Simple Benchmark for LLM Code Interpreter Security Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T14:25:15.861206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:25:15.861206Z digest=sha256:f62759138b0d292c1a387400c0cd44f3ddc8d2e32acbc8bef4655049652c4b5a

Observation 5e17f278-5263-41ab-abdf-1c10041c4853 · inbound

HLSDebugger: Identification and Correction of Logic Bugs in HLS Code with LLM Solutions cites this paper.

HLSDebugger: Identification and Correction of Logic Bugs in HLS Code with LLM Solutions Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T12:46:40.769641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:46:40.769641Z digest=sha256:e7fce9ec5cf0ea527391f7f953cd1b052d0920613d68812755fa8591d7cafc84

Observation 138b9b54-5a36-400d-9835-0d97c8b54655 · inbound

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs cites this paper.

Let's Revise Step-by-Step: A Unified Local Search Framework for Code Generation with LLMs Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T22:09:46.518940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:09:46.518940Z digest=sha256:ff4d11f36a250e072b8081bf9c56fc4f459fe6494c154c3e5e4839671dd4a94a

Observation e1d02a06-62f1-441f-a554-2ccef80e06e9 · inbound

ReCode: Improving LLM-based Code Repair with Fine-Grained Retrieval-Augmented Generation cites this paper.

ReCode: Improving LLM-based Code Repair with Fine-Grained Retrieval-Augmented Generation Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T11:39:14.083993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T11:39:14.083993Z digest=sha256:e873fb354cd7b2e26250daca95841ec292433469ef0a2f3bc27d91be1885fa22

Observation e7b0c769-9aa3-437d-9e2f-29051d296fb2 · inbound

In Line with Context: Repository-Level Code Generation via Context Inlining cites this paper.

In Line with Context: Repository-Level Code Generation via Context Inlining Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-05-16T18:08:12.808019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T18:04:56.915339Z digest=sha256:6188f5436ffd72bf9e04d91cfd2e3c82d3f24176ab08189a47e7c286252aaa5a

Observation e6e73bf9-4857-48a0-bdd4-8c8853180c78 · inbound

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs? cites this paper.

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs? Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:20:17.842508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T20:17:52.610119Z digest=sha256:5aeb7a50fa8ae775471c88578c96adb0779376f7a45d86af89ef06f92f9f8a97

Observation 423b9b4e-3673-43b2-8a38-ecff1cf43351 · inbound

Detection Time Distribution Predicted Using Absorbing Boundary Conditions and Imaginary Potentials cites this paper.

Detection Time Distribution Predicted Using Absorbing Boundary Conditions and Imaginary Potentials Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-13T20:28:38.179448Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T20:28:38.179448Z digest=sha256:393972bfaf2157e1bbc88896f2f63851e8719774e850b4b09298a2f1fb8261b9

Observation 647f40af-d011-4e14-a6bc-f4ae8690cf05 · inbound

Dynamic analysis enhances issue resolution cites this paper.

Dynamic analysis enhances issue resolution Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:48:24.895620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T00:45:55.460545Z digest=sha256:b8016ef172f4f34414096eb09bde9ea1461cf0039a27488b19a89dd3daad202f

Observation 4e80d75e-b4bc-487a-a323-cd5751583e95 · inbound

AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search cites this paper.

AdverMCTS: Combating Pseudo-Correctness in Code Generation via Adversarial Monte Carlo Tree Search Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:40:57.597712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-10T16:35:16.056397Z digest=sha256:56d06c3c8e93dc623de006f68809bb14655629d35340e70ea8052d975f7efdda

Observation 3123ca8f-b908-45bc-b825-786b8f5e1c02 · inbound

How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks cites this paper.

How Many Tries Does It Take? Iterative Self-Repair in LLM Code Generation Across Model Scales and Benchmarks Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:06:00.395039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:14:52.614444Z digest=sha256:b68331ff09a62268c5abc1fdbb869f6773ab099ea533e05b24890810d88a84ef

Observation 151509f8-ab1c-4d19-952d-34cb0ad2a0c3 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:20:57.021636Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T01:47:39.926540Z digest=sha256:10cf289be8a85f9672a050f5fcf7a47327b9059b977ea91269356bfbd09f3e2d

Observation 9bbb7300-3ad0-4605-b27a-e1643094d53b · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:19:15.120133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T23:15:44.550045Z digest=sha256:4431b848de9b2c592c1423f1a1cc75cdbabebd8112c5735e50de88e358710c1f

Observation 08a2cae3-f3c1-48f9-a53e-f567bd35fa8e · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:25:07.438777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T23:23:42.883286Z digest=sha256:79baadcc5055e77a3e95285415f60e2bd0cb140f314b9a1828f623343338f4f2

Observation d59e8a9a-00b7-4d77-b013-f328f93e4991 · inbound

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning cites this paper.

StepCodeReasoner: Aligning Code Reasoning with Stepwise Execution Traces via Reinforcement Learning Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T05:32:20.209348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-13T05:27:37.521421Z digest=sha256:ffe5456f2fe73267977f2b1fb8db98fc3aa8c39b20332734b553200cc3cb4072

Observation ee9575bd-67af-4c33-97d7-552400cb5328 · inbound

Prompt Optimization for LLM Code Generation via Reinforcement Learning cites this paper.

Prompt Optimization for LLM Code Generation via Reinforcement Learning Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T08:53:10.295304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T08:49:36.986452Z digest=sha256:3424ea62840bfaa04aed46fd2b29ab1c25350f7d51185ce82ba588068757a319

Observation 9b9dc526-1674-4f4b-8bc2-571b80a7c23a · inbound

How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval cites this paper.

How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:26:12.311180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T21:14:01.070601Z digest=sha256:cd19f0dc978386430ae9aeeefb423d3d946a99acdcf83a18f53c919be6e08626

Observation 09042a0b-9425-4295-bc19-d8f1fb10914f · inbound

SrDetection: A Self-Referential Framework for Data Leakage Detection in Code Large Language Models cites this paper.

SrDetection: A Self-Referential Framework for Data Leakage Detection in Code Large Language Models Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:24:18.354146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-30T06:20:02.046020Z digest=sha256:cf092e096177f35e4054fc37ac6d8ec32326c6661601267360a074f948c3fde7