REVIEW 3 cited by
Leveraging Print Debugging to Improve Code Generation in Large Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Large language models (LLMs) have made significant progress in code generation tasks, but their performance in tackling programming problems with complex data structures and algorithms remains suboptimal. To address this issue, we propose an in-context learning approach that guides LLMs to debug by using a "print debugging" method, which involves inserting print statements to trace and analysing logs for fixing the bug. We collect a Leetcode problem dataset and evaluate our method using the Leetcode online judging system. Experiments with GPT-4 demonstrate the effectiveness of our approach, outperforming rubber duck debugging in easy and medium-level Leetcode problems by 1.5% and 17.9%.
Forward citations
Cited by 3 Pith papers
-
Examining $H$-Closed Ducci Sequences on $\mathbb{Z}_m^n$
The authors prove that for several families of moduli, cyclically shifting the starting tuple does not change its Ducci cycle, and they tabulate many more such cases.
-
InfiGUIAgent: A Multimodal Generalist GUI Agent with Native Reasoning and Reflection
A 2B multimodal agent trained with two-stage supervised fine-tuning and synthesized hierarchical/reflection reasoning achieves competitive results on ScreenSpot and AndroidWorld.
-
The Current Challenges of Software Engineering in the Era of Large Language Models
The paper reports 26 challenges in LLM-based software engineering, grouped into seven aspects, derived from a structured discussion among 24 academics and practitioners.
Discussion (0). Continue with ORCID to comment.