REVIEW 2 cited by
Efficient Classification of Student Help Requests in Programming Courses Using Large Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The accurate classification of student help requests with respect to the type of help being sought can enable the tailoring of effective responses. Automatically classifying such requests is non-trivial, but large language models (LLMs) appear to offer an accessible, cost-effective solution. This study evaluates the performance of the GPT-3.5 and GPT-4 models for classifying help requests from students in an introductory programming class. In zero-shot trials, GPT-3.5 and GPT-4 exhibited comparable performance on most categories, while GPT-4 outperformed GPT-3.5 in classifying sub-categories for requests related to debugging. Fine-tuning the GPT-3.5 model improved its performance to such an extent that it approximated the accuracy and consistency across categories observed between two human raters. Overall, this study demonstrates the feasibility of using LLMs to enhance educational systems through the automated classification of student needs.
Forward citations
Cited by 2 Pith papers
-
Beyond the Hype: A Comprehensive Review of Current Trends in Generative AI Research, Teaching Practices, and Tools
Computing educators are adopting GenAI faster than they are formalizing policies, and both educators and developers see code reading, evaluation, and problem decomposition as rising in importance over syntax recall.
-
Analysis of Student-LLM Interaction in a Software Engineering Project
Analysis of student-LLM conversations and code in a 13-week software engineering course finds ChatGPT preferred over Copilot and conversational prompting yields lower-complexity code.
Discussion (0). Continue with ORCID to comment.