REVIEW 2 cited by
CodeHelp: Using Large Language Models with Guardrails for Scalable Support in Programming Classes
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Computing educators face significant challenges in providing timely support to students, especially in large class settings. Large language models (LLMs) have emerged recently and show great promise for providing on-demand help at a large scale, but there are concerns that students may over-rely on the outputs produced by these models. In this paper, we introduce CodeHelp, a novel LLM-powered tool designed with guardrails to provide on-demand assistance to programming students without directly revealing solutions. We detail the design of the tool, which incorporates a number of useful features for instructors, and elaborate on the pipeline of prompting strategies we use to ensure generated outputs are suitable for students. To evaluate CodeHelp, we deployed it in a first-year computer and data science course with 52 students and collected student interactions over a 12-week period. We examine students' usage patterns and perceptions of the tool, and we report reflections from the course instructor and a series of recommendations for classroom use. Our findings suggest that CodeHelp is well-received by students who especially value its availability and help with resolving errors, and that for instructors it is easy to deploy and complements, rather than replaces, the support that they provide to students.
Forward citations
Cited by 2 Pith papers
-
Seeing the Forest and the Trees: Solving Visual Graph and Tree Based Data Structure Problems using Large Multimodal Models
On a newly generated benchmark, multimodal models solve up to 87.6% of visual tree problems and 56.2% of visual graph problems, undercutting the idea that diagrams make exam questions AI-proof.
-
Oversight in Action: Experiences with Instructor-Moderated LLM Responses in an Online Discussion Forum
An instructor-moderated LLM bot for discussion forums reduced self-reported instructor workload in one course, but the evaluation lacks student feedback and a workload baseline.
Discussion (0). Continue with ORCID to comment.