Pith. sign in

REVIEW 2 cited by

Large Language Models (GPT) for automating feedback on programming assignments

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2307.00150 v1 pith:GV6V4SN3 submitted 2023-06-30 cs.HC cs.AI

classification cs.HCcs.AI
keywords feedbackhintsexperimentalstudentsassignmentswerelessprogramming
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Addressing the challenge of generating personalized feedback for programming assignments is demanding due to several factors, like the complexity of code syntax or different ways to correctly solve a task. In this experimental study, we automated the process of feedback generation by employing OpenAI's GPT-3.5 model to generate personalized hints for students solving programming assignments on an automated assessment platform. Students rated the usefulness of GPT-generated hints positively. The experimental group (with GPT hints enabled) relied less on the platform's regular feedback but performed better in terms of percentage of successful submissions across consecutive attempts for tasks, where GPT hints were enabled. For tasks where the GPT feedback was made unavailable, the experimental group needed significantly less time to solve assignments. Furthermore, when GPT hints were unavailable, students in the experimental condition were initially less likely to solve the assignment correctly. This suggests potential over-reliance on GPT-generated feedback. However, students in the experimental condition were able to correct reasonably rapidly, reaching the same percentage correct after seven submission attempts. The availability of GPT hints did not significantly impact students' affective state.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Reflection-Satisfaction Tradeoff: Investigating Impact of Reflection on Student Engagement with AI-Generated Programming Hints

    cs.CY 2025-12 conditional novelty 6.0 of 10

    Reflection prompts that produce deeper student reflection are associated with lower satisfaction with AI-generated hints, with no measurable gain in immediate problem-solving performance.

  2. Narrowing the Gap: Supervised Fine-Tuning of Open-Source LLMs as a Viable Alternative to Proprietary Models for Pedagogical Tools

    cs.CY 2025-07 conditional novelty 5.0 of 10

    Fine-tuned open-source models, especially Qwen3-4B, explain C compiler errors at a quality close to GPT-4.1 on expert-judged metrics.

Pith tools