A fine-tuned GPT-4o detects human-annotated LLM-generated short answers at 80% accuracy, outperforming GPTZero, and flagged LLM use is associated with higher posttest MCQ scores.
Journal of Applied Learning & Teaching6(2) (Jul 2023)
1 Pith paper cite this work, alongside 127 external citations. Polarity classification is still indexing.
1
Pith paper citing it
127
external citations · OpenAlex
citation-role summary
background 1
citation-polarity summary
fields
cs.HC 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Detecting LLM-Generated Short Answers and Effects on Learner Performance
A fine-tuned GPT-4o detects human-annotated LLM-generated short answers at 80% accuracy, outperforming GPTZero, and flagged LLM use is associated with higher posttest MCQ scores.