A generate-evaluate-regenerate loop with GPT-4o raised rubric scores for feedback on 208 quiz responses, but the second-round evaluation was done by the same model that rewrote the feedback, and the abstract misreports one non-significant result.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.HC 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
From First Draft to Final Insight: A Multi-Agent Approach for Feedback Generation
A generate-evaluate-regenerate loop with GPT-4o raised rubric scores for feedback on 208 quiz responses, but the second-round evaluation was done by the same model that rewrote the feedback, and the abstract misreports one non-significant result.