GPT-4o essay scoring shows larger errors for non-native English writers when the model correctly infers that they are non-native, but no meaningful gender effect.
Creative Edu cation 15(7), 1499–1523 (2024)
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Does the Prompt-based Large Language Model Recognize Students' Demographics and Introduce Bias in Essay Scoring?
GPT-4o essay scoring shows larger errors for non-native English writers when the model correctly infers that they are non-native, but no meaningful gender effect.