A preregistered 2x3 factorial experiment on 262 real survey questions shows GPT-4.0 and a survey-expert persona make ChatGPT flag more, and different, survey question problems than GPT-3.5 or no persona.
Title resolution pending
1 Pith paper cite this work, alongside 10 external citations. Polarity classification is still indexing.
1
Pith paper citing it
10
external citations · OpenAlex
citation-role summary
background 1
citation-polarity summary
fields
stat.ME 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Generative AI as a Safety Net for Survey Question Refinement
A preregistered 2x3 factorial experiment on 262 real survey questions shows GPT-4.0 and a survey-expert persona make ChatGPT flag more, and different, survey question problems than GPT-3.5 or no persona.