LLM-generated reviews on real pre-revision submissions are longer, more positive, and less score-calibrated than human reviews, and aggregate quality scores alone overestimate their quality.
Title resolution pending
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CY 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
AI-Assisted Peer Review Across Research Communities: From Reviewer AI Policies to LLM Review Quality
LLM-generated reviews on real pre-revision submissions are longer, more positive, and less score-calibrated than human reviews, and aggregate quality scores alone overestimate their quality.