LLMs score resumes higher than human experts and differ in how they adapt to company context, but the study's statistical claims are undermined by non-independent observations, a tiny human sample, and no released data.
All statistical analyses were conducted using appropriate software with significance levels and confidence intervals reported for all hypothesis tests
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Signal or Noise? Evaluating Large Language Models in Resume Screening Across Contextual Variations and Human Expert Benchmarks
LLMs score resumes higher than human experts and differ in how they adapt to company context, but the study's statistical claims are undermined by non-independent observations, a tiny human sample, and no released data.