A benchmark of LLMs versus a proprietary hiring model on ~10,000 real candidate-job pairs reports the proprietary model wins on accuracy and fairness, while all tested LLMs show racial and intersectional bias.
Man is to computer programmer as woman is to homemaker? debiasing word embeddings
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
Evaluating the Promise and Pitfalls of LLMs in Hiring Decisions
A benchmark of LLMs versus a proprietary hiring model on ~10,000 real candidate-job pairs reports the proprietary model wins on accuracy and fairness, while all tested LLMs show racial and intersectional bias.