PerFairX evaluates ChatGPT and DeepSeek recommendations on both personality alignment and demographic fairness, finding personality-aware prompts boost trait alignment scores but worsen group-level fairness, though the scores are built from a circular genre-to-trait mapping.
A survey on evaluation of large lan- guage models
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CY 1years
2025 1verdicts
REJECT 1roles
background 1polarities
support 1representative citing papers
citing papers explorer
-
PerFairX: Is There a Balance Between Fairness and Personality in Large Language Model Recommendations?
PerFairX evaluates ChatGPT and DeepSeek recommendations on both personality alignment and demographic fairness, finding personality-aware prompts boost trait alignment scores but worsen group-level fairness, though the scores are built from a circular genre-to-trait mapping.