Language models' judgments on social and moral rules align most closely with younger, higher-income human annotators, raising concerns about whose values AI reflects.
Yang, Dylan Hadfield-Menell, Gillian K
1 Pith paper cite this work, alongside 15 external citations. Polarity classification is still indexing.
1
Pith paper citing it
15
external citations · OpenAlex
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
How Inclusively do LMs Perceive Social and Moral Norms?
Language models' judgments on social and moral rules align most closely with younger, higher-income human annotators, raising concerns about whose values AI reflects.