Synthetic preference data generated from the Integrated Values Survey preserves cross-country value differences and enables DPO fine-tuning that improves LLM cultural alignment across five countries.
Title resolution pending
3 Pith papers cite this work. Polarity classification is still indexing.
years
2026 3representative citing papers
Frontier LLMs consistently output Western-style individualist advice on personal dilemmas even when prompted with non-Western cultural contexts, exceeding survey-measured local values by an average of 0.76 points on a 1-5 scale.
This survey examines applications of social choice theory to aggregating human feedback in AI alignment, identifying failure modes and expanding design options for disagreement.
citing papers explorer
-
PLURAL: A Global Dataset for Value Alignment
Synthetic preference data generated from the Integrated Values Survey preserves cross-country value differences and enables DPO fine-tuning that improves LLM cultural alignment across five countries.
-
When AI Speaks, Whose Values Does It Express? A Cross-Cultural Audit of Individualism-Collectivism Bias in Large Language Models
Frontier LLMs consistently output Western-style individualist advice on personal dilemmas even when prompted with non-Western cultural contexts, exceeding survey-measured local values by an average of 0.76 points on a 1-5 scale.
-
AI Alignment From Social Choice Perspectives
This survey examines applications of social choice theory to aggregating human feedback in AI alignment, identifying failure modes and expanding design options for disagreement.