A randomized audit of five LLM APIs finds verified survey-country metadata improves held-out response forecasts, while disclosing that a country label was randomly assigned does not reliably attenuate its influence.
InProceedings of the First Work- shop on Multilingual Multicultural Evaluation, 23–34
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.AI 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Signal or Spurious Cue? A Randomized Audit of Survey-Country Metadata in LLM Social Inference
A randomized audit of five LLM APIs finds verified survey-country metadata improves held-out response forecasts, while disclosing that a country label was randomly assigned does not reliably attenuate its influence.