Across 204 languages, model scale has little effect on zero-shot multilingual performance, improves two-shot classification linearly, and helps translation mainly for instruction-tuned models.
This is reflected in how these models perform in the languages of the different resource levels
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
The Impact of Model Scaling on Seen and Unseen Language Performance
Across 204 languages, model scale has little effect on zero-shot multilingual performance, improves two-shot classification linearly, and helps translation mainly for instruction-tuned models.