REVIEW 2 cited by
On the Scaling Laws of Geographical Representation in Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Language models have long been shown to embed geographical information in their hidden representations. This line of work has recently been revisited by extending this result to Large Language Models (LLMs). In this paper, we propose to fill the gap between well-established and recent literature by observing how geographical knowledge evolves when scaling language models. We show that geographical knowledge is observable even for tiny models, and that it scales consistently as we increase the model size. Notably, we observe that larger language models cannot mitigate the geographical bias that is inherent to the training data.
Forward citations
Cited by 2 Pith papers
-
From Data to Knowledge: Evaluating How Efficiently Language Models Learn Facts
Language models trained on the same data differ most on rare facts; larger models and LLaMA architectures appear more sample-efficient at factual recall.
-
Fantastic Biases (What are They) and Where to Find Them
A survey that defines bias broadly, catalogs commonly discussed AI and NLP biases, and reviews methods to detect and mitigate them.
Discussion (0). Continue with ORCID to comment.