Pith. sign in

REVIEW 1 cited by

On the Scaling Laws of Geographical Representation in Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.19406 v2 pith:LV3A7PY7 submitted 2024-02-29 cs.CL cs.AI

classification cs.CLcs.AI
keywords modelsgeographicallanguagebeenknowledgescalingbiascannot
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Language models have long been shown to embed geographical information in their hidden representations. This line of work has recently been revisited by extending this result to Large Language Models (LLMs). In this paper, we propose to fill the gap between well-established and recent literature by observing how geographical knowledge evolves when scaling language models. We show that geographical knowledge is observable even for tiny models, and that it scales consistently as we increase the model size. Notably, we observe that larger language models cannot mitigate the geographical bias that is inherent to the training data.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Fantastic Biases (What are They) and Where to Find Them

    cs.CL 2024-11 conditional novelty 3.0 of 10

    A survey that defines bias broadly, catalogs commonly discussed AI and NLP biases, and reviews methods to detect and mitigate them.

Pith tools