Pith. sign in

REVIEW 1 cited by

When Dialects Collide: How Socioeconomic Mixing Affects Language Use

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2307.10016 v2 pith:7RHRKEFV submitted 2023-07-19 physics.soc-ph cs.CLcs.SI

classification physics.soc-phcs.CLcs.SI
keywords socioeconomicstandardareasclassesdatadifferentincomelanguage
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The socioeconomic background of people and how they use standard forms of language are not independent, as demonstrated in various sociolinguistic studies. However, the extent to which these correlations may be influenced by the mixing of people from different socioeconomic classes remains relatively unexplored from a quantitative perspective. In this work we leverage geotagged tweets and transferable computational methods to map deviations from standard English on a large scale, in seven thousand administrative areas of England and Wales. We combine these data with high-resolution income maps to assign a proxy socioeconomic indicator to home-located users. Strikingly, across eight metropolitan areas we find a consistent pattern suggesting that the more different socioeconomic classes mix, the less interdependent the frequency of their departures from standard grammar and their income become. Further, we propose an agent-based model of linguistic variety adoption that sheds light on the mechanisms that produce the observations seen in the data.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Entropy and type-token ratio in gigaword corpora

    cs.CL 2024-11 conditional novelty 5.0 of 10

    Word entropy and type-token ratio in billion-token corpora are linked by an asymptotic formula built from Zipf and Heaps laws, confirmed across English, Spanish and Turkish texts.

Pith tools