REVIEW 6 cited by
Language Models as Models of Language
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
This chapter critically examines the potential contributions of modern language models to theoretical linguistics. Despite their focus on engineering goals, these models' ability to acquire sophisticated linguistic knowledge from mere exposure to data warrants a careful reassessment of their relevance to linguistic theory. I review a growing body of empirical evidence suggesting that language models can learn hierarchical syntactic structure and exhibit sensitivity to various linguistic phenomena, even when trained on developmentally plausible amounts of data. While the competence/performance distinction has been invoked to dismiss the relevance of such models to linguistic theory, I argue that this assessment may be premature. By carefully controlling learning conditions and making use of causal intervention methods, experiments with language models can potentially constrain hypotheses about language acquisition and competence. I conclude that closer collaboration between theoretical linguists and computational researchers could yield valuable insights, particularly in advancing debates about linguistic nativism.
Forward citations
Cited by 6 Pith papers
-
Implicit Causality-biases in humans and LLMs as a tool for benchmarking LLM discourse capabilities
Most tested LLMs fail to reproduce human implicit causality biases in coreference, coherence, and referring-expression form, even when they show partial coreference effects.
-
Randomly Sampled Language Reasoning Problems Elucidate Limitations of In-Context Learning
On randomly sampled 3-state DFA language tasks, foundation LLMs underperform n-gram baselines under pure in-context-learning prompts.
-
Do Large Language Models Advocate for Inferentialism?
Large language models process language in ways that align with inferentialist, anti-representationalist semantics, and they may be understood through a consensus theory of truth grounded in human feedback.
-
Linear representations of grammaticality in neural language models
Grammaticality is linearly decodable from language model sentence representations and generalizes across phenomena and languages in larger models.
-
Beyond Human-Like Processing: Large Language Models Perform Equivalently on Forward and Backward Scientific Text
GPT-2 models trained on character-reversed neuroscience text perform as well on a neuroscience abstract-selection benchmark as models trained on normal text, despite higher perplexity.
-
The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models
A systematic review of 337 articles shows Transformers handle formal syntax well but perform worse and more variably at the syntax-semantics interface, with the field over-reliant on English and BERT.
Discussion (0). Continue with ORCID to comment.