REVIEW 4 cited by
Generative Monoculture in Large Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We introduce {\em generative monoculture}, a behavior observed in large language models (LLMs) characterized by a significant narrowing of model output diversity relative to available training data for a given task: for example, generating only positive book reviews for books with a mixed reception. While in some cases, generative monoculture enhances performance (e.g., LLMs more often produce efficient code), the dangers are exacerbated in others (e.g., LLMs refuse to share diverse opinions). As LLMs are increasingly used in high-impact settings such as education and web search, careful maintenance of LLM output diversity is essential to ensure a variety of facts and perspectives are preserved over time. We experimentally demonstrate the prevalence of generative monoculture through analysis of book review and code generation tasks, and find that simple countermeasures such as altering sampling or prompting strategies are insufficient to mitigate the behavior. Moreover, our results suggest that the root causes of generative monoculture are likely embedded within the LLM's alignment processes, suggesting a need for developing fine-tuning paradigms that preserve or promote diversity.
Forward citations
Cited by 4 Pith papers
-
Token-Level Entropy Reveals Demographic Disparities in Large Language Models
The abstract's claim of demographic entropy disparities is directly contradicted by the paper's own results section, which finds token sampling uncertainty does not explain homogeneity bias.
-
Correlated Errors in Large Language Models
Large language models from different providers and architectures often make the same errors, and more accurate models are especially likely to share mistakes.
-
We're Different, We're the Same: Creative Homogeneity Across LLMs
Across three divergent-thinking tests, responses from seven LLM families were substantially more similar to one another than responses from 102 humans were to one another.
-
An approach to systemic risks of AI through the lens of emergence, collective action problems, and externalities
Systemic AI risks are presented as emergent threats to public goods, driven chiefly by collective action problems and complex externalities, amplified by concentration, feedback, and information gaps.
Discussion (0). Continue with ORCID to comment.