Pith. sign in

REVIEW 4 cited by

Generative Monoculture in Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.02209 v1 pith:FQ27K776 submitted 2024-07-02 cs.CL cs.AI

classification cs.CLcs.AI
keywords generativemonoculturellmsdiversitybehaviorbookcodelanguage
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We introduce {\em generative monoculture}, a behavior observed in large language models (LLMs) characterized by a significant narrowing of model output diversity relative to available training data for a given task: for example, generating only positive book reviews for books with a mixed reception. While in some cases, generative monoculture enhances performance (e.g., LLMs more often produce efficient code), the dangers are exacerbated in others (e.g., LLMs refuse to share diverse opinions). As LLMs are increasingly used in high-impact settings such as education and web search, careful maintenance of LLM output diversity is essential to ensure a variety of facts and perspectives are preserved over time. We experimentally demonstrate the prevalence of generative monoculture through analysis of book review and code generation tasks, and find that simple countermeasures such as altering sampling or prompting strategies are insufficient to mitigate the behavior. Moreover, our results suggest that the root causes of generative monoculture are likely embedded within the LLM's alignment processes, suggesting a need for developing fine-tuning paradigms that preserve or promote diversity.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Token-Level Entropy Reveals Demographic Disparities in Large Language Models

    cs.CL 2025-01 unverdicted novelty 7.0 of 10

    The abstract's claim of demographic entropy disparities is directly contradicted by the paper's own results section, which finds token sampling uncertainty does not explain homogeneity bias.

  2. Correlated Errors in Large Language Models

    cs.CL 2025-06 conditional novelty 6.0 of 10

    Large language models from different providers and architectures often make the same errors, and more accurate models are especially likely to share mistakes.

  3. We're Different, We're the Same: Creative Homogeneity Across LLMs

    cs.CY 2025-01 conditional novelty 6.0 of 10

    Across three divergent-thinking tests, responses from seven LLM families were substantially more similar to one another than responses from 102 humans were to one another.

  4. An approach to systemic risks of AI through the lens of emergence, collective action problems, and externalities

    cs.CY 2026-07 conditional novelty 5.0 of 10

    Systemic AI risks are presented as emergent threats to public goods, driven chiefly by collective action problems and complex externalities, amplified by concentration, feedback, and information gaps.

Pith tools