Pith. sign in

REVIEW 3 cited by

Generative Monoculture in Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.02209 v1 pith:FQ27K776 submitted 2024-07-02 cs.CL cs.AI

Generative Monoculture in Large Language Models

classification cs.CL cs.AI
keywords generativemonoculturellmsdiversitybehaviorbookcodelanguage
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

We introduce {\em generative monoculture}, a behavior observed in large language models (LLMs) characterized by a significant narrowing of model output diversity relative to available training data for a given task: for example, generating only positive book reviews for books with a mixed reception. While in some cases, generative monoculture enhances performance (e.g., LLMs more often produce efficient code), the dangers are exacerbated in others (e.g., LLMs refuse to share diverse opinions). As LLMs are increasingly used in high-impact settings such as education and web search, careful maintenance of LLM output diversity is essential to ensure a variety of facts and perspectives are preserved over time. We experimentally demonstrate the prevalence of generative monoculture through analysis of book review and code generation tasks, and find that simple countermeasures such as altering sampling or prompting strategies are insufficient to mitigate the behavior. Moreover, our results suggest that the root causes of generative monoculture are likely embedded within the LLM's alignment processes, suggesting a need for developing fine-tuning paradigms that preserve or promote diversity.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. LLM Jaggedness Unlocks Scientific Creativity

    cs.AI 2026-05 unverdicted novelty 6.0

    Jagged capabilities in LLMs for scientific idea generation can be leveraged through inference-time ensembles to outperform individual models.

  2. An approach to systemic risks of AI through the lens of emergence, collective action problems, and externalities

    cs.CY 2026-07 conditional novelty 5.0

    Systemic AI risks are presented as emergent threats to public goods, driven chiefly by collective action problems and complex externalities, amplified by concentration, feedback, and information gaps.

  3. LLM Jaggedness Unlocks Scientific Creativity

    cs.AI 2026-05 unverdicted novelty 5.0

    LLMs exhibit jagged scientific creativity across models, prompts, and domains, and this unevenness can be leveraged via model ensembles to outperform any single model on idea generation.