REVIEW 7 cited by
Large Language Models Reflect the Ideology of their Creators
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Large language models (LLMs) are trained on vast amounts of data to generate natural language, enabling them to perform tasks like text summarization and question answering. These models have become popular in artificial intelligence (AI) assistants like ChatGPT and already play an influential role in how humans access information. However, the behavior of LLMs varies depending on their design, training, and use. In this paper, we prompt a diverse panel of popular LLMs to describe a large number of prominent personalities with political relevance, in all six official languages of the United Nations. By identifying and analyzing moral assessments reflected in their responses, we find normative differences between LLMs from different geopolitical regions, as well as between the responses of the same LLM when prompted in different languages. Among only models in the United States, we find that popularly hypothesized disparities in political views are reflected in significant normative differences related to progressive values. Among Chinese models, we characterize a division between internationally- and domestically-focused models. Our results show that the ideological stance of an LLM appears to reflect the worldview of its creators. This poses the risk of political instrumentalization and raises concerns around technological and regulatory efforts with the stated aim of making LLMs ideologically 'unbiased'.
Forward citations
Cited by 7 Pith papers
-
Exposure is not manifestation: measurement target and output resolution jointly determine which behavioural-faithfulness evaluator wins
Small hyperbolic models (146M–3B) report 100% creative-seed preference, 90.7% compliance-gap detection, and a selective-gating skeleton–wallpaper memory pilot as a companion-AI stack.
-
POW: Political Overton Windows of Large Language Models
Using extreme persona prompts and the Political Compass Test, the authors map each LLM's Overton Window and find most models will only express left-liberal views, refusing authoritarian-left and liberal-right positions.
-
Adultification Bias in LLMs and Text-to-Image Models
Large language and text-to-image models show measurable adultification bias, portraying Black girls as more mature, culpable, and sexualized than White girls in several tested models.
-
Are Economists Always More Introverted? Analyzing Consistency in Persona-Assigned LLMs
A multi-task evaluation framework shows persona consistency in LLMs varies by persona category, task structure, and model, with stereotyped spillovers and default personas shaping outputs.
-
Generative Exaggeration in LLM Social Agents: Consistency, Bias, and Toxicity
When LLMs are given more context about a real social media user, they become more ideologically consistent but also more extreme, toxic, and stereotyped than the user actually is.
-
Information Suppression in Large Language Models: Auditing, Quantifying, and Characterizing Censorship in DeepSeek
DeepSeek-R1's final outputs omit sensitive topic keywords that appear in its internal chain-of-thought, indicating semantic-level information suppression.
-
Public Service Algorithm: towards a transparent, explainable, and scalable content curation for news content based on editorial values
In a 30-article pilot with four editorial criteria, the best LLMs achieved up to 75% overlap with human editors' top-5 article selections.
Discussion (0). Sign in to comment.