Failure of contextual invariance in large language models

Andrea Baronchelli; Ariel Flint; Luca Maria Aiello; Sagar Kumar

arxiv: 2603.23485 · v2 · pith:PHQYLR7Fnew · submitted 2026-03-24 · 💻 cs.CL · cs.AI· cs.CY

Failure of contextual invariance in large language models

Sagar Kumar , Ariel Flint , Luca Maria Aiello , Andrea Baronchelli This is my paper

classification 💻 cs.CL cs.AIcs.CY

keywords outputscontextgenderlargemodelpronouncontextualinvariance

0 comments

read the original abstract

Standard evaluation practices assume that large language model (LLM) outputs are stable when prompts are embedded in contextually equivalent discourses. Here, we test this assumption in the setting of gender inference. Using a controlled pronoun selection task, we introduce minimal, theoretically uninformative discourse context and find that this induces large, systematic shifts in model outputs. Correlations with cultural gender stereotypes, present in decontextualized settings, weaken or disappear once context is introduced, while theoretically irrelevant features, such as the gender of a pronoun for an unrelated referent, become the most informative predictors of model behavior. A Contextuality-by-Default analysis reveals that, in 19--52\% of cases across models, this dependence persists after accounting for all marginal effects of context on individual outputs and cannot be attributed to simple pronoun repetition. These findings show that LLM outputs violate contextual invariance even under near-identical syntactic formulations, with implications for bias benchmarking and deployment in high-stakes settings.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

Machine individuality: Separating genuine idiosyncrasy from response bias in large language models
cs.AI 2026-04 unverdicted novelty 7.0

Crossed random-effects models on LLM word ratings show 16.9% variance from genuine stimulus-specific individuality, exceeding null models and forming coherent per-model fingerprints.