Big Five inventories fail to capture meaningful differences or recover the five-factor structure in LLMs, with only 3% variance between models and four facets collapsing (r >= .92).
Published as , year=
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
years
2026 2representative citing papers
Stable personality vectors in LLMs function as intrinsic guardrails, with ablation increasing emergent misalignment above 40% and amplification reducing it below 3%, enabling zero-shot transfer from aligned to corrupted models.
citing papers explorer
-
Personality Without Persons? A Psychometric Critique of Big Five Testing in Large Language Models
Big Five inventories fail to capture meaningful differences or recover the five-factor structure in LLMs, with only 3% variance between models and four facets collapsing (r >= .92).
-
Intrinsic Guardrails: How Semantic Geometry of Personality Interacts with Emergent Misalignment in LLMs
Stable personality vectors in LLMs function as intrinsic guardrails, with ablation increasing emergent misalignment above 40% and amplification reducing it below 3%, enabling zero-shot transfer from aligned to corrupted models.