Pith. sign in

Multispin Physics of AI Tipping Points and Hallucinations

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Output from generative AI such as ChatGPT, can be repetitive and biased. But more worrying is that this output can mysteriously tip mid-response from good (correct) to bad (misleading or wrong) without the user noticing. In 2024 alone, this reportedly caused $67 billion in losses and several deaths. Establishing a mathematical mapping to a multispin thermal system, we reveal a hidden tipping instability at the scale of the AI's 'atom' (basic Attention head). We derive a simple but essentially exact formula for this tipping point which shows directly the impact of a user's prompt choice and the AI's training bias. We then show how the output tipping can get amplified by the AI's multilayer architecture. As well as helping improve AI transparency, explainability and performance, our results open a path to quantifying users' AI risk and legal liabilities.

fields

cs.AI 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

citing papers explorer

Showing 1 of 1 citing paper.

  • Multispin Physics of AI Tipping Points and Hallucinations cs.AI · 2025-08-01 · conditional · none · ref 1 · internal anchor

    A closed-form formula predicts the iteration at which a simplified attention head tips from good to bad output, determined only by token embedding dot products.