Applying IIT 3.0/4.0 Φ estimates to LLM hidden-state sequences from Theory of Mind tests finds no robust statistical evidence of 'consciousness' phenomena, with span representations usually explaining score differences better than Φ.
1499–1509
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Can "consciousness" be observed from large language model (LLM) internal states? Dissecting LLM representations obtained from Theory of Mind test with Integrated Information Theory and Span Representation analysis
Applying IIT 3.0/4.0 Φ estimates to LLM hidden-state sequences from Theory of Mind tests finds no robust statistical evidence of 'consciousness' phenomena, with span representations usually explaining score differences better than Φ.