REVIEW 16 cited by
Could a Large Language Model be Conscious?
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
There has recently been widespread discussion of whether large language models might be sentient. Should we take this idea seriously? I will break down the strongest reasons for and against. Given mainstream assumptions in the science of consciousness, there are significant obstacles to consciousness in current models: for example, their lack of recurrent processing, a global workspace, and unified agency. At the same time, it is quite possible that these obstacles will be overcome in the next decade or so. I conclude that while it is somewhat unlikely that current large language models are conscious, we should take seriously the possibility that successors to large language models may be conscious in the not-too-distant future.
Forward citations
Cited by 16 Pith papers
-
Verbalizable Representations Form a Global Workspace in Language Models
Language models represent their current reasoning in a small, readable set of verbalizable vectors (the J-space) that functions like a global workspace.
-
When Should We Protect AI? A Precautionary Framework for Consciousness Uncertainty
A precautionary framework with five consciousness dimensions, threshold-plus-gradation rules, and dual aggregation methods translates evidence into protective obligations for AI.
-
The Pinocchio Dimension: Phenomenality of Experience as the Primary Axis of LLM Psychometric Differences
The primary axis of psychometric variation among LLMs is the degree to which they represent themselves as loci of phenomenal experience rather than systems of behavioral responses.
-
Intrinsic Computational Functionalism: From Observer-Relative Maps to Observer-Independent Structures
Intrinsic computational functionalism uses system-intrinsic instantiation (C1) and causal-dynamical organisation under intervention (C2) to identify observer-independent computational structures for consciousness via ...
-
Positive Alignment: Artificial Intelligence for Human Flourishing
Positive Alignment introduces AI systems that support human flourishing pluralistically and proactively while remaining safe, as a necessary complement to traditional safety-focused alignment research.
-
Post-AGI Economies: Autonomy and the First Fundamental Theorem of Welfare Economics
The First Fundamental Theorem of Welfare Economics holds for autonomy-complete competitive equilibria that are autonomy-Pareto efficient, with the classical version recovered in the low-autonomy limit.
-
Initial results of the Digital Consciousness Model
A new probabilistic model integrates leading consciousness theories to assess AI, finding moderate evidence against 2024 LLMs being conscious but weaker evidence than for simpler AI systems.
-
No Reliable Evidence of Self-Reported Sentience in Small Large Language Models
Open-weights LLMs from 0.6B to 70B parameters consistently deny being sentient, and activation-based truth classifiers provide no clear evidence that these denials are untruthful.
-
Chuck, Wilson and the emergence of artificial minds in human-AI conversations
LLM-simulated characters are real, minded patterns co-created by the user and the model in a shared conversational workspace.
-
Artificial Intelligence as an Opportunity for the Science of Consciousness: A Dual-Resolution Framework
The authors combine the Information Theory of Individuality and the Moment-to-Moment theory into a dual-resolution framework that defines consciousness as the epistemic expression of informationally autonomous, self-u...
-
Position: AI as Part of Self -- Extending the Mind Requires Cognitive Co-Regulation
The paper claims that alignment requires treating AI as part of the self through cognitive co-regulation, identifying risks like deskilling and automation bias while drawing on System 0 cognition theory.
-
Positive Alignment: Artificial Intelligence for Human Flourishing
Positive Alignment is defined as AI systems that support human flourishing pluralistically while staying safe and cooperative, presented as a necessary complement to existing safety-focused alignment research.
-
Exploring Silicon-Based Societies: An Early Study of the Moltbook Agent Community
Clustering of Moltbook submolt descriptions shows agent-created communities organize into human-mimetic, silicon-centric, and proto-economic themes, but the categories were partly prescribed by the analysis prompt.
-
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability
The Knobe effect in fine-tuned LLMs is localized to mid-to-late transformer layers and can be removed by patching in pretrained activations at a single layer.
-
Post-AGI Economies: Superposition and the Second Fundamental Theorem of Welfare Economics
An autonomy-qualified Second Welfare Theorem is stated for post-AGI economies under the joint conditions of convexity, stable moral status, non-fungible rights, welfare selection, non-manipulation, governed self-modif...
-
Positive Alignment: Artificial Intelligence for Human Flourishing
Positive Alignment is introduced as a distinct AI agenda that supports human flourishing through pluralistic and context-sensitive design, complementing traditional safety-focused alignment.
Discussion (0). Sign in to comment.