Pith. sign in

REVIEW 14 cited by

The Consciousness Prior

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1709.08568 v2 pith:2HPWVFZD submitted 2017-09-25 cs.LG cs.AIstat.ML

The Consciousness Prior

classification cs.LG cs.AIstat.ML
keywords consciouspriorconceptsformconsciousnessfactorhigh-levellanguage
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

A new prior is proposed for learning representations of high-level concepts of the kind we manipulate with language. This prior can be combined with other priors in order to help disentangling abstract factors from each other. It is inspired by cognitive neuroscience theories of consciousness, seen as a bottleneck through which just a few elements, after having been selected by attention from a broader pool, are then broadcast and condition further processing, both in perception and decision-making. The set of recently selected elements one becomes aware of is seen as forming a low-dimensional conscious state. This conscious state is combining the few concepts constituting a conscious thought, i.e., what one is immediately conscious of at a particular moment. We claim that this architectural and information-processing constraint corresponds to assumptions about the joint distribution between high-level concepts. To the extent that these assumptions are generally true (and the form of natural language seems consistent with them), they can form a useful prior for representation learning. A low-dimensional thought or conscious state is analogous to a sentence: it involves only a few variables and yet can make a statement with very high probability of being true. This is consistent with a joint distribution (over high-level concepts) which has the form of a sparse factor graph, i.e., where the dependencies captured by each factor of the factor graph involve only very few variables while creating a strong dip in the overall energy function. The consciousness prior also makes it natural to map conscious states to natural language utterances or to express classical AI knowledge in a form similar to facts and rules, albeit capturing uncertainty as well as efficient search mechanisms implemented by attention mechanisms.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 14 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Hidden APIs in Language Models: Discovering Reusable Causal Interfaces from Forked Futures

    cs.AI 2026-07 conditional novelty 7.0

    A Shared hidden-state interface beats Local, Mixture, and Distributed alternatives under held-out causal description length in Qwen2.5-1.5B and Llama-3-8B, with transplantation and mediation evidence of reuse.

  2. Verbalizable Representations Form a Global Workspace in Language Models

    cs.CL 2026-07 conditional novelty 7.0

    Language models represent their current reasoning in a small, readable set of verbalizable vectors (the J-space) that functions like a global workspace.

  3. Bayesian updates from coalgebraic determinisation

    cs.LO 2026-06 unverdicted novelty 7.0

    Unifilarisation of stochastic Mealy machines is an instance of coalgebraic determinisation over monads with support structure, producing causal stochastic behaviours rather than Moore-style output distributions.

  4. EasyLens: A Training-Free Plug-and-Play Subtle-Lesion Representation Amplifier for Medical Vision-Language Models

    cs.CV 2026-06 conditional novelty 7.0

    Prior-minimal multi-agent RL agents develop indexical encoding, persistent self-state, and an echo-mismatch self-monitoring circuit that vanishes when the echo affordance is removed during training.

  5. ViperGPT: Visual Inference via Python Execution for Reasoning

    cs.CV 2023-03 unverdicted novelty 7.0

    ViperGPT generates executable Python code to compose pre-trained vision-and-language modules into programs that answer visual queries, reaching state-of-the-art results with no additional training.

  6. EasyLens: A Training-Free Plug-and-Play Subtle-Lesion Representation Amplifier for Medical Vision-Language Models

    cs.CV 2026-06 unverdicted novelty 6.0

    EasyLens introduces a plug-and-play amplifier that uses pathology-anatomy prototypes and morphology-guided residual enhancement to boost subtle-lesion cues in frozen medical VLMs.

  7. Generative Recursive Reasoning

    cs.AI 2026-05 unverdicted novelty 6.0

    GRAM is a latent-variable generative model that performs recursive reasoning via stochastic trajectories, trained with amortized variational inference to support multi-hypothesis reasoning and unconditional generation.

  8. Generative Recursive Reasoning

    cs.AI 2026-05 unverdicted novelty 6.0

    GRAM turns recursive latent reasoning into a generative probabilistic model via stochastic trajectories and amortized variational inference, claiming better performance on structured reasoning tasks than deterministic...

  9. When to Think Fast and Slow? AMOR: Adaptive Entropy Gate for Hybrid Models

    cs.AI 2026-01 unverdicted novelty 6.0

    AMOR uses output entropy to gate attention in recurrent hybrids, matching full attention performance at roughly 22% attention invocations across 180M-1.5B models.

  10. Dynamical Mechanisms for Coordinating Long-term Working Memory Based on the Precision of Spike-timing in Cortical Neurons

    q-bio.NC 2025-12 conditional novelty 5.0

    Cortical traveling waves may use precise spike timing and temporary STDP to form a second network that sustains long-term working memory for hours.

  11. Thoughtseeds as Latent Causes: A Dual-Process Computational Phenomenology of Focused-Attention Meditation

    q-bio.NC 2026-07 reject novelty 4.0

    A three-layer active-inference simulation of focused-attention meditation reproduces expert–novice differences, yet the key differences are pre-specified in hand-set priors rather than derived.

  12. Emergent Language as an Approach to Conscious AI

    cs.CL 2026-06 unverdicted novelty 4.0

    Agents in a minimal multi-agent RL setup develop self-referential communication and an echo-mismatch detection circuit that emerges from environmental affordances rather than task structure or architecture.

  13. Dynamical Mechanisms for Coordinating Long-term Working Memory Based on the Precision of Spike-timing in Cortical Neurons

    q-bio.NC 2025-12 unverdicted novelty 4.0

    Coordinated millisecond-precision spike timing and traveling waves may enable a second tier of cortical activity for long-term working memory via sustained STDP.

  14. AI Consciousness and Existential Risk

    cs.AI 2025-11 unverdicted novelty 2.0

    Consciousness does not directly predict AI existential risk unlike intelligence, though it may indirectly affect risk through alignment or capability requirements.