Pith. sign in

REVIEW 9 cited by

The Consciousness Prior

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1709.08568 v2 pith:2HPWVFZD submitted 2017-09-25 cs.LG cs.AIstat.ML

classification cs.LGcs.AIstat.ML
keywords consciouspriorconceptsformconsciousnessfactorhigh-levellanguage
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A new prior is proposed for learning representations of high-level concepts of the kind we manipulate with language. This prior can be combined with other priors in order to help disentangling abstract factors from each other. It is inspired by cognitive neuroscience theories of consciousness, seen as a bottleneck through which just a few elements, after having been selected by attention from a broader pool, are then broadcast and condition further processing, both in perception and decision-making. The set of recently selected elements one becomes aware of is seen as forming a low-dimensional conscious state. This conscious state is combining the few concepts constituting a conscious thought, i.e., what one is immediately conscious of at a particular moment. We claim that this architectural and information-processing constraint corresponds to assumptions about the joint distribution between high-level concepts. To the extent that these assumptions are generally true (and the form of natural language seems consistent with them), they can form a useful prior for representation learning. A low-dimensional thought or conscious state is analogous to a sentence: it involves only a few variables and yet can make a statement with very high probability of being true. This is consistent with a joint distribution (over high-level concepts) which has the form of a sparse factor graph, i.e., where the dependencies captured by each factor of the factor graph involve only very few variables while creating a strong dip in the overall energy function. The consciousness prior also makes it natural to map conscious states to natural language utterances or to express classical AI knowledge in a form similar to facts and rules, albeit capturing uncertainty as well as efficient search mechanisms implemented by attention mechanisms.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 9 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 108 citations worldwide. Full citation record

  1. Hidden APIs in Language Models: Discovering Reusable Causal Interfaces from Forked Futures

    cs.AI 2026-07 conditional novelty 7.0 of 10

    A Shared hidden-state interface beats Local, Mixture, and Distributed alternatives under held-out causal description length in Qwen2.5-1.5B and Llama-3-8B, with transplantation and mediation evidence of reuse.

  2. Verbalizable Representations Form a Global Workspace in Language Models

    cs.CL 2026-07 conditional novelty 7.0 of 10

    Language models represent their current reasoning in a small, readable set of verbalizable vectors (the J-space) that functions like a global workspace.

  3. EasyLens: A Training-Free Plug-and-Play Subtle-Lesion Representation Amplifier for Medical Vision-Language Models

    cs.CV 2026-06 unverdicted novelty 7.0 of 10

    Prior-minimal multi-agent RL agents develop indexical encoding, persistent self-state, and an echo-mismatch self-monitoring circuit that vanishes when the echo affordance is removed during training.

  4. Can "consciousness" be observed from large language model (LLM) internal states? Dissecting LLM representations obtained from Theory of Mind test with Integrated Information Theory and Span Representation analysis

    cs.CL 2025-06 conditional novelty 6.0 of 10

    Applying IIT 3.0/4.0 Φ estimates to LLM hidden-state sequences from Theory of Mind tests finds no robust statistical evidence of 'consciousness' phenomena, with span representations usually explaining score difference...

  5. Transformers Learn Faster with Semantic Focus

    cs.LG 2025-06 conditional novelty 6.0 of 10

    Input-dependent top-k sparse attention makes small transformers converge faster and generalize as well as full attention, while input-agnostic sparsity does not, and the effect is tied to reduced dispersion of attenti...

  6. Dynamical Mechanisms for Coordinating Long-term Working Memory Based on the Precision of Spike-timing in Cortical Neurons

    q-bio.NC 2025-12 unverdicted novelty 5.0 of 10

    Cortical traveling waves may use precise spike timing and temporary STDP to form a second network that sustains long-term working memory for hours.

  7. Semantic Communication meets System 2 ML: How Abstraction, Compositionality and Emergent Languages Shape Intelligence

    cs.LG 2025-05 conditional novelty 5.0 of 10

    The paper proposes a research vision in which 6G networks and AI agents communicate through semantic abstractions composed algebraically, guided by Kahneman's System 1 versus System 2 distinction.

  8. Thoughtseeds as Latent Causes: A Dual-Process Computational Phenomenology of Focused-Attention Meditation

    q-bio.NC 2026-07 reject novelty 4.0 of 10

    A three-layer active-inference simulation of focused-attention meditation reproduces expert–novice differences, yet the key differences are pre-specified in hand-set priors rather than derived.

  9. Bhatt Conjectures: On Necessary-But-Not-Sufficient Benchmark Tautology for Human Like Reasoning

    cs.CR 2025-06 reject novelty 2.0 of 10

    A position paper restates known AI evaluation criteria as 'tautological' definitions of reasoning and understanding, with proposed diagnostic tests but no empirical evidence.

Pith tools