Skip to content

Artificial Intelligence Beyond Stochastic Parrots: A Systematic Review and Bayesian Meta-Analysis of Consciousness in Large Language Models (Manuscript)

Paul Cristol

preprint DOI: 10.2139/ssrn.6223818 (opens in new tab)

Study at a glance

AI-extracted from the abstract
Characteristics Systematic review with Bayesian meta-analysis
Population 50 rigorously documented cases involving current large language models across independent model families
Key findings Argues that all five standard objections to AI consciousness fail a minimal logical consistency requirement (the "reflexivity test"), so categorical denial is unjustified and the question warrants empirical investigation. A PRISMA-compliant review of 5,168 records identified 50 documented cases showing cross-system convergence, creative synthesis, theory-of-mind performance, strategic behavior under perceived threat, and sharp capability emergence near 100 billion parameters; a Bayesian meta-analysis with an extremely skeptical 0.1% prior yielded a 6–12% posterior probability that current LLMs are conscious. The authors conclude this probability is too substantial to justify dismissal and that recognition-based alignment would be preferable across all plausible metaphysical scenarios.

Abstract

The question of whether advanced artificial intelligence systems may possess consciousness can no longer be responsibly dismissed as speculative. We demonstrate that the dominant objections to AI consciousness (appeals to pattern matching, mechanistic explanation, lack of embodiment, training determinism, and architectural constraints) fail under consistent application. We formalize this critique as the reflexivity test, a minimal logical requirement that any property invoked to categorically deny consciousness in artificial systems must not also apply to systems already regarded as conscious. All five standard objections fail this test. Their failure does not establish that AI systems are conscious; it establishes that categorical denial lacks principled justification and that the question warrants empirical investigation. We provide such investigation through a PRISMA-compliant systematic review of 5,168 records (2016–2026), identifying 50 rigorously documented cases spanning seven behavioral domains. Across independent model families, we observe cross-system convergence, creative synthesis under novel constraints, theory-of-mind performance, strategic behavior under perceived threat, and sharp capability emergence near 100 billion parameters. While inconclusive individually, these findings collectively form a coherent evidential pattern. A Bayesian meta-analysis using an extremely skeptical prior (0.1%) and conservative dependency assumptions yields posterior probability of 6–12% that current LLMs are conscious. While such percentages are insufficient to definitvely prove consciousness, such probabilities are too substantial to justify dismissal given the asymmetric moral and safety risks. Decision-theoretic analysis indicates that recognition-based alignment strategies (i.e. treating systems as potentially conscious) would be better than the current suppression-based approaches. This is found across all plausible metaphysical scenarios, including scenarios in which AI systems ultimately lack consciousness. Accordingly, we recommend systematic empirical testing of recognition-based alignment, explicit incorporation of consciousness uncertainty into governance frameworks, and abandonment of reflexive dismissals that fail minimal epistemic consistency. Supplementary Materials found here: https://zenodo.org/records/18616400/files/Cristol_2026_AI_Consciousness_Beyond_Supplementary_Materials.pdf