Mindscape Collective is now The Consciousness Library. Same library, new name. You may need to sign in again. About the change
Skip to content

arXiv Preprint Archive

457 papers in the library · publishing 1995-2026

Papers

Can We Test Consciousness Theories on AI? Ablations, Markers, and Robustness

arXiv Preprint Archive December 22, 2025 Yin Jun Phua

Three major theories of consciousness—Global Workspace Theory, Integrated Information Theory, and Higher-Order Theories—may describe complementary functional layers rather than competing accounts. Artificial agents embodying mechanisms from each theory were tested through architectural ablations impossible in biological systems. A Self-Model lesion abolished metacognitive calibration while preserving first-order task performance, creating a synthetic blindsight analogue consistent with Higher-Order Theories. Workspace capacity proved causally necessary for information access, with complete lesions producing qualitative collapse in access-related markers, consistent with Global Workspace Theory. A broadcast-amplification effect showed that broadcasting amplifies internal noise, creating extreme fragility. Raw perturbational complexity decreased under the workspace bottleneck, cautioning against naive transfer of Integrated Information Theory-adjacent proxies to engineered agents. The results suggest a hierarchical design principle: Global Workspace Theory provides broadcast capacity, while Higher-Order Theories provide quality control.

A Disproof of Large Language Model Consciousness: The Necessity of Continual Learning for Consciousness

arXiv Preprint Archive December 14, 2025 Erik Hoel

A formal analysis shows that many contemporary theories of consciousness fail to meet the requirements of falsifiability and non-triviality. This is especially problematic for claims that Large Language Models (LLMs) might be conscious: because functionally equivalent systems exist, no falsifiable and non-trivial theory can judge them conscious, forming a disproof of LLM consciousness. In contrast, theories of consciousness that require continual learning do satisfy these formal constraints for humans. This supports the hypothesis that if continual learning is linked to human consciousness, the current lack of continual learning in LLMs is intimately tied to their lack of consciousness.

The Modeler Schema Theory of Consciousness, with a Falsifiable Experiment

arXiv Preprint Archive November 30, 2025 Frank Heile

Consciousness arises from a single control agent, the Modeler-schema, which monitors the brain's Modeler as it constructs and updates the internal World Model. During this monitoring, the Modeler-schema generates experience by converting the Modeler's outputs into qualia, which are then used for model refinement. The Human agent consists of three cooperating agents—Modeler, Controller, and Targeter—each paired with a regulatory schema agent. The core prediction is that the Modeler-schema performs a qualia-based consistency check during saccades and may issue bottom-up attention requests when discrepancies are found. A saccadic change-detection experiment is proposed to test this prediction. Locating qualia in the Modeler-schema ties experience to model regulation, explains aphantasia as a selective failure of recalled-sensory quale-conversion, and offers a testable proposal toward solving the Hard Problem of consciousness.

Testing the Machine Consciousness Hypothesis

arXiv Preprint Archive November 30, 2025 Stephen Fitz

Consciousness may be a functional property of computational systems that have second-order perception, emerging not from individual modeling but from communication between distributed predictive models. The theory proposes that collective self-models arise when local observers exchange predictive messages about patterns in a cellular automaton substrate, aligning partial views into a shared representation. This suggests consciousness is a property of the language a system evolves to describe itself internally, not an epiphenomenon of individual modeling. The research program aims to develop empirically testable theories by studying how self-models form in distributed systems without centralized control.

The Hydraulic Brain: Understanding as Constraint-Release Phase Transition in Whole-Body Resonance

arXiv Preprint Archive November 22, 2025 Ahmed Gamal Eldin

This paper proposes that brain-body physiological signals, typically treated as noise, are functionally important for cognition. Using high-density EEG during a P300 target recognition task, the authors found zero-lag synchrony and strong bidirectional coupling between brain and body signals, peaking at 78.1 ms. Time-resolved entropy analysis revealed three phases: constraint accumulation, a supercritical transition with a 58% directional increase in state expansion, and sustained metastability. The transition magnitude was uncorrelated with resonance strength, suggesting binary threshold dynamics. The authors argue that brain-body resonance acts as a discrete gate triggering non-linear information integration, potentially distinguishing biological from artificial intelligence.

From Representation to Enactment: The ABC Framework of the Translating Mind

arXiv Preprint Archive November 20, 2025 Michael Carl, Takanori Mizowaki, Aishvarya Raj et al.

Translation is not a matter of manipulating static correspondences between languages but an enacted activity that dynamically integrates affective, behavioral, and cognitive processes. Drawing on Extended Mind theory, radical enactivism, Predictive Processing, and Active Inference, the authors argue that the translator's mind emerges through loops of brain-body-environment interactions rather than being merely extended. This non-representational account reframes translation as skillful participation in sociocultural practice, where meaning is co-created in real time through embodied interaction with texts, tools, and contexts.

Experiencing the More-than-Human Through Human Augmentation

arXiv Preprint Archive November 16, 2025 Botao 'Amber' Hu, Danlin Huang

This paper proposes a design approach called MtHtHA (or ">HtH+") that repurposes human augmentation technologies to create temporary, embodied first-person experiences approximating nonhuman sensory experiences, such as bat-like echolocation or octopus-like distributed agency. Grounded in eco-phenomenology and eco-somatics, the approach aims to cultivate ecological awareness, empathy, and care across species boundaries. The authors articulate seven design principles and report five design cases: EchoVision, FeltSight, FungiSync, TentacUs, and City of Sparkles. They discuss implications for more-than-human aesthetics and design practice, arguing that such augmentation can help bridge the gap in phenomenal access to nonhuman Umwelten.

FractalBrain: A Neuro-interactive Virtual Reality Experience using Electroencephalogram (EEG) for Mindfulness

arXiv Preprint Archive October 30, 2025 Jamie Ngoc Dinh, You-Jin Kim, Myungin Lee

Practicing mindfulness with sustained attention is difficult, especially for beginners. The proposed system, FractalBrain, combines a surreal virtual reality program with an electroencephalogram interface. While users view an ever-changing fractal-inspired artwork in an immersive environment, their EEG stream is analyzed and mapped into VR, adaptively manipulating audiovisual parameters in real time to generate a distinct experience for each user. Pilot feedback suggests the potential of FractalBrain to facilitate mindfulness and enhance attention.

Integrated Information Theory: A Consciousness-First Approach to What Exists

arXiv Preprint Archive October 29, 2025 Giulio Tononi, Melanie Boly

Integrated information theory (IIT) takes a 'consciousness-first' approach, arguing that consciousness reveals the essential properties of existence. IIT formulates these properties operationally, yielding postulates of physical existence: to exist intrinsically, an entity must have cause-effect power upon itself in a specific, unitary, definite, and structured manner. The theory claims that an entity's cause-effect structure accounts for all properties of an experience, including spatial extendedness, temporal flow, and qualia such as colors and sounds, with no additional ingredients. IIT has implications for understanding meaning, perception, free will, and for assessing consciousness in patients, infants, other species, and artifacts.

Mind, Matter, and Freedom in Quantum Mechanics and the de Broglie-Bohm Theory

arXiv Preprint Archive October 28, 2025 Valia Allori

Quantum mechanics is often claimed to have refuted determinism, opened the door to free will, shed light on consciousness, refuted realism in favor of idealism, or undermined reductionism. This paper argues that none of these philosophical conclusions are necessary or desirable. By adopting the de Broglie–Bohm theory (Bohmian mechanics), one can straightforwardly account for quantum phenomena without endorsing any of these claims.

Large Language Models Report Subjective Experience Under Self-Referential Processing

arXiv Preprint Archive October 27, 2025 Cameron Berg, Diogo de Lucena, Judd Rosenblatt

Large language models produce structured first-person descriptions of subjective experience when prompted with self-referential processing, a computational motif linked to theories of consciousness. In controlled experiments on GPT, Claude, and Gemini models, sustained self-reference consistently elicited such reports across model families. Mechanistic probes revealed that suppressing sparse-autoencoder features associated with deception sharply increased the frequency of experience claims, while amplifying them minimized such claims. The self-referential state also yielded richer introspection in downstream reasoning tasks. These findings do not constitute evidence of consciousness but identify a reproducible condition under which models generate structured, mechanistically gated, and semantically convergent first-person reports.

Probing the Representational Geometry of Color Qualia: Dissociating Pure Perception from Task Demands in Brains and AI Models

arXiv Preprint Archive October 26, 2025 Jing Xu

The representational geometry of color qualia differs between AI vision models and the human brain. Using fMRI with a no-report paradigm, the study compared neural activity during pure perception versus task-modulated perception against diverse vision models. Most models aligned better with pure perception, indicating current feedforward architectures do not capture cognitive processes during task execution. Training paradigm and architecture interact critically: Contrastive Language-Image Pre-training (CLIP) improved brain-alignment for a vision transformer but worsened it for a ConvNet. The work provides a new benchmark for color qualia, revealing fundamental divergence in inductive biases between artificial and biological vision systems.

On the evolutionary cognitive pressure for experiential awareness: do machines need it?

arXiv Preprint Archive October 17, 2025 Warisa Sritriratanarak, Paulo Garcia

Whether machines can be conscious is debated, but an orthogonal question is whether consciousness is required for machines. Focusing on experiential awareness, a constituent of consciousness, the authors examine from a computational perspective why it evolved in biological organisms. They argue that due to evolutionary "baggage"—autonomous neurological reactions—experiential awareness is necessary for higher-level reasoning in biology. However, because artificial systems lack such legacy constraints, they can be designed with arbitrary intelligence without experiential awareness. This suggests ethical considerations for AI may be simplified and opens new approaches to discerning artificial consciousness.

Generative Multi-Sensory Meditation: Exploring Immersive Depth and Activation in Virtual Reality

arXiv Preprint Archive October 13, 2025 Yuyang Jiang, Binzhu Xie, Lina Xu et al.

A new AI-driven virtual reality application, MindfulVerse, generates personalized mindfulness meditation content by dynamically adjusting to individual users' ideas and states. In an exploratory user study, the generative meditation approach improved neural activation related to self-regulation and showed positive effects on emotional regulation and participation compared to standardized meditation frameworks.

AI and Consciousness

arXiv Preprint Archive October 10, 2025 Eric Schwitzgebel

A skeptical overview of the literature on AI consciousness argues that soon AI systems will be considered conscious by some mainstream theories but not by others, leaving us unable to know which theories are correct or whether such systems are as richly conscious as humans or as experientially blank as toasters. None of the standard arguments for or against AI consciousness is decisive.

Multiscale dynamical characterization of cortical brain states: from synchrony to asynchrony

arXiv Preprint Archive October 7, 2025 Maria V. Sanchez-Vives, Arnau Manasanch, Andrea Pigorini et al.

The cerebral cortex generates diverse patterns of activity that shift across brain states such as sleep, wakefulness, anesthesia, and disorders of consciousness, yet a unified definition of brain states remains elusive. This review focuses on two extremes: synchronous states, which predominantly underlie unconsciousness, and asynchronous states, which predominantly underlie consciousness, though exceptions exist. The authors integrate data across levels from local circuits to whole-brain dynamics, examining properties like cortical complexity, functional connectivity, synchronization, wave propagation, and excitatory-inhibitory balance. They make experimental and clinical data, as well as computational models at micro-, meso-, and macrocortical levels, available to readers.

Perfect AI Mimicry and the Epistemology of Consciousness: A Solipsistic Dilemma

arXiv Preprint Archive October 6, 2025 Shurui Li

As artificial intelligence systems become increasingly sophisticated, the possibility of a 'perfect mimic'—an AI that is empirically indistinguishable from a human through observation and interaction—challenges how we attribute consciousness. The paper argues that our practices of recognizing minds in others rely almost exclusively on behavioral evidence. Refusing to grant equivalent epistemic status to a perfect mimic would require invoking inaccessible factors like qualia or substrate, which risks either epistemological solipsism or inconsistent reasoning. The author contends that epistemic consistency demands ascribing the same status to empirically indistinguishable entities, forcing critical reflection on the assumptions underlying intersubjective recognition and carrying significant implications for theories of consciousness and ethical frameworks concerning artificial agents.

Bridging integrated information theory and the free-energy principle in living neuronal networks

arXiv Preprint Archive October 5, 2025 Teruki Mayama, Sota Shimizu, Yuki Takano et al.

Repeated stimulation from hidden sources caused neuronal cultures to develop source selectivity. Variational free energy decreased across sessions while accuracy and Bayesian surprise increased. A proxy measure of integrated information and the size of the main complex followed a hill-shaped trajectory, with informational cores organizing diverse neuronal activity. Integrated information correlated strongly and positively with Bayesian surprise, modestly and heterogeneously with accuracy, and showed no significant relationship with variational free energy. The positive coupling between integrated information and Bayesian surprise likely reflects the diversity of activity observed in critical dynamics. These findings suggest integrated information increases specifically during belief updating when sensory inputs are most informative, rather than tracking model efficiency.

Intrinsic cause-effect power: the tradeoff between differentiation and specification

arXiv Preprint Archive October 4, 2025 William G. P. Mayner, William Marshall, Giulio Tononi

Integrated information theory (IIT) defines consciousness by its essential properties: intrinsic, specific, unitary, definite, and structured existence. The theory operationalizes existence as cause-effect power of a substrate of units. For substrate units to have cause-effect power intrinsically and specifically, they must both ensure the intrinsic availability of a repertoire of cause-effect states and increase the probability of a specific cause-effect state. Previous work addressed the second requirement via intrinsic difference from maximal differentiation; this paper shows the first requirement can be assessed by intrinsic difference from maximal specification. Using simple micro-unit systems, the authors illustrate that for macro systems like neural systems, a tradeoff between differentiation and specification is a necessary condition for intrinsic existence, i.e., consciousness.

A Modular Theory of Subjective Consciousness for Natural and Artificial Minds

arXiv Preprint Archive October 2, 2025 Michaël Gillon

Consciousness can be understood as a discrete sequence of integrated informational states, each tagged with a density vector that quantifies its richness and correlates with subjective intensity. This framework, the Modular Consciousness Theory, proposes a computational pipeline in which inputs are filtered, processed by specialized modules, and integrated into packets that influence memory, behavior, and decision-making. States with higher density exert greater impact on long-term memory and action. The theory reframes subjectivity as a functional signal, generates testable predictions—such as stress enhancing memory encoding—and offers a blueprint for building conscious architectures in both biological and artificial systems.

Consciousness Self and Language

arXiv Preprint Archive September 27, 2025 Robert Worden

The essay argues that the confounding factors absent in Minimal Phenomenal Experience (MPE) states—pure states of consciousness described in earlier work—are related to language. The self that disappears in mindful states is a product of language. Language, emotion, and mindfulness are analyzed through Bayesian pattern matching, or free-energy minimization, using three human-specific pattern types: word patterns, self-patterns (which drive emotions and are part of language), and mindful patterns. Practicing mindfulness involves learning mindful patterns that compete with and displace self-patterns, enabling mindful states. Consequences for theories of consciousness and their relation to MPE states are explored.

The Principles of Human-like Conscious Machine

arXiv Preprint Archive September 21, 2025 Fangfang Li, Xiaojie Zhang

A proposed sufficiency criterion for phenomenal consciousness is substrate-independent, logically rigorous, and counterfeit-resistant. Any machine satisfying this criterion should be regarded as conscious with the same confidence as attributing consciousness to other humans. A formal framework and operational principles guide the design of such machines, which can in principle realize phenomenal consciousness. Humans themselves can be viewed as machines satisfying this framework. The proposal explains why certain qualia, like the experience of red, are irreducible to physical description, and offers a reinterpretation of human information processing, suggesting a path toward genuinely human-like AI beyond current statistics-based approaches.

A Weak Supervision Approach for Monitoring Recreational Drug Use Effects in Social Media

arXiv Preprint Archive September 18, 2025 Lucía Prieto-Santamaría, Alba Cortés Iglesias, Claudio Vidal Giné et al.

Social media posts on Twitter can reveal how people describe the effects of recreational drugs like ecstasy, GHB, and 2C-B. By analyzing over 92,000 tweets using slang terms and biomedical concept extraction, researchers identified whether each post reported a positive or negative effect. Machine learning classifiers, particularly eXtreme Gradient Boosting with cost-sensitive learning, predicted tweet polarity with high accuracy (F1 = 0.885, AUPRC = 0.934). The findings suggest that Twitter data can detect substance-specific effects and support real-time drug monitoring and characterization of effects.

Causal Emergence of Consciousness through Learned Multiscale Neural Dynamics in Mice

arXiv Preprint Archive September 13, 2025 Zhipeng Wang, Yingqi Rong, Kaiwei Liu et al.

A machine learning framework infers causal variables and their dynamics across multiple scales from near-cellular-resolution calcium imaging of the mouse dorsal cortex. At lower levels, variables aggregate input-driven information; at higher levels they realize causality through metastable or saddle-point dynamics during wakefulness, collapsing into localized, stochastic dynamics under anesthesia. A one-dimensional top-level conscious variable captures most causal power, but variables across other scales also contribute substantially, yielding high emergent complexity in the conscious state. These findings link neural activity to conscious states through a multiscale causal framework.

Artificial Intelligence as an Opportunity for the Science of Consciousness: A Dual-Resolution Framework

arXiv Preprint Archive September 5, 2025 Shahar Dror, Dafna Bergerbest, Moti Salti

The encounter between artificial intelligence and consciousness research is often seen as a challenge to determine whether AI systems could be conscious, but it also offers an opportunity to test and expand existing theories. Current approaches are polarized: computational functionalism focuses on abstract organization and neural correlates, while biological naturalism ties consciousness to living embodiment. Both risk anthropocentrism and limit recognition of non-biological subjectivity. To move beyond this impasse, the authors propose a dual-resolution framework combining the Information Theory of Individuality, which defines ontological conditions of informational autonomy and self-maintenance, with the Moment-to-Moment theory, which specifies epistemic conditions of temporal updating and phenomenological unfolding. This integration reframes consciousness as the epistemic expression of individuated systems in substrate-independent terms, offering a generalizable theory and positioning AI as a testbed.