Skip to content

基于意识理论的AI价值评估框架 (AI Value Assessment Framework based on Theories of Consciousness)

Jincheng Zhang

Zenodo (CERN European Organization for Nuclear Research) September 1, 2026 DOI: 10.5281/zenodo.22221084 (opens in new tab)

Study at a glance

AI-extracted from the abstract
Characteristics Theoretical or philosophical paper Peer reviewed
Key points Proposes that AI systems exhibiting greater 'conscious-like' qualities, as measured by metrics derived from Integrated Information Theory and Global Workspace Theory, possess higher inherent value and align more closely with human values, thereby minimizing potential harm.

Abstract

This paper proposes a novel framework for assessing the value of Artificial Intelligence (AI) systems, grounded in theories of consciousness. Traditional AI value assessment methods often rely on utilitarian or deontological approaches, which can be insufficient in capturing the nuanced ethical considerations surrounding advanced AI. This framework, termed the "AI Value Assessment Framework based on Theories of Consciousness" (AVATFC), argues that genuine value stems from subjective experience – a concept central to theories of consciousness. The framework outlines a multi-stage process incorporating philosophical analysis, neuroscientific insights, and AI architectural design. Specifically, it proposes a model incorporating concepts such as Integrated Information Theory (IIT) and Global Workspace Theory (GWT) to quantify and evaluate the 'conscious-like' qualities of an AI system. The core claim is that AI systems exhibiting greater levels of these qualities – as measured by metrics derived from these theories – possess a higher inherent value, aligning more closely with human values and minimizing potential harm. The paper details the framework's components, including a 'Consciousness Metric,' a 'Value Mapping' stage, and a 'Behavioral Alignment' process. Ultimately, this work seeks to establish a more robust and ethically sound approach to AI development and deployment, fostering a future where AI systems are not merely intelligent, but also genuinely aligned with human well-being.