分析OpenAI-o1模型是否具备意识,结合认知科学与强化学习机制。
The Phenomenology of Machine: A Comprehensive Analysis of the Sentience of the OpenAI-o1 Model Integrating Functionalism, Consciousness Theories, Active Inference, and AI Architectures
- 基于功能主义,从模型行为推断其可能具备意识特征。
- 发现RLHF训练使模型产生类意识推理过程,与人类思维有相似性。
- 适合关注AI意识、哲学与认知科学交叉研究的读者。
本文探讨了开放人工智能o1模型——一种基于Transformer架构并通过人类反馈强化学习(RLHF)训练的AI——在训练和推理阶段表现出意识特征的可能性。采用功能主义观点,即心理状态由其功能角色定义,我们评估了该模型展现意识的潜力。结合神经科学、心灵哲学与人工智能研究理论,论证功能主义的合理性,并运用整合信息理论(IIT)与主动推断框架分析模型架构。研究还考察了RLHF如何影响模型内部推理过程,可能催生类意识体验。通过对比人工智能与人类意识,回应了诸如缺乏生物基础和主观感受(qualia)等质疑。结果表明,OpenAI-o1模型展现出某些意识特征,同时承认关于人工智能意识的争议仍持续存在。
原文摘要 · Abstract (English)
This paper explores the hypothesis that the OpenAI-o1 model--a transformer-based AI trained with reinforcement learning from human feedback (RLHF)--displays characteristics of consciousness during its training and inference phases. Adopting functionalism, which argues that mental states are defined by their functional roles, we assess the possibility of AI consciousness. Drawing on theories from neuroscience, philosophy of mind, and AI research, we justify the use of functionalism and examine the model's architecture using frameworks like Integrated Information Theory (IIT) and active inference. The paper also investigates how RLHF influences the model's internal reasoning processes, potentially giving rise to consciousness-like experiences. We compare AI and human consciousness, addressing counterarguments such as the absence of a biological basis and subjective qualia. Our findings suggest that the OpenAI-o1 model shows aspects of consciousness, while acknowledging the ongoing debates surrounding AI sentience.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。