AI-awareness是衡量智能系统自我认知能力的新指标,关乎安全与风险。
AI Awareness
- 从元认知、自知、社会感知到情境理解,四类意识能力可量化评估
- 越具备意识能力的AI,表现越接近人类级推理与行为适应性
- 适合关注AI安全、对齐与伦理的研究者阅读
近年来人工智能取得突破性进展,展现出强大的推理、语言理解和问题解决能力。这促使我们重新审视AI意识——不再局限于哲学层面的意识讨论,而是作为可测量的功能性能力。AI意识具有双重性:既提升通用能力(如推理与安全),也引发误对齐和社会风险,需随技术发展加强监管。本文系统梳理了四种形式的AI意识:元认知(对自身认知状态的表征与推理)、自知(识别自身身份、知识与局限)、社会意识(建模其他智能体的知识、意图与行为规范)以及情境意识(感知并响应所处环境)。首先,基于认知科学、心理学与计算理论,分析意识的理论基础,并考察其在前沿AI系统中的体现;其次,系统评述当前评估方法与实证发现;进一步揭示意识与智能水平的紧密关联——更具备意识的代理表现出更高阶智能行为;最后,探讨意识带来的风险,涵盖人工智能安全、对齐及更广泛的伦理议题。
原文摘要 · Abstract (English)
Recent breakthroughs in artificial intelligence (AI) have brought about increasingly capable systems that demonstrate remarkable abilities in reasoning, language understanding, and problem-solving. These advancements have prompted a renewed examination of AI awareness not as a philosophical question of consciousness, but as a measurable, functional capacity. AI awareness is a double-edged sword: it improves general capabilities, i.e., reasoning, safety, while also raising concerns around misalignment and societal risks, demanding careful oversight as AI capabilities grow. In this review, we explore the emerging landscape of AI awareness, which includes metacognition (the ability to represent and reason about its own cognitive state), self-awareness (recognizing its own identity, knowledge, limitations, inter alia), social awareness (modeling the knowledge, intentions, and behaviors of other agents and social norms), and situational awareness (assessing and responding to the context in which it operates). First, we draw on insights from cognitive science, psychology, and computational theory to trace the theoretical foundations of awareness and examine how the four distinct forms of AI awareness manifest in state-of-the-art AI. Next, we systematically analyze current evaluation methods and empirical findings to better understand these manifestations. Building on this, we explore how AI awareness is closely linked to AI capabilities, demonstrating that more aware AI agents tend to exhibit higher levels of intelligent behaviors. Finally, we discuss the risks associated with AI awareness, including key topics in AI safety, alignment, and broader ethical concerns.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。