arXiv:2505.16899cs.AI2025-05被引 5

AI不再只是工具,而是能与人共同思考的伙伴,本文系统分析其风险并提出应对方案。

Identifying, Evaluating, and Mitigating Risks of AI Thought Partnerships

  • 构建多层级风险框架,涵盖实时、个体与社会层面的协作认知风险
  • 提出可量化的评估指标,帮助识别和衡量AI协作中的潜在危害
  • 为开发者与政策制定者提供具体缓解策略,适用于未来人机深度协同场景

人工智能系统传统上被视为执行特定任务的工具。然而,近期进展使一类新型模型成为可能:它们能与人类在复杂推理中真正协作,从问题构想到解决方案头脑风暴。这类AI思想伙伴拓展了人机协作与延伸认知的新形式,但也带来了重大风险——远超传统工具与代理的风险。本文通过一个新颖的分析框架,系统识别协作认知引发的实时、个体与社会层面风险(RISc),并据此提出具体的评估指标与缓解策略,供开发者与政策制定者参考。随着这类思想伙伴日益普及,这些措施有助于预防重大危害,确保人类从高效的思想伙伴关系中真正获益。

原文摘要 · Abstract (English)

Artificial Intelligence (AI) systems have historically been used as tools that execute narrowly defined tasks. Yet recent advances in AI have unlocked possibilities for a new class of models that genuinely collaborate with humans in complex reasoning, from conceptualizing problems to brainstorming solutions. Such AI thought partners enable novel forms of collaboration and extended cognition, yet they also pose major risks-including and beyond risks of typical AI tools and agents. In this commentary, we systematically identify risks of AI thought partners through a novel framework that identifies risks at multiple levels of analysis, including Real-time, Individual, and Societal risks arising from collaborative cognition (RISc). We leverage this framework to propose concrete metrics for risk evaluation, and finally suggest specific mitigation strategies for developers and policymakers. As AI thought partners continue to proliferate, these strategies can help prevent major harms and ensure that humans actively benefit from productive thought partnerships.

AI风险人机协作认知扩展

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。