arXiv:2604.25982cs.LGcs.AI2026-04

梳理前沿人工智能风险治理的未解难题,指明各方责任。

Open Problems in Frontier AI Risk Management

  • 按风险管理全流程梳理开放问题,分三类:共识缺失、框架冲突、执行不足。
  • 识别出开发者、监管者等不同主体在解决各类问题中的最优角色。
  • 不提具体方案,而是提供可更新的协作指南,助研究与治理协同。

前沿人工智能既放大了现有风险,也带来了本质全新的挑战。技术快速迭代导致科学共识难以形成,新兴的安全实践常与既有风险管理体系错位甚至破坏其有效性。为此,本文系统性地揭示前沿人工智能风险治理中的开放问题。采用问题导向方法,从风险规划、识别、分析、评估到缓解,逐环节梳理文献,发现未解决的挑战及最适于应对的行动方。根据问题性质,将开放问题分为三类:(a)科学或技术共识缺失;(b)与既有风险管理体系不一致或构成挑战;(c)虽有共识与对齐但实施不力。通过映射这些问题并定位关键行动者——包括开发者、部署方、监管机构、标准组织、研究人员及第三方评估者——本工作旨在明确推进方向,推动对前沿人工智能风险治理的稳健且有意义的共识。本文不提出具体解决方案,而是作为面向未来的议题设定参考文献,辅以动态在线资源库,支持协调、减少重复,并引导后续研究与治理努力。

原文摘要 · Abstract (English)

Frontier AI both amplifies existing risks and introduces qualitatively novel challenges. Not only is there a notable lack of stable scientific consensus resulting from the rapid pace of technological change, but emerging frontier AI safety practices are often misaligned with, or may undermine, established risk management frameworks. To address these challenges, we systematically surface open problems in frontier AI risk management. Adopting a problem-oriented approach, we examine each stage of the risk management process - risk planning, identification, analysis, evaluation, and mitigation - through a structured review of the literature, identifying unresolved challenges and the actors best positioned to address them. Recognising that different types of open problems call for different responses, we classify open problems according to whether they reflect (a) a lack of scientific or technical consensus, (b) misalignment with, or challenges to, established risk management frameworks, or (c) shortcomings in implementation despite apparent consensus and alignment. By mapping these open problems and identifying the actors best positioned to address them - including developers, deployers, regulators, standards bodies, researchers, and third-party evaluators - this work aims to clarify where progress is needed to enable robust and meaningful consensus on frontier AI risk management.The paper does not propose specific solutions; instead, it provides a problem-oriented, agenda-setting reference document, complemented by a living online repository, intended to support coordination, reduce duplication, and guide future research and governance efforts.

AI治理风险管控前沿AI开放问题

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。