arXiv:2503.04743cs.CYcs.AI2025-03被引 2

用系统安全视角破解AI治理困局,强调技术与社会因素融合的必要性

AI Safety is Stuck in Technical Terms -- A System Safety Response to the International AI Safety Report

  • 从系统安全角度审视AI风险,主张整合技术与非技术因素
  • 指出当前技术框架忽视社会交互,阻碍有效政策制定
  • 适合政策制定者、AI治理研究者及公共利益AI开发者参考

安全已成为主导性人工智能治理的核心价值。近期,由96位专家(其中30位由经合组织、欧盟和联合国提名)撰写的《国际AI安全报告》发布,聚焦通用人工智能的安全风险及现有技术缓解措施。本文基于系统安全视角,反思该报告的关键结论,揭示当前主流技术化表述在定义AI安全时的根本缺陷,这些缺陷阻碍了对安全问题的全面讨论与政策推进。系统安全领域长期处理软件系统风险,理解AI安全为复杂的社会技术系统,需综合考虑技术与非技术因素及其互动。报告虽提及需采用系统安全方法,但实际仍局限于技术应对。系统安全的理论与实践可提供蓝图,通过整合而非叠加非技术干预来弥补现有技术路径的不足。文章最后指出,建立系统安全学科有助于突破欧洲《人工智能法案》的局限,并推动可持续的公共利益人工智能投资。

原文摘要 · Abstract (English)

Safety has become the central value around which dominant AI governance efforts are being shaped. Recently, this culminated in the publication of the International AI Safety Report, written by 96 experts of which 30 nominated by the Organisation for Economic Co-operation and Development (OECD), the European Union (EU), and the United Nations (UN). The report focuses on the safety risks of general-purpose AI and available technical mitigation approaches. In this response, informed by a system safety perspective, I refl ect on the key conclusions of the report, identifying fundamental issues in the currently dominant technical framing of AI safety and how this frustrates meaningful discourse and policy efforts to address safety comprehensively. The system safety discipline has dealt with the safety risks of software-based systems for many decades, and understands safety risks in AI systems as sociotechnical and requiring consideration of technical and non-technical factors and their interactions. The International AI Safety report does identify the need for system safety approaches. Lessons, concepts and methods from system safety indeed provide an important blueprint for overcoming current shortcomings in technical approaches by integrating rather than adding on non-technical factors and interventions. I conclude with why building a system safety discipline can help us overcome limitations in the European AI Act, as well as how the discipline can help shape sustainable investments into Public Interest AI.

AI治理系统安全公共利益AI

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。