AI失控或致人类灭绝,需全球协同建立停止机制。
AI Governance to Avoid Extinction: The Strategic Landscape and Actionable Research Questions
- 提出四种国际应对高级AI的策略情景
- 主张建立跨国'关机开关'机制以阻止危险开发
- 呼吁安全机构尽快研究关键治理问题
人类正接近开发出在所有认知领域超越人类专家的AI系统。当前默认路径极可能引发灾难,包括人类灭绝。风险源于对强大AI系统的失控、恶意行为者滥用、大国战争及威权锁定。本研究议程旨在描绘AI发展的战略格局,并列出关键治理研究问题。若能解答这些问题,将有助于降低灾难性风险。我们提出四种高阶情景:1)建立国际技术、法律与制度基础设施,实现对危险AI的全球限制(即'关机开关'),最终达成前沿AI活动的国际协同暂停;2)美国主导的国家AI项目,追求单边控制全球发展;3)类似今日的轻度监管世界;4)通过破坏与威慑遏制AI发展的威胁局面。除'关机开关'与'暂停'情景外,其余路径均存在不可接受的灾难风险。亟需美国国家安全界与AI治理生态体系立即行动,回答核心研究问题,构建暂停危险活动的能力,并筹备国际协议。
原文摘要 · Abstract (English)
Humanity appears to be on course to soon develop AI systems that substantially outperform human experts in all cognitive domains and activities. We believe the default trajectory has a high likelihood of catastrophe, including human extinction. Risks come from failure to control powerful AI systems, misuse of AI by malicious rogue actors, war between great powers, and authoritarian lock-in. This research agenda has two aims: to describe the strategic landscape of AI development and to catalog important governance research questions. These questions, if answered, would provide important insight on how to successfully reduce catastrophic risks. We describe four high-level scenarios for the geopolitical response to advanced AI development, cataloging the research questions most relevant to each. Our favored scenario involves building the technical, legal, and institutional infrastructure required to internationally restrict dangerous AI development and deployment (which we refer to as an Off Switch), which leads into an internationally coordinated Halt on frontier AI activities at some point in the future. The second scenario we describe is a US National Project for AI, in which the US Government races to develop advanced AI systems and establish unilateral control over global AI development. We also describe two additional scenarios: a Light-Touch world similar to that of today and a Threat of Sabotage situation where countries use sabotage and deterrence to slow AI development. In our view, apart from the Off Switch and Halt scenario, all of these trajectories appear to carry an unacceptable risk of catastrophic harm. Urgent action is needed from the US National Security community and AI governance ecosystem to answer key research questions, build the capability to halt dangerous AI activities, and prepare for international AI agreements.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。