arXiv:2602.08565cs.HCcs.AI2026-02被引 2

用AI代理模拟长期风险,结合专家判断提升系统性预见能力

Agent-Supported Foresight for AI Systemic Risks: AI Agents for Breadth, Experts for Judgment

  • 用未来之轮方法构建AI代理,模拟四种不同成熟度的AI应用风险
  • 每种应用生成86至110个后果,提炼出27至47个独特风险
  • 混合流程:代理拓展系统性覆盖,人类提供情境判断,适合政策制定者

AI影响评估常聚焦短期风险,因人类判断随时间推移而衰减,体现柯林里奇困境:前瞻性最需时知识却最少。为应对长期系统性风险,我们提出一种可扩展的方法,通过战略远见中的未来之轮(Futures Wheel)在仿真中运行AI代理。实验涵盖四个处于不同技术成熟度等级(TRL)的AI应用:聊天陪伴机器人(TRL 9,成熟)、AI玩具(TRL 7,中等)、哀悼机器人(TRL 5,低)、死亡应用(TRL 2,概念)。每项应用进行30次代理运行,共产生86-110个后果,归纳为27-47个独特风险。为对比代理输出与人类视角,我们收集了290名领域专家和7位领袖的评估,并组织42名专家与42名普通人的未来之轮研讨。结果显示,代理在多轮中生成大量系统性后果;相较之下,专家识别的风险较少,更具体但系统性较弱,认为更可能发生;普通人则提出更多情绪化关切,普遍缺乏系统性。我们建议采用混合远见流程:代理负责扩大系统覆盖,人类提供情境校准。数据集已公开:https://social-dynamics.net/ai-risks/foresight。

原文摘要 · Abstract (English)

AI impact assessments often stress near-term risks because human judgment degrades over longer horizons, exemplifying the Collingridge dilemma: foresight is most needed when knowledge is scarcest. To address long-term systemic risks, we introduce a scalable approach that simulates in-silico agents using the strategic foresight method of the Futures Wheel. We applied it to four AI uses spanning Technology Readiness Levels (TRLs): Chatbot Companion (TRL 9, mature), AI Toy (TRL 7, medium), Griefbot (TRL 5, low), and Death App (TRL 2, conceptual). Across 30 agent runs per use, agents produced 86-110 consequences, condensed into 27-47 unique risks. To benchmark the agent outputs against human perspectives, we collected evaluations from 290 domain experts and 7 leaders, and conducted Futures Wheel sessions with 42 experts and 42 laypeople. Agents generated many systemic consequences across runs. Compared with these outputs, experts identified fewer risks, typically less systemic but judged more likely, whereas laypeople surfaced more emotionally salient concerns that were generally less systemic. We propose a hybrid foresight workflow, wherein agents broaden systemic coverage, and humans provide contextual grounding. Our dataset is available at: https://social-dynamics.net/ai-risks/foresight.

AI风险系统性风险远见模拟人机协同

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。