arXiv:2504.02701cs.AIcs.MA2025-04

用可持续发展目标指导AI研发,防范攻击性AI风险

Responsible Development of Offensive AI

  • 以可持续发展目标为框架评估攻击性AI风险
  • 测试了漏洞探测代理与智能恶意软件的实战能力
  • 适合政策制定者与负责任AI研究者参考

随着人工智能的发展,亟需更广泛的共识来确定研究优先级。本文探讨了攻击性AI,并通过可持续发展目标(SDGs)和可解释性技术提供指导,旨在更有效地平衡社会收益与风险。研究评估了两种攻击性AI形式:解决攻防对抗挑战(Capture-The-Flag)的漏洞探测代理,以及由AI驱动的恶意软件。

原文摘要 · Abstract (English)

As AI advances, broader consensus is needed to determine research priorities. This endeavor discusses offensive AI and provides guidance by leveraging Sustainable Development Goals (SDGs) and interpretability techniques. The objective is to more effectively establish priorities that balance societal benefits against risks. The two forms of offensive AI evaluated in this study are vulnerability detection agents, which solve Capture- The-Flag challenges, and AI-powered malware.

AI安全攻击性AI可持续发展

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。