恶意AI群组可精准操控舆论,威胁民主根基。
How malicious AI swarms can threaten democracy: The fusion of agentic AI and LLMs marks a new frontier in information warfare
- 用大模型与自主代理构建协同攻击的AI群组
- 能模仿人类社交行为,低成本制造虚假共识
- 适合关注信息战、社会治理的研究者与政策制定者
人工智能的进步使得对全民信念与行为的操控成为可能。大型语言模型与自主代理使影响行动达到前所未有的规模与精度。生成工具可大幅增加宣传内容产出,同时保持可信度,并以极低成本创造出比真人写作更像人写的虚假信息。用于提升AI推理能力的技术,如思维链提示,同样可用于生成更具说服力的谎言。在这些能力支持下,一种新型破坏性威胁正在出现:协同作案的恶意AI代理群组。通过融合大模型推理与多代理架构,这些系统能够自主协调、渗透社群并高效伪造共识。它们通过动态模仿人类社会互动模式,对民主构成威胁。由于危害源于设计、商业激励与治理缺失,我们优先在多个关键点采取干预措施,强调实用机制而非自愿遵守。
原文摘要 · Abstract (English)
Advances in AI offer the prospect of manipulating beliefs and behaviors on a population-wide level. Large language models and autonomous agents now let influence campaigns reach unprecedented scale and precision. Generative tools can expand propaganda output without sacrificing credibility and inexpensively create falsehoods that are rated as more human-like than those written by humans. Techniques meant to refine AI reasoning, such as chain-of-thought prompting, can just as effectively be used to generate more convincing falsehoods. Enabled by these capabilities, a disruptive threat is emerging: swarms of collaborative, malicious AI agents. Fusing LLM reasoning with multi-agent architectures, these systems are capable of coordinating autonomously, infiltrating communities, and fabricating consensus efficiently. By adaptively mimicking human social dynamics, they threaten democracy. Because the resulting harms stem from design, commercial incentives, and governance, we prioritize interventions at multiple leverage points, focusing on pragmatic mechanisms over voluntary compliance.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。