小模型可自动生成政治宣传内容,且效果受人设影响大于模型本身。
AI Propaganda factories with language models
- 用小模型模拟不同人设生成连贯政治言论
- 对抗回复使极端内容增多,立场更坚定
- 适合关注舆论战防御与自动化传播检测者
如今,基于AI的影响操作已可在普通硬件上端到端运行。我们发现小型语言模型能生成连贯、具有人格特征的政治信息,并可通过自动化方式评估而无需人工标注。两个行为发现:一是‘人设大于模型’——人设设计对行为解释力超过模型身份;二是‘回应作为压力源’——当需反驳对立观点时,意识形态认同增强,极端内容比例上升。结果表明,完全自动化的宣传内容生成已可被大型和小型行动方实现。因此,防御策略应从限制模型访问转向聚焦对话检测与破坏传播网络。讽刺的是,这些操作的高一致性本身反而成为可被识别的信号。
原文摘要 · Abstract (English)
AI-powered influence operations can now be executed end-to-end on commodity hardware. We show that small language models produce coherent, persona-driven political messaging and can be evaluated automatically without human raters. Two behavioural findings emerge. First, persona-over-model: persona design explains behaviour more than model identity. Second, engagement as a stressor: when replies must counter-arguments, ideological adherence strengthens and the prevalence of extreme content increases. We demonstrate that fully automated influence-content production is within reach of both large and small actors. Consequently, defence should shift from restricting model access towards conversation-centric detection and disruption of campaigns and coordination infrastructure. Paradoxically, the very consistency that enables these operations also provides a detection signature.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。