用游戏模拟网络论坛,看大模型如何被用来造谣或辟谣。
Sword and Shield: Uses and Strategies of LLMs in Navigating Disinformation
- 设计类狼人杀游戏,测试不同角色用大模型策略。
- 发现角色身份决定大模型用途:造谣者用它撒谎,管理员用它查证。
- 适合研究信息战、平台治理和大模型安全的学者看。
大型语言模型(LLMs)在对抗虚假信息方面呈现出双重挑战。这些能够大规模生成类人文本的强大工具,既可能被用于制造复杂且有说服力的虚假信息,也具备增强检测与缓解策略的潜力。本文通过受狼人杀启发的通信博弈,模拟在线论坛环境,以25名参与者进行实验,分析欺骗者、管理员和普通用户如何利用大模型实现各自目标。研究揭示了不同角色对大模型的不同使用方式及其策略效果,强调理解其在该语境下的有效性至关重要。最后讨论了对未来发展及平台设计的启示,倡导在赋能用户、建立信任的同时,防范大模型辅助虚假信息的风险。
原文摘要 · Abstract (English)
The emergence of Large Language Models (LLMs) presents a dual challenge in the fight against disinformation. These powerful tools, capable of generating human-like text at scale, can be weaponised to produce sophisticated and persuasive disinformation, yet they also hold promise for enhancing detection and mitigation strategies. This paper investigates the complex dynamics between LLMs and disinformation through a communication game that simulates online forums, inspired by the game Werewolf, with 25 participants. We analyse how Disinformers, Moderators, and Users leverage LLMs to advance their goals, revealing both the potential for misuse and combating disinformation. Our findings highlight the varying uses of LLMs depending on the participants' roles and strategies, underscoring the importance of understanding their effectiveness in this context. We conclude by discussing implications for future LLM development and online platform design, advocating for a balanced approach that empowers users and fosters trust while mitigating the risks of LLM-assisted disinformation.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。