arXiv:2603.19649cs.SIcs.AI2026-03被引 7

用大模型模拟社交平台,提前测试政策效果,避免现实部署出错。

PolicySim: An LLM-Based Agent Social Simulation Sandbox for Proactive Policy Optimization

  • 用SFT和DPO训练用户代理,还原真实平台行为模式。
  • 引入上下文关联的强化学习机制,动态适应网络结构变化。
  • 可预测政策影响,适合平台安全团队做预演评估。

社交平台是信息传播的核心,用户行为与平台干预共同塑造观点。然而,推荐算法和内容过滤等干预策略可能无意中加剧回音室效应和群体极化,带来重大社会风险。现有方法多依赖上线后的反应式A/B测试,风险发现滞后且成本高。基于大模型的社会仿真提供了部署前的替代方案,但当前方法在真实建模平台干预及反馈机制方面仍有不足。为此,我们提出PolicySim——一个用于主动评估与优化干预政策的LLM驱动社会仿真沙盒。该系统通过两个核心组件实现用户行为与平台干预的双向动态建模:(1) 基于监督微调(SFT)和直接偏好优化(DPO)训练的用户代理模块,实现平台特定的行为真实性;(2) 采用消息传递的上下文老虎机算法的自适应干预模块,以捕捉动态网络结构。实验表明,PolicySim可在微观与宏观层面准确模拟平台生态,并支持有效干预策略生成。

原文摘要 · Abstract (English)

Social platforms serve as central hubs for information exchange, where user behaviors and platform interventions jointly shape opinions. However, intervention policies like recommendation and content filtering, can unintentionally amplify echo chambers and polarization, posing significant societal risks. Proactively evaluating the impact of such policies is therefore crucial. Existing approaches primarily rely on reactive online A/B testing, where risks are identified only after deployment, making risk identification delayed and costly. LLM-based social simulations offer a promising pre-deployment alternative, but current methods fall short in realistically modeling platform interventions and incorporating feedback from the platform. Bridging these gaps is essential for building actionable frameworks to assess and optimize platform policies. To this end, we propose PolicySim, an LLM-based social simulation sandbox for the proactive assessment and optimization of intervention policies. PolicySim models the bidirectional dynamics between user behavior and platform interventions through two key components: (1) a user agent module refined via supervised fine-tuning (SFT) and direct preference optimization (DPO) to achieve platform-specific behavioral realism; and (2) an adaptive intervention module that employs a contextual bandit with message passing to capture dynamic network structures. Experiments show that PolicySim can accurately simulate platform ecosystems at both micro and macro levels and support effective intervention policy.

社会仿真大模型应用政策评估平台治理

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。