新基准测试评估顶尖伪造视频检测模型在真实场景下的表现
TalkingHeadBench: A Multi-Modal Benchmark & Analysis of Talking-Head DeepFake Detection
- 构建多模型多生成器的综合评测框架,涵盖学术与商业级伪造技术
- 发现现有检测模型在身份和生成器分布变化下泛化能力普遍不足
- 提供可视化分析工具,揭示检测器常见失败模式与偏差
由先进生成模型驱动的说话头深伪视频技术迅速发展,其逼真度已对媒体、政治和金融等领域构成重大风险。然而,当前的深伪说话头检测基准仍依赖过时的生成器,难以反映最新进展,且缺乏对模型鲁棒性和泛化能力的深入分析。我们提出TalkingHeadBench,一个全面的多模型、多生成器基准与精选数据集,用于评估前沿检测器在最先进生成器上的表现。数据集包含由领先学术与商业模型生成的深伪视频,并设计了精心构造的测试协议,以评估在身份和生成器特征分布偏移下的泛化能力。我们对多种检测方法(包括CNN、视觉变换器和时序模型)进行了基准测试,并通过Grad-CAM可视化进行误差分析,揭示常见失败模式与检测器偏差。TalkingHeadBench已在Hugging Face公开,所有数据划分与协议均可访问。该基准旨在推动更鲁棒、泛化性更强的检测模型研究,应对快速演进的生成技术挑战。
原文摘要 · Abstract (English)
The rapid advancement of talking-head deepfake generation fueled by advanced generative models has elevated the realism of synthetic videos to a level that poses substantial risks in domains such as media, politics, and finance. However, current benchmarks for deepfake talking-head detection fail to reflect this progress, relying on outdated generators and offering limited insight into model robustness and generalization. We introduce TalkingHeadBench, a comprehensive multi-model multi-generator benchmark and curated dataset designed to evaluate the performance of state-of-the-art detectors on the most advanced generators. Our dataset includes deepfakes synthesized by leading academic and commercial models and features carefully constructed protocols to assess generalization under distribution shifts in identity and generator characteristics. We benchmark a diverse set of existing detection methods, including CNNs, vision transformers, and temporal models, and analyze their robustness and generalization capabilities. In addition, we provide error analysis using Grad-CAM visualizations to expose common failure modes and detector biases. TalkingHeadBench is hosted on https://huggingface.co/datasets/luchaoqi/TalkingHeadBench with open access to all data splits and protocols. Our benchmark aims to accelerate research towards more robust and generalizable detection models in the face of rapidly evolving generative techniques.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。