arXiv:2508.00748cs.CVcs.AI2025-08中稿 · the IEEE Internati…被引 5

用人脸动态特征识别换脸虚拟人,防冒用更安全

Is It Really You? Exploring Biometric Verification Scenarios in Photorealistic Talking-Head Avatar Videos

  • 基于面部关键点动态变化建模,无需依赖外观相似性
  • 在真实数据集上实现80%的识别准确率(AUC)
  • 为虚拟形象身份验证提供可解释、轻量级的技术方案

随着逼真说话头像在虚拟会议、游戏和社交平台中日益普及,其带来的安全风险也愈发突出。攻击者可窃取用户头像,保留其外貌与声音,使视觉或听觉层面难以辨别真伪。本文探讨在该类场景下进行生物特征验证的可行性,核心问题是:当头像外观完全仿照本人时,是否可通过个体独特的面部动作模式实现可靠身份识别?为此,我们构建了一个基于先进单张图像生成模型GAGAvatar的真实头像视频数据集,包含真实与伪造视频。同时提出一种轻量级、可解释的时空图卷积网络架构,结合时间注意力池化机制,仅使用面部关键点捕捉动态表情特征。实验表明,仅依靠面部运动线索即可实现有效身份验证,AUC值接近80%。所提出的基准测试与系统已开源,旨在推动对虚拟形象通信系统中行为生物特征防御技术的重视。

原文摘要 · Abstract (English)

Photorealistic talking-head avatars are becoming increasingly common in virtual meetings, gaming, and social platforms. These avatars allow for more immersive communication, but they also introduce serious security risks. One emerging threat is impersonation: an attacker can steal a user's avatar, preserving his appearance and voice, making it nearly impossible to detect its fraudulent usage by sight or sound alone. In this paper, we explore the challenge of biometric verification in such avatar-mediated scenarios. Our main question is whether an individual's facial motion patterns can serve as reliable behavioral biometrics to verify their identity when the avatar's visual appearance is a facsimile of its owner. To answer this question, we introduce a new dataset of realistic avatar videos created using a state-of-the-art one-shot avatar generation model, GAGAvatar, with genuine and impostor avatar videos. We also propose a lightweight, explainable spatio-temporal Graph Convolutional Network architecture with temporal attention pooling, that uses only facial landmarks to model dynamic facial gestures. Experimental results demonstrate that facial motion cues enable meaningful identity verification with AUC values approaching 80%. The proposed benchmark and biometric system are available for the research community in order to bring attention to the urgent need for more advanced behavioral biometric defenses in avatar-based communication systems.

生物特征识别虚拟形象对抗攻击行为识别

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。