arXiv:2507.10542cs.GRcs.AI2025-07International Conf…被引 18

用局部表情贴片提升3D虚拟人面部细节,实现实时高保真动画。

ScaffoldAvatar: High-Fidelity Gaussian Avatars with Patch Expressions

  • 通过局部贴片表达控制3D高斯点云动态生成
  • 在3K分辨率下实现高质量渲染与快速收敛
  • 适合需要精细面部表情的沉浸式交互场景

生成高保真、实时动画的写实3D人脸虚拟人对沉浸式远程通信和电影制作至关重要。尤其在近距离展现微表情和细部动作时挑战巨大。本文提出ScaffoldAvatar,将局部表情特征与3D高斯溅射结合,通过基于贴片的几何人脸模型提取局部表达,并利用Scaffold-GS的层级结构锚点,将这些表达映射为局部皮肤外观与运动,实时合成3D高斯点云,条件依赖于贴片表达和视角。采用基于颜色的密集化与渐进训练策略,在3K分辨率下实现高质量结果与更快收敛。ScaffoldAvatar在保持自然动态的同时,持续达到当前最佳性能,可真实呈现多样化的面部表情与风格,支持实时运行。

原文摘要 · Abstract (English)

Generating high-fidelity real-time animated sequences of photorealistic 3D head avatars is important for many graphics applications, including immersive telepresence and movies. This is a challenging problem particularly when rendering digital avatar close-ups for showing character's facial microfeatures and expressions. To capture the expressive, detailed nature of human heads, including skin furrowing and finer-scale facial movements, we propose to couple locally-defined facial expressions with 3D Gaussian splatting to enable creating ultra-high fidelity, expressive and photorealistic 3D head avatars. In contrast to previous works that operate on a global expression space, we condition our avatar's dynamics on patch-based local expression features and synthesize 3D Gaussians at a patch level. In particular, we leverage a patch-based geometric 3D face model to extract patch expressions and learn how to translate these into local dynamic skin appearance and motion by coupling the patches with anchor points of Scaffold-GS, a recent hierarchical scene representation. These anchors are then used to synthesize 3D Gaussians on-the-fly, conditioned by patch-expressions and viewing direction. We employ color-based densification and progressive training to obtain high-quality results and faster convergence for high resolution 3K training images. By leveraging patch-level expressions, ScaffoldAvatar consistently achieves state-of-the-art performance with visually natural motion, while encompassing diverse facial expressions and styles in real time.

3D虚拟人高斯溅射面部动画

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。