arXiv:2604.09835cs.CVcs.AI2026-04

专注人脸细节的全身高斯人像生成方法,提升表情与几何保真度。

F3G-Avatar : Face Focused Full-body Gaussian Avatar

  • 分体式双分支结构,分别建模身体形变与人脸精细调整
  • 在AvatarReX数据集上人脸视图达PSNR 26.243/SSIM 0.964
  • 适合需要高保真人脸动画的虚拟角色应用

现有全身高斯人像方法多关注全局重建质量,常忽略面部几何与表情细节。根源在于面部表征能力有限,难以建模高频的姿势相关形变。为此,我们提出F3G-Avatar,一种基于多视角RGB视频与回归姿态/形状参数的全身、人脸感知人像合成方法。从着装的Momentum Human Rig(MHR)模板出发,渲染前后位置图并经双分支架构解码为3D高斯:一分支捕捉姿势依赖的非刚性形变,另一分支聚焦于头部几何与外观精修。生成的高斯融合后通过线性混合皮肤(LBS)绑定,并以可微高斯点阵渲染。训练结合重建与感知目标,引入人脸专用对抗损失以增强近景真实感。实验表明,该方法在渲染质量上表现优异,于AvatarReX数据集上人脸视图达到PSNR 26.243 / SSIM 0.964 / LPIPS 0.084。消融实验进一步验证了MHR模板与人脸精修分支的有效性。F3G-Avatar提供了一条实用且高质量的可动画全身人像合成路径。

原文摘要 · Abstract (English)

Existing full-body Gaussian avatar methods primarily optimize global reconstruction quality and often fail to preserve fine-grained facial geometry and expression details. This challenge arises from limited facial representational capacity that causes difficulties in modeling high-frequency pose-dependent deformations. To address this, we propose F3G-Avatar, a full-body, face-aware avatar synthesis method that reconstructs animatable human representations from multi-view RGB video and regressed pose/shape parameters. Starting from a clothed Momentum Human Rig (MHR) template, front/back positional maps are rendered and decoded into 3D Gaussians through a two-branch architecture: a body branch that captures pose-dependent non-rigid deformations and a face-focused deformation branch that refines head geometry and appearance. The predicted Gaussians are fused, posed with linear blend skinning (LBS), and rendered with differentiable Gaussian splatting. Training combines reconstruction and perceptual objectives with a face-specific adversarial loss to enhance realism in close-up views. Experiments demonstrate strong rendering quality, with face-view performance reaching PSNR/SSIM/LPIPS of 26.243/0.964/0.084 on the AvatarReX dataset. Ablations further highlight contributions of the MHR template and the face-focused deformation. F3G-Avatar provides a practical, high-quality pipeline for realistic, animatable full-body avatar synthesis.

全身建模高斯人体人脸细节

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。