arXiv:2501.01589cs.CVcs.GR2025-01被引 10

从单目视频重建可分离的穿衣人体,支持动画与换装应用

D$^3$-Human: Dynamic Disentangled Digital Human from Monocular Video

  • 结合显式与隐式表示,用SDF和新提出的hmSDF分离可见躯干与衣物
  • 在遮挡区域保持几何合理性,实现高质量可分离穿衣人体重建
  • 适合需要精细人体-服装分离的动画制作与虚拟试衣场景

我们提出D$^3$-Human,一种从单目视频重建动态可分离数字人体几何的方法。以往基于单目视频的人体重建多聚焦于未解耦的穿衣人体或仅重建服装,难以直接应用于动画制作等场景。其核心挑战在于衣物对身体造成的遮挡,需确保可见区域细节与不可见区域的合理性。为此,我们结合显式与隐式表示:将可见区域重建为有向距离场(SDF),并提出一种新型人体流形有向距离场(hmSDF)以分割可见衣物与可见身体,再融合可见与不可见身体部分。大量实验表明,相比现有方案,D$^3$-Human能实现不同服装下高质量可分离的人体重建,并可直接用于服装迁移与动画生成。

原文摘要 · Abstract (English)

We introduce D$^3$-Human, a method for reconstructing Dynamic Disentangled Digital Human geometry from monocular videos. Past monocular video human reconstruction primarily focuses on reconstructing undecoupled clothed human bodies or only reconstructing clothing, making it difficult to apply directly in applications such as animation production. The challenge in reconstructing decoupled clothing and body lies in the occlusion caused by clothing over the body. To this end, the details of the visible area and the plausibility of the invisible area must be ensured during the reconstruction process. Our proposed method combines explicit and implicit representations to model the decoupled clothed human body, leveraging the robustness of explicit representations and the flexibility of implicit representations. Specifically, we reconstruct the visible region as SDF and propose a novel human manifold signed distance field (hmSDF) to segment the visible clothing and visible body, and then merge the visible and invisible body. Extensive experimental results demonstrate that, compared with existing reconstruction schemes, D$^3$-Human can achieve high-quality decoupled reconstruction of the human body wearing different clothing, and can be directly applied to clothing transfer and animation.

数字人人体重建可分离单目视频

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。