PoseMaster统一3D姿态风格化与生成,提升精度与效率。
PoseMaster: A Unified 3D Native Framework for Stylized Pose Generation
- 将姿态风格化与3D生成统一在框架内,避免2D转3D误差累积。
- 直接使用3D骨架引导,实现更精准的空间关系建模。
- 支持自动绑定生成可动画角色资产,适合游戏/影视制作。
姿态风格化旨在生成与目标姿态对齐的风格化内容,是2D、3D和视频领域的基础任务。现有3D方法多采用级联流程:先用2D基础模型操控图像姿态,再映射到3D表示,但该范式限制了3D姿态风格化的精度与多样性。为此,我们提出一种新范式,将姿态风格化与3D生成统一于一个连贯框架中,减少累积误差,提升模型效率与效果。不同于以往使用2D骨骼图像作为引导,我们直接利用3D骨骼,因其能更准确地表达3D空间与拓扑关系,显著增强模型在姿态风格化上的表现力。此外,我们构建了一个大规模的「图像-骨骼-网格」三元组数据集,使模型能联合学习身份保持与几何对齐。大量实验表明,PoseMaster在定性与定量指标上均显著优于当前最优方法。由于生成的3D网格与条件骨骼具有严格空间对齐,结合自动绑定模型后,可直接生成可动画资产,展现出在自动化角色绑定中的巨大潜力。
原文摘要 · Abstract (English)
Pose stylization, which aims to synthesize stylized content aligning with target poses, serves as a fundamental task across 2D, 3D, and video domains. In the 3D realm, prevailing approaches typically rely on a cascade pipeline: first manipulating the image pose via 2D foundation models and subsequently lifting it into 3D representations. However, this paradigm limits the precision and diversity of the 3d pose stylization. To this end, we propose a novel paradigm for 3D pose stylization that unifies pose stylization and 3D generation within a cohesive framework. This integration minimizes the risk of cumulative errors and enhances the model's efficiency and effectiveness. In addition, diverging from previous works that typically utilize 2D skeleton images as guidance, we directly utilize the 3D skeleton because it can provide a more accurate representation of 3D spatial and topological relationships, which significantly enhances the model's capacity to achieve richer and more precise pose stylization. Moreover, we develop a scalable data engine to construct a large-scale dataset of ''Image-Skeleton-Mesh'' triplets, enabling the model to jointly learn identity preservation and geometric alignment. Extensive experiments demonstrate that PoseMaster significantly outperforms state-of-the-art methods in both qualitative and quantitative metrics. Owing to the strict spatial alignment between the generated 3D meshes and the conditioning skeletons, PoseMaster enables the direct creation of animatable assets when coupled with automated skinning models, highlighting its compelling potential for automated character rigging.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。