arXiv:2603.28152cs.CV2026-03

让2D图像编辑拥有3D真实感,拖拽即可精准变形。

ObjectMorpher: 3D-Aware Image Editing via Deformable 3DGS Models

  • 用3D高斯溅射重建目标物体,实现可交互的3D编辑
  • 拖动控制点后通过ARAP约束保持形状合理,生成效果逼真
  • 适合需要精细、真实感图像修改的设计师和研究者

在图像编辑中实现精确的物体级控制仍具挑战:2D方法缺乏3D感知,常导致模糊或不合理的结果;现有3D感知方法依赖复杂优化或不完整的单目重建。我们提出ObjectMorpher,一个统一且交互式的框架,将模糊的2D编辑转化为基于几何的操作。该框架通过图像到3D生成器将目标实例提升至可编辑的3D高斯溅射(3DGS)空间,支持快速、身份保持的操纵。用户拖动控制点,采用基于图的非刚性变形并施加尽可能刚性(ARAP)约束,确保形状与姿态变化物理合理。复合扩散模块协调光照、颜色与边界,实现无缝重融合。在多种类别上,ObjectMorpher均实现了细粒度、逼真的编辑,在KID、LPIPS、SIFID指标及用户偏好上优于2D拖拽与3D-aware基线方法。

原文摘要 · Abstract (English)

Achieving precise, object-level control in image editing remains challenging: 2D methods lack 3D awareness and often yield ambiguous or implausible results, while existing 3D-aware approaches rely on heavy optimization or incomplete monocular reconstructions. We present ObjectMorpher, a unified, interactive framework that converts ambiguous 2D edits into geometry-grounded operations. ObjectMorpher lifts target instances with an image-to-3D generator into editable 3D Gaussian Splatting (3DGS), enabling fast, identity-preserving manipulation. Users drag control points; a graph-based non-rigid deformation with as-rigid-as-possible (ARAP) constraints ensures physically sensible shape and pose changes. A composite diffusion module harmonizes lighting, color, and boundaries for seamless reintegration. Across diverse categories, ObjectMorpher delivers fine-grained, photorealistic edits with superior controllability and efficiency, outperforming 2D drag and 3D-aware baselines on KID, LPIPS, SIFID, and user preference.

3D编辑图像生成高斯溅射交互式

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。