arXiv:2608.24169cs.CVcs.GR2026-08

让AI像真人一样在Blender里直接改3D模型,不生成代码也不重做整体。

ViSculpt: Visual-Centric Agentic Geometry Editing

论文配图:ViSculpt: Visual-Centric Agentic Geometry Editing
图 1 · 摘自论文原文
  • 用多智能体模拟人类艺术家,通过观察视窗和模拟操作来编辑模型。
  • 在基准测试中能理解自然语言指令并局部修改网格,保持原模型整体特征。
  • 无需训练,直接在Blender界面操作,适合专业3D设计流程使用。

3D几何编辑是图形管线中关键但耗时的环节,需艺术家将创意转化为复杂软件中的精确操作。大语言模型(LLMs)在基于脚本的3D创作中展现潜力,但在感知驱动的现有网格编辑中效果有限,因执行需保持视觉一致性且未修改区域应保留。我们提出一种视觉中心、无需训练的多智能体系统,通过模拟人类艺术家的迭代工作流,在Blender中直接编辑现有3D网格。该系统不生成脚本或重构几何体,而是通过模拟用户交互,由多模态LLM智能体观察视口、推理当前网格状态并执行局部修改。在精心构建的基准测试中,初步证据表明该方法可遵循自然语言指令,完成典型局部网格编辑,并保留输入资产的整体身份。结果揭示了一种语言驱动3D编辑的新范式:在原生3D编辑流程中对现有网格进行直接就地修改。本研究为专业图形软件中的视觉中心型智能体几何编辑迈出探索性一步。

原文摘要 · Abstract (English)

3D geometry editing is a critical yet labor-intensive part of the graphics pipeline, requiring artists to translate creative intent into precise operations in complex professional software. Large language models (LLMs) have shown promise for script-based 3D creation, but script generation is less suited to perception-driven editing of arbitrary existing meshes, where execution must remain visually grounded and untouched regions should be preserved. We present a \emph{visual-centric}, training-free multi-agent system that edits existing 3D meshes directly in Blender by emulating the iterative workflow of human artists. Rather than generating scripts or regenerating geometry, our system operates through the Blender GUI: multimodal LLM agents observe the viewport, reason about the current mesh state, and execute localized edits through simulated user interactions. Experiments on a curated benchmark provide initial evidence that this agentic approach can follow natural language instructions, perform representative localized mesh edits, and preserve the overall identity of the input asset. Our results highlight a complementary regime for language-driven 3D editing: direct in-place modification of existing meshes within the native 3D editing workflow. We view this work as an exploratory step toward visual-centric agentic geometry editing in professional graphics software.

3D编辑多智能体视觉引导Blender

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。