用多智能体系统让AI从草图自动生成可迭代的CAD模型
From Idea to CAD: A Language Model-Driven Multi-Agent System for Collaborative Design
- 构建视觉语言模型驱动的多智能体协作系统,分工负责需求、建模与质检
- 仅需草图或文字描述,即可自动生成参数化CAD模型并支持用户迭代优化
- 适合工业设计师与3D打印爱好者,降低专业门槛
使用计算机辅助设计(CAD)创建数字模型需要深厚的专业知识。在工业产品开发中,该过程通常涉及工程师团队,涵盖需求工程、CAD建模和质量保证。我们提出一种基于视觉语言模型(VLM)的多智能体系统,模拟团队结构,具备对参数化CAD工具及文档的访问能力。系统包含需求工程、CAD工程和基于视觉的质量保证三个智能体,可从草图或文本描述自动生成模型,并支持用户在迭代验证循环中协同优化。该方法有望提升设计效率,适用于行业专家与3D打印爱好者。我们在多个设计任务中验证了架构潜力,并通过若干消融实验展示了各组件的贡献。
原文摘要 · Abstract (English)
Creating digital models using Computer Aided Design (CAD) is a process that requires in-depth expertise. In industrial product development, this process typically involves entire teams of engineers, spanning requirements engineering, CAD itself, and quality assurance. We present an approach that mirrors this team structure with a Vision Language Model (VLM)-based Multi Agent System, with access to parametric CAD tooling and tool documentation. Combining agents for requirements engineering, CAD engineering, and vision-based quality assurance, a model is generated automatically from sketches and/ or textual descriptions. The resulting model can be refined collaboratively in an iterative validation loop with the user. Our approach has the potential to increase the effectiveness of design processes, both for industry experts and for hobbyists who create models for 3D printing. We demonstrate the potential of the architecture at the example of various design tasks and provide several ablations that show the benefits of the architecture's individual components.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。