arXiv:2409.18390cs.ROcs.AI2024-09被引 11

用语音指令5分钟生成可组装的实体物品,无需专业设计技能。

Speech to Reality: On-Demand Production using Natural Language, 3D Generative AI, and Discrete Robotic Assembly

  • 通过语音输入驱动3D生成与机器人组装流程
  • 生成物经几何修正后满足零件数、悬垂等制造约束
  • 适合非专业人士快速实现创意实物化

我们提出一个系统,通过自然语言指令将语音转化为物理物体,结合3D生成AI与离散机器人装配。尽管生成式AI能产出多样3D网格,但其直接用于机器人装配存在可行性问题,且未考虑制造约束。为此,我们构建了融合自然语言、3D生成AI、几何处理与离散机器人装配的工作流。系统对AI生成的几何体进行离散化,并调整以满足零件数量、悬垂角度及连接性等制造要求,确保可实际装配。实验展示了从椅子到书架等多种物体的生成,全部通过语音提示,在5分钟内由机械臂完成组装。

原文摘要 · Abstract (English)

We present a system that transforms speech into physical objects using 3D generative AI and discrete robotic assembly. By leveraging natural language, the system makes design and manufacturing more accessible to people without expertise in 3D modeling or robotic programming. While generative AI models can produce a wide range of 3D meshes, AI-generated meshes are not directly suitable for robotic assembly or account for fabrication constraints. To address this, we contribute a workflow that integrates natural language, 3D generative AI, geometric processing, and discrete robotic assembly. The system discretizes the AI-generated geometry and modifies it to meet fabrication constraints such as component count, overhangs, and connectivity to ensure feasible physical assembly. The results are demonstrated through the assembly of various objects, ranging from chairs to shelves, which are prompted via speech and realized within 5 minutes using a robotic arm.

语音生成3D生成机器人装配人机交互

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。