让普通人通过说话和手势与AI协作建3D模型
3Description: An Intuitive Human-AI Collaborative 3D Modeling Approach
- 用语音和手势描述生成3D模型,降低建模门槛
- 基于OpenAI和MediaPipe的AI技术实现自然交互
- 适合设计初学者、教育场景及创意工作者
本文提出3Description,一种面向非专业人士的直观人机协同3D建模方法。该方法结合自然语言处理与计算机视觉技术(基于OpenAI和MediaPipe),使用户可通过语音和手势描述创建并调整3D模型。研究通过定性分析、产品评估与用户测试验证其有效性。系统为基于网页的应用,具备跨平台特性,旨在提升3D建模的可访问性与易用性。在人工智能与新兴媒体融合背景下,3Description不仅推动更包容、友好的设计流程,促进更多人参与未来3D世界构建,还强调人类与AI的共同创造,避免过度依赖技术,保留人类创造力。
原文摘要 · Abstract (English)
This paper presents 3Description, an experimental human-AI collaborative approach for intuitive 3D modeling. 3Description aims to address accessibility and usability challenges in traditional 3D modeling by enabling non-professional individuals to co-create 3D models using verbal and gesture descriptions. Through a combination of qualitative research, product analysis, and user testing, 3Description integrates AI technologies such as Natural Language Processing and Computer Vision, powered by OpenAI and MediaPipe. Recognizing the web has wide cross-platform capabilities, 3Description is web-based, allowing users to describe the desired model and subsequently adjust its components using verbal and gestural inputs. In the era of AI and emerging media, 3Description not only contributes to a more inclusive and user-friendly design process, empowering more people to participate in the construction of the future 3D world, but also strives to increase human engagement in co-creation with AI, thereby avoiding undue surrender to technology and preserving human creativity.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。