机器人2分钟生成个性肖像画,还能互动调整风格。
SketcherX: AI-Driven Interactive Robotic drawing with Diffusion model and Vectorization Techniques
- 用扩散模型+向量优化,让机器人画出有艺术感的线条图。
- 两分钟内完成高质量肖像,支持实时人机交互修改。
- 适合对交互式艺术创作感兴趣的设计师与开发者。
我们提出SketcherX,一种通过人机互动实现个性化肖像绘制的新型机器人系统。不同于依赖模拟打印的传统机器人艺术系统,SketcherX将人脸图像转化为具有独特人类艺术风格的矢量绘图。系统包含两个6轴机械臂:面部机器人配备头戴摄像头和大语言模型(LLM),实现实时交互;绘画机器人则采用微调后的Stable Diffusion模型、ControlNet及视觉-语言模型,动态生成风格化图像。主要贡献包括开发定制的向量低秩适配模型(LoRA),可无缝切换多种艺术风格,并采用成对微调方法提升笔触质量与风格准确性。实验表明,系统可在两分钟内生成高质量个性化肖像,展示了其在机器人创造性领域的潜力。该研究推动了机器人作为创作过程主动参与者的范式变革,为未来人机协同艺术合作开辟新路径。
原文摘要 · Abstract (English)
We introduce SketcherX, a novel robotic system for personalized portrait drawing through interactive human-robot engagement. Unlike traditional robotic art systems that rely on analog printing techniques, SketcherX captures and processes facial images to produce vectorized drawings in a distinctive, human-like artistic style. The system comprises two 6-axis robotic arms : a face robot, which is equipped with a head-mounted camera and Large Language Model (LLM) for real-time interaction, and a drawing robot, utilizing a fine-tuned Stable Diffusion model, ControlNet, and Vision-Language models for dynamic, stylized drawing. Our contributions include the development of a custom Vector Low Rank Adaptation model (LoRA), enabling seamless adaptation to various artistic styles, and integrating a pair-wise fine-tuning approach to enhance stroke quality and stylistic accuracy. Experimental results demonstrate the system's ability to produce high-quality, personalized portraits within two minutes, highlighting its potential as a new paradigm in robotic creativity. This work advances the field of robotic art by positioning robots as active participants in the creative process, paving the way for future explorations in interactive, human-robot artistic collaboration.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。