让AI实时理解草图的结构和意义,实现人机协同创作。
Real-Time Intuitive AI Drawing System for Collaboration: Enhancing Human Creativity through Formal and Contextual Intent Integration
- 双路解析:同时分析线条几何与语义主题,融合生成意图。
- 低延迟两阶段生成,支持多人同步在触屏上协作作画。
- 适合艺术创作、设计协作场景,尤其助力非专业用户表达创意。
本文提出一种实时生成式绘图系统,能够同时解析并融合形式意图(草图的结构、构图与风格特征)与上下文意图(基于视觉内容推断的语义与主题意义),形成统一的转换流程。与传统依赖文本提示的生成系统不同,本方法同时分析底层直观几何特征(如线条轨迹、比例、空间布局)与通过视觉-语言模型提取的高层语义线索。两种意图信号在多阶段生成流水线中联合调控,结合轮廓保持的结构控制与风格及内容感知的图像合成。系统采用触屏界面与分布式推理架构,实现低延迟的两阶段转换,并支持多人在共享画布上的同步协作。该平台使不同艺术水平的参与者均可参与同步共创,重新定义人机交互为协同创造与相互增强的过程。
原文摘要 · Abstract (English)
This paper presents a real-time generative drawing system that interprets and integrates both formal intent - the structural, compositional, and stylistic attributes of a sketch - and contextual intent - the semantic and thematic meaning inferred from its visual content - into a unified transformation process. Unlike conventional text-prompt-based generative systems, which primarily capture high-level contextual descriptions, our approach simultaneously analyzes ground-level intuitive geometric features such as line trajectories, proportions, and spatial arrangement, and high-level semantic cues extracted via vision-language models. These dual intent signals are jointly conditioned in a multi-stage generation pipeline that combines contour-preserving structural control with style- and content-aware image synthesis. Implemented with a touchscreen-based interface and distributed inference architecture, the system achieves low-latency, two-stage transformation while supporting multi-user collaboration on shared canvases. The resulting platform enables participants, regardless of artistic expertise, to engage in synchronous, co-authored visual creation, redefining human-AI interaction as a process of co-creation and mutual enhancement.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。