arXiv:2604.03448cs.CVcs.AI2026-04中稿 · CVPR

用扩散模型实现3秒内精准编辑角色表情,无噪点且兼容Photoshop

ExpressEdit: Fast Editing of Stylized Facial Expressions with Diffusion Models in Photoshop

  • 基于扩散模型的插件,3秒完成表情编辑
  • 无全局噪声和像素漂移,保持图像质量
  • 适配专业美术师,支持故事化表达生成

角色面部表情是视觉叙事的核心。当前AI图像编辑模型虽能辅助艺术创作,但常引入全局噪声和像素漂移,难以融入专业软件流程。为此,我们提出ExpressEdit——一个完全开源的Photoshop插件,彻底避免主流模型的常见缺陷,且与Liquify等原生操作无缝协同。该插件在单块消费级GPU上3秒内完成表情编辑,显著快于主流专有模型。为支持多样叙事需求,我们构建了包含135个表情标签的综合性表达数据库,每条均附带示例故事与图像,用于检索增强生成。代码与数据集均已开源,以推动未来研究与艺术探索。

原文摘要 · Abstract (English)

Facial expressions of characters are a vital component of visual storytelling. While current AI image editing models hold promise for assisting artists in the task of stylized expression editing, these models introduce global noise and pixel drift into the edited image, preventing the integration of these models into professional image editing software and workflows. To bridge this gap, we introduce ExpressEdit, a fully open-source Photoshop plugin that is free from common artifacts of proprietary image editing models and robustly synergizes with native Photoshop operations such as Liquify. ExpressEdit seamlessly edits an expression within 3 seconds on a single consumer-grade GPU, significantly faster than popular proprietary models. Moreover, to support the generation of diverse expressions according to different narrative needs, we compile a comprehensive expression database of 135 expression tags enriched with example stories and images designed for retrieval-augmented generation. We open source the code and dataset to facilitate future research and artistic exploration.

表情编辑扩散模型Photoshop插件实时生成

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。