无需标注、掩码或训练,仅靠语言指令就能精准编辑图像。
Hands-off Image Editing: Language-guided Editing without any Task-specific Labeling, Masking or even Training
- 纯语言指令驱动,无需任何任务特定标注或训练
- 在多个编辑任务上表现接近现有顶尖方法
- 适合希望快速部署且无标注数据的开发者使用
指令引导的图像编辑旨在根据给定图像和指令生成符合要求的修改结果。当前最先进的方法通常依赖于特定任务的标注、掩码或训练,面临可扩展性差和领域迁移困难的问题。本文提出一种新方法,完全摆脱任务特定监督,不需标注、掩码或训练。实验表明,该方法性能优异,达到与现有最优方法相当的效果,在多种编辑任务中展现出强大潜力。
原文摘要 · Abstract (English)
Instruction-guided image editing consists in taking an image and an instruction and deliverring that image altered according to that instruction. State-of-the-art approaches to this task suffer from the typical scaling up and domain adaptation hindrances related to supervision as they eventually resort to some kind of task-specific labelling, masking or training. We propose a novel approach that does without any such task-specific supervision and offers thus a better potential for improvement. Its assessment demonstrates that it is highly effective, achieving very competitive performance.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。