用隐式布局蒸馏生成更流畅透明的动态贴纸。
ILDiff: Generate Transparent Animated Stickers by Implicit Layout Distillation
- 通过隐式布局蒸馏建模动态贴纸的透明通道。
- 在0.32M高质量数据集上实现更细腻平滑的透明效果。
- 适合需要高精度透明动画的视觉创作人群。
高质量动态贴纸通常包含透明通道,但现有视频生成模型常忽略此特性。当前方法分为视频抠像与基于扩散模型两类:前者在贴纸半开放区域表现不佳,后者多用于单图建模,导致动态贴纸出现局部闪烁。本文提出ILDiff方法,通过隐式布局蒸馏生成动态透明通道,解决了半开放区域坍塌与缺乏时序信息的问题。同时构建了包含0.32M高质样本的透明动态贴纸数据集(TASD),为该领域提供数据支持。大量实验表明,ILDiff在透明通道的精细度与平滑性上优于Matting Anything和Layer Diffusion等方法。代码与数据集将公开于https://xiaoyuan1996.github.io。
原文摘要 · Abstract (English)
High-quality animated stickers usually contain transparent channels, which are often ignored by current video generation models. To generate fine-grained animated transparency channels, existing methods can be roughly divided into video matting algorithms and diffusion-based algorithms. The methods based on video matting have poor performance in dealing with semi-open areas in stickers, while diffusion-based methods are often used to model a single image, which will lead to local flicker when modeling animated stickers. In this paper, we firstly propose an ILDiff method to generate animated transparent channels through implicit layout distillation, which solves the problems of semi-open area collapse and no consideration of temporal information in existing methods. Secondly, we create the Transparent Animated Sticker Dataset (TASD), which contains 0.32M high-quality samples with transparent channel, to provide data support for related fields. Extensive experiments demonstrate that ILDiff can produce finer and smoother transparent channels compared to other methods such as Matting Anything and Layer Diffusion. Our code and dataset will be released at link https://xiaoyuan1996.github.io.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。