用扩散模型同时做异常检测与生成,打破传统分离思路。
Anomaly Detection and Generation with Diffusion Models: A Survey
- 将异常检测与生成结合,形成互相增强的闭环
- 利用生成弥补异常数据稀缺,提升检测效果
- 覆盖图像、时序、多模态等多类数据,适合跨领域研究
异常检测(AD)在网络安全、金融、医疗和工业制造等领域至关重要,用于识别偏离正常模式的异常行为。近年来,扩散模型(DMs)因其能学习复杂数据分布并生成高质量样本,成为无监督异常检测的有力工具。本文全面综述了基于扩散模型的异常检测与生成(ADGDM),从理论基础到实际应用,涵盖图像、视频、时间序列、表格数据及多模态数据。与以往将检测与生成分开讨论的综述不同,本文强调二者内在协同关系:生成可缓解异常样本稀缺问题,而检测则提供反馈以提升生成质量,实现双向增强。我们构建了基于异常评分机制、条件策略和架构设计的详细分类体系,分析各类方法的优劣。最后讨论了可扩展性与计算效率等关键挑战,并展望了高效架构、新型条件策略及与基础模型(如视觉-语言模型、大语言模型)融合等未来方向。本综述旨在推动研究人员在多元场景中创新应用扩散模型解决异常检测问题。
原文摘要 · Abstract (English)
Anomaly detection (AD) plays a pivotal role across diverse domains, including cybersecurity, finance, healthcare, and industrial manufacturing, by identifying unexpected patterns that deviate from established norms in real-world data. Recent advancements in deep learning, specifically diffusion models (DMs), have sparked significant interest due to their ability to learn complex data distributions and generate high-fidelity samples, offering a robust framework for unsupervised AD. In this survey, we comprehensively review anomaly detection and generation with diffusion models (ADGDM), presenting a tutorial-style analysis of the theoretical foundations and practical implementations and spanning images, videos, time series, tabular, and multimodal data. Crucially, unlike existing surveys that often treat anomaly detection and generation as separate problems, we highlight their inherent synergistic relationship. We reveal how DMs enable a reinforcing cycle where generation techniques directly address the fundamental challenge of anomaly data scarcity, while detection methods provide critical feedback to improve generation fidelity and relevance, advancing both capabilities beyond their individual potential. A detailed taxonomy categorizes ADGDM methods based on anomaly scoring mechanisms, conditioning strategies, and architectural designs, analyzing their strengths and limitations. We final discuss key challenges including scalability and computational efficiency, and outline promising future directions such as efficient architectures, conditioning strategies, and integration with foundation models (e.g., visual-language models and large language models). By synthesizing recent advances and outlining open research questions, this survey aims to guide researchers and practitioners in leveraging DMs for innovative AD solutions across diverse applications.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。