arXiv:2601.19430cs.CV2026-01被引 4

构建细粒度可解释AI生成图像检测基准,揭示现有模型依赖不可解释特征。

Unveiling Perceptual Artifacts: A Fine-Grained Benchmark for Interpretable AI-Generated Image Detection

  • 提出像素级标注的X-AIGD基准,涵盖低层失真、高层语义与认知反事实三类伪影
  • 实证发现现有检测器几乎不依赖感知伪影,即使在基础失真层面也如此
  • 显式对齐注意力与伪影区域能提升检测器可解释性与泛化能力

当前AI生成图像(AIGI)检测方法多依赖二分类区分真实与合成图像,却缺乏可解释或可信的决策依据。这一局限源于现有基准覆盖伪影类型不足且缺乏细粒度定位标注。为此,我们提出细粒度可解释AIGI检测基准X-AIGD,提供像素级、分类化的感知伪影标注,涵盖低层失真、高层语义与认知层反事实三类。该标注支持细粒度可解释性评估,深化对模型决策机制的理解。通过广泛实验,我们获得关键发现:(1) 现有检测器对感知伪影几乎无依赖,即使在最基础的失真层级亦然;(2) 尽管可训练检测特定伪影,其判断仍严重依赖不可解释特征;(3) 显式对齐模型注意力与伪影区域可显著提升检测器的可解释性与泛化能力。数据与代码已公开于https://github.com/Coxy7/X-AIGD。

原文摘要 · Abstract (English)

Current AI-Generated Image (AIGI) detection approaches predominantly rely on binary classification to distinguish real from synthetic images, often lacking interpretable or convincing evidence to substantiate their decisions. This limitation stems from existing AIGI detection benchmarks, which, despite featuring a broad collection of synthetic images, remain restricted in their coverage of artifact diversity and lack detailed, localized annotations. To bridge this gap, we introduce a fine-grained benchmark towards eXplainable AI-Generated image Detection, named X-AIGD, which provides pixel-level, categorized annotations of perceptual artifacts, spanning low-level distortions, high-level semantics, and cognitive-level counterfactuals. These comprehensive annotations facilitate fine-grained interpretability evaluation and deeper insight into model decision-making processes. Our extensive investigation using X-AIGD provides several key insights: (1) Existing AIGI detectors demonstrate negligible reliance on perceptual artifacts, even at the most basic distortion level. (2) While AIGI detectors can be trained to identify specific artifacts, they still substantially base their judgment on uninterpretable features. (3) Explicitly aligning model attention with artifact regions can increase the interpretability and generalization of detectors. The data and code are available at: https://github.com/Coxy7/X-AIGD.

可解释性图像检测伪影分析AI生成

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。