arXiv:2606.02402cs.CV2026-06中稿 · ICML

定位长视频中被篡改的片段并给出可解释原因

Explainable Forensics of Manipulated Segments in Untrimmed Long Videos

论文配图:Explainable Forensics of Manipulated Segments in Untrimmed Long Videos
图 1 · 摘自论文原文
  • 从粗到细检测长视频中的伪造片段边界
  • 在1.2万+视频上验证了定位准确率提升
  • 适合关注AI视频造假分析与可解释性的研究者

AI生成视频技术迅猛发展,带来内容创作便利的同时,也加剧了在长视频中局部篡改引发的虚假信息风险。现有方法多针对短独立片段,难以应对真实场景下伪造内容嵌入于大量真实画面中的情况。为此,本文提出时序性AI生成片段定位与解释任务,旨在实现真实性判断、时间定位和可解释分析。构建了TASLE基准数据集,包含12,472个未剪辑长视频,涵盖多样篡改模式及丰富标注信号,包括时间边界、真伪标签与片段级推理依据。提出MSLoc基线模型,结合边界敏感的提案生成模块实现高效长视频扫描,并通过多模态大模型(MLLM)精修模块实现精准边界定位与可解释推理。实验验证了该方法的有效性,强调了片段级可解释取证对长视频AI生成内容分析的重要性。数据集与代码已开源。

原文摘要 · Abstract (English)

The rapid advancement of AI-driven video generation has transformed content creation, while simultaneously increasing the risk of misinformation through localized manipulations in long-form videos. Existing video forensic methods predominantly operate on short, independent clips, and thus fail to capture realistic scenarios where AI-generated content is sparsely embedded within otherwise authentic footage. To bridge this gap, we formulate the task of Temporal AI-Generated Segment Localization and Explanation, which targets authenticity detection, temporal localization, and interpretable analysis of manipulated segments in untrimmed long videos. We further introduce TASLE, a large-scale benchmark comprising 12,472 untrimmed videos with diverse manipulation patterns and rich annotation signals, including temporal boundaries, authenticity labels, and segment-level rationales. In addition, we propose MSLoc, a coarse-to-fine forensic baseline that combines a boundary-sensitive proposal generation module for efficient long-video scanning with an MLLM-based refinement module for precise boundary localization and interpretable reasoning. Experiments validate the effectiveness of the proposed baseline, highlighting the importance of segment-level explainable forensics for long-form AI-generated video analysis. Our dataset and code are publicly available at https://debby-0527.github.io/TASLE.

视频取证可解释性长视频伪造检测

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。