UniMark统一识别AI生成内容,支持隐写与显式标记。
UniMark: Artificial Intelligence Generated Content Identification Toolkit
- 模块化引擎统一处理文本、图像、音频、视频多模态内容
- 首创隐写+显式标记双策略,兼顾版权保护与合规要求
- 提供图像/视频/音频三类标准评测基准,评估更规范
人工智能生成内容的迅速泛滥引发了信任危机并催生了迫切的监管需求。现有识别工具存在碎片化问题,且缺乏对可见合规标记的支持。为此,我们提出开源统一框架 UniMark,具备跨文本、图像、音频、视频模态的抽象能力。关键创新在于提出一种新型双操作策略,原生支持隐写标记(用于版权保护)和显式标记(用于监管合规)。此外,我们构建了标准化评估体系,包含图像、视频、音频三个专用评测基准,确保性能评估严谨可靠。该工具架起了先进算法与工程实现之间的桥梁,推动更透明、安全的数字生态建设。
原文摘要 · Abstract (English)
The rapid proliferation of Artificial Intelligence Generated Content has precipitated a crisis of trust and urgent regulatory demands. However, existing identification tools suffer from fragmentation and a lack of support for visible compliance marking. To address these gaps, we introduce the \textbf{UniMark}, an open-source, unified framework for multimodal content governance. Our system features a modular unified engine that abstracts complexities across text, image, audio, and video modalities. Crucially, we propose a novel dual-operation strategy, natively supporting both \emph{Hidden Watermarking} for copyright protection and \emph{Visible Marking} for regulatory compliance. Furthermore, we establish a standardized evaluation framework with three specialized benchmarks (Image/Video/Audio-Bench) to ensure rigorous performance assessment. This toolkit bridges the gap between advanced algorithms and engineering implementation, fostering a more transparent and secure digital ecosystem.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。