arXiv:2512.12090cs.CVcs.CR2025-12被引 2

通过选择性参数位移实现高效鲁棒的视频水印,兼顾隐蔽性与抗篡改能力。

SPDMark: Selective Parameter Displacement for Robust Video Watermarking

  • 在生成模型中仅修改部分参数嵌入水印,利用低秩适配实现高效控制。
  • 水印可高精度恢复,对剪辑、压缩等常见篡改保持鲁棒性。
  • 适合需要内容溯源的生成式视频应用,如版权保护与真实性验证。

高质量视频生成模型的兴起加剧了对可靠视频水印方案的需求,以追踪生成视频的来源。现有基于事后或生成过程中的水印方法难以同时满足隐蔽性、鲁棒性和计算效率。本文提出一种新型生成内视频水印框架SPDMark(发音为'SpeedMark'),基于视频扩散模型的选择性参数位移。通过修改生成模型中的一小部分参数来嵌入水印,将位移建模为层间基础偏移的加性组合,最终组合由水印密钥索引。为提升参数效率,采用低秩适配(LoRA)实现基础偏移。训练阶段联合优化基础偏移与水印提取器,最小化消息恢复、感知相似性和时间一致性损失。检测和定位时序篡改时,使用密码学哈希函数从基础密钥生成帧级水印信息。提取时采用最大二分匹配算法,即使在时间顺序被篡改的情况下仍能恢复正确帧序。在文本到视频和图像到视频生成模型上的评估表明,SPDMark能生成几乎不可察觉的水印,并实现高精度恢复,同时对多种常见视频操作具有强鲁棒性。

原文摘要 · Abstract (English)

The advent of high-quality video generation models has amplified the need for robust watermarking schemes that can be used to reliably detect and track the provenance of generated videos. Existing video watermarking methods based on both post-hoc and in-generation approaches fail to simultaneously achieve imperceptibility, robustness, and computational efficiency. This work introduces a novel framework for in-generation video watermarking called SPDMark (pronounced `SpeedMark') based on selective parameter displacement of a video diffusion model. Watermarks are embedded into the generated videos by modifying a subset of parameters in the generative model. To make the problem tractable, the displacement is modeled as an additive composition of layer-wise basis shifts, where the final composition is indexed by the watermarking key. For parameter efficiency, this work specifically leverages low-rank adaptation (LoRA) to implement the basis shifts. During the training phase, the basis shifts and the watermark extractor are jointly learned by minimizing a combination of message recovery, perceptual similarity, and temporal consistency losses. To detect and localize temporal modifications in the watermarked videos, we use a cryptographic hashing function to derive frame-specific watermark messages from the given base watermarking key. During watermark extraction, maximum bipartite matching is applied to recover the correct frame order, even from temporally tampered videos. Evaluations on both text-to-video and image-to-video generation models demonstrate the ability of SPDMark to generate imperceptible watermarks that can be recovered with high accuracy and also establish its robustness against a variety of common video modifications.

视频水印扩散模型生成内容溯源

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。