arXiv:2605.03547cs.CVcs.AI2026-05中稿 · LREC 2026

首个评估视觉语言模型版权内容遗忘的基准,解决模型侵权风险。

Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models

论文配图:Erase Persona, Forget Lore: Benchmarking Multimodal Copyright Unlearning in Large Vision Language Models
图 1 · 摘自论文原文
  • 用程序生成的合成数据构建多模态版权内容遗忘测试集
  • 验证了遗忘效果与通用能力之间的权衡关系
  • 适合关注AI版权合规与模型安全的研究者

大规模视觉语言模型(LVLM)在网页规模数据上训练,可能记忆并复现受版权保护的视觉内容(如角色、徽标),带来显著风险。机器遗忘提供了一种训练后移除特定内容的路径,但其在复杂多模态设置下的有效性评估仍是开放问题。现有方法往往缺乏鲁棒性或无法捕捉跨模态概念遗忘的细微差别。为此,我们提出CoVUBench基准,首个专为评估LVLM中版权内容遗忘而设计的框架。该基准采用程序生成的合法合成数据,结合构图变化和多样域表现的系统性视觉变异,确保评估的真实性和鲁棒性。全面的多模态评估协议从版权持有方视角衡量遗忘有效性,从部署方视角评估通用能力保留情况。通过严格测量这一关键权衡,CoVUBench提供了标准化工具,推动负责任且有效的遗忘方法发展。

原文摘要 · Abstract (English)

Large Vision-Language Models (LVLMs), trained on web-scale data, risk memorizing and regenerating copyrighted visual content such as characters and logos, creating significant challenges. Machine unlearning offers a path to mitigate these risks by removing specific content post-training, but evaluating its effectiveness, especially in the complex multimodal setting of LVLMs, remains an open problem. Current evaluation methods often lack robustness or fail to capture the nuances of cross-modal concept erasure. To address this critical gap, we introduce the CoVUBench benchmark, the first framework specifically designed for evaluating copyright content unlearning in LVLMs. CoVUBench utilizes procedurally generated, legally safe synthetic data coupled with systematic visual variations spanning compositional changes and diverse domain manifestations to ensure realistic and robust evaluation of unlearning generalization. Our comprehensive multimodal evaluation protocol assesses both forgetting efficacy from the copyright holder perspective and the preservation of general model utility from the deployer viewpoint. By rigorously measuring this crucial trade-off, CoVUBench provides a standardized tool to advance the development of responsible and effective unlearning methods for LVLMs.

视觉语言模型版权保护机器遗忘多模态

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。