arXiv:2410.14975cs.CVcs.AI2024-10中稿 · ICLR被引 4

通过自生成概念建议提升视觉语言模型的分布外检测能力

Reflexive Guidance: Improving OoDD in Vision-Language Models via Self-Guided Image-Adaptive Concept Generation

  • 提出自引导提示方法,利用模型自身生成适配图像的概念
  • 实验显示该方法显著提升视觉语言模型在分类与分布外检测上的表现
  • 适合关注大模型可信性与安全部署的研究者与开发者

随着互联网规模数据训练的基础模型展现出卓越泛化能力,其应用范围不断扩展。然而,这些模型的可信度仍待深入研究,尤其是大视觉语言模型(LVLMs)在分布外检测(OoDD)方面的能力尚未充分探索。我们评估并分析了多种专有及开源LVLMs的OoDD性能,揭示了模型如何通过自然语言响应表达置信度。为此,我们提出一种自引导提示方法——反射式引导(ReGuide),通过自生成的图像适应性概念建议来增强模型的OoDD能力。实验结果表明,ReGuide能有效提升当前主流LVLM在图像分类与分布外检测任务中的表现。样本图像、提示与响应数据已公开于https://github.com/daintlab/ReGuide。

原文摘要 · Abstract (English)

With the recent emergence of foundation models trained on internet-scale data and demonstrating remarkable generalization capabilities, such foundation models have become more widely adopted, leading to an expanding range of application domains. Despite this rapid proliferation, the trustworthiness of foundation models remains underexplored. Specifically, the out-of-distribution detection (OoDD) capabilities of large vision-language models (LVLMs), such as GPT-4o, which are trained on massive multi-modal data, have not been sufficiently addressed. The disparity between their demonstrated potential and practical reliability raises concerns regarding the safe and trustworthy deployment of foundation models. To address this gap, we evaluate and analyze the OoDD capabilities of various proprietary and open-source LVLMs. Our investigation contributes to a better understanding of how these foundation models represent confidence scores through their generated natural language responses. Furthermore, we propose a self-guided prompting approach, termed Reflexive Guidance (ReGuide), aimed at enhancing the OoDD capability of LVLMs by leveraging self-generated image-adaptive concept suggestions. Experimental results demonstrate that our ReGuide enhances the performance of current LVLMs in both image classification and OoDD tasks. The lists of sampled images, along with the prompts and responses for each sample are available at https://github.com/daintlab/ReGuide.

视觉语言模型分布外检测可信人工智能自引导

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。