arXiv:2505.09747cs.CYcs.AI2025-05被引 1

提出'健康不信任'概念,强调合理质疑是建立真正可信AI的关键。

Healthy Distrust in AI systems

  • 将'健康不信任'定义为对可能损害个人利益的AI使用应持有的理性审慎态度
  • 指出盲目信任AI可能削弱人的自主性,合理质疑反而是信任的前提
  • 适用于政策制定者、伦理审查员及关注算法公平性的研究者

在'可信赖AI'的口号下,当前多数AI研究聚焦于设计能激发人类信任的系统与使用方式,从而促进采纳。然而,当个体受到AI系统影响时,仅靠系统设计未必能赢得其信任——尤其当系统嵌入的社会环境中存在与个人利益相冲突的机制时,这种不信任是正当且必要的,甚至有助于构建真正的信任。本文提出'健康不信任'这一概念,用以描述在特定情境下对AI使用实践所持的合理、审慎立场。通过综合计算机科学、社会学、历史、心理学与哲学中的信任与不信任理论,本文指出当前研究中尚未填补的空白,并将健康不信任构想为尊重人类自主性的关键组成部分。

原文摘要 · Abstract (English)

Under the slogan of trustworthy AI, much of contemporary AI research is focused on designing AI systems and usage practices that inspire human trust and, thus, enhance adoption of AI systems. However, a person affected by an AI system may not be convinced by AI system design alone -- neither should they, if the AI system is embedded in a social context that gives good reason to believe that it is used in tension with a person's interest. In such cases, distrust in the system may be justified and necessary to build meaningful trust in the first place. We propose the term "healthy distrust" to describe such a justified, careful stance towards certain AI usage practices. We investigate prior notions of trust and distrust in computer science, sociology, history, psychology, and philosophy, outline a remaining gap that healthy distrust might fill and conceptualize healthy distrust as a crucial part for AI usage that respects human autonomy.

AI伦理信任机制人类自主性

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。