arXiv:2604.08226cs.AIcs.HC2026-04被引 1

为临床AI建立可评估的智能框架,明确其在不同场景下的可靠性边界。

Grounding Clinical AI Competency in Human Cognition Through the Clinical World Model and Skill-Mix Framework

  • 构建患者-医生-生态三元交互的临床世界模型,统一认知基础。
  • 提出八维技能矩阵,覆盖疾病、阶段、角色等维度,生成百亿级能力坐标。
  • 强调每个坐标需独立验证,避免泛化误判,适合医疗AI研发与监管者使用。

任何智能体的能力都受限于其对运行环境的正式认知。当前临床AI缺乏这样的认知框架。现有方法仅孤立地关注评估、监管或系统设计,缺乏连接各方的共同临床世界模型。本文提出临床世界模型,将医疗照护形式化为患者、提供者与生态系统之间的三元互动。为形式化人类或人工智能如何将信息转化为临床行动,我们基于经验证的临床认知原则,构建了三类并行决策架构。临床AI技能混合框架通过八个维度定义能力:五个描述临床能力空间(疾病、阶段、照护场景、提供者角色、任务),三个刻画AI如何介入人类推理(授权范围、面向对象、锚定层级)。这些维度的组合产生数十亿个独特的能力坐标。核心结构启示是:单一坐标的验证对其他坐标性能几乎无参考价值,因此该能力空间不可约简。该框架为临床AI提供了跨利益相关方的共同语言,用于指定、评估与界定其能力边界。通过显式呈现这一结构,本框架将领域核心问题从‘AI是否有效’转变为‘在哪些能力坐标上已证实可靠,对谁可靠’。

原文摘要 · Abstract (English)

The competency of any intelligent agent is bounded by its formal account of the world in which it operates. Clinical AI lacks such an account. Existing frameworks address evaluation, regulation, or system design in isolation, without a shared model of the clinical world to connect them. We introduce the Clinical World Model, a framework that formalizes care as a tripartite interaction among Patient, Provider, and Ecosystem. To formalize how any agent, whether human or artificial, transforms information into clinical action, we develop parallel decision-making architectures for providers, patients, and AI agents, grounded in validated principles of clinical cognition. The Clinical AI Skill-Mix operationalizes competency through eight dimensions. Five define the clinical competency space (condition, phase, care setting, provider role, and task) and three specify how AI engages human reasoning (assigned authority, agent facing, and anchoring layer). The combinatorial product of these dimensions yields a space of billions of distinct competency coordinates. A central structural implication is that validation within one coordinate provides minimal evidence for performance in another, rendering the competency space irreducible. The framework supplies a common grammar through which clinical AI can be specified, evaluated, and bounded across stakeholders. By making this structure explicit, the Clinical World Model reframes the field's central question from whether AI works to in which competency coordinates reliability has been demonstrated, and for whom.

临床AI认知建模能力评估

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。