提出按阶段评估医疗对话合规性,确保关键信息按序到位且可审计。
All Required, In Order: Phase-Level Evaluation for AI-Human Dialogue in Healthcare and Beyond
- 按对话阶段结构化检查临床义务是否按序完成
- 在呼吸病史和福利验证案例中实现可审计的证据链
- 适合医疗AI开发与临床合规审查人员使用
对话式AI正逐步支持临床工作,但现有评估方法忽略了合规性依赖于对话全过程。本文提出义务信息阶段结构化合规评估(OIP-SCE),通过检查每个必需的临床义务是否在正确阶段被满足,并提供清晰证据供临床医生审核,使复杂规则变得可实践、可审计。该方法在呼吸病史采集和福利验证两个案例中验证,将政策转化为可共享、可执行的步骤。通过赋予临床医生对检查内容的控制权,并为工程师提供明确规范,OIP-SCE构建了单一、可审计的评估界面,使AI能力与临床工作流对齐,支持安全、常规的应用。
原文摘要 · Abstract (English)
Conversational AI is starting to support real clinical work, but most evaluation methods miss how compliance depends on the full course of a conversation. We introduce Obligatory-Information Phase Structured Compliance Evaluation (OIP-SCE), an evaluation method that checks whether every required clinical obligation is met, in the right order, with clear evidence for clinicians to review. This makes complex rules practical and auditable, helping close the gap between technical progress and what healthcare actually needs. We demonstrate the method in two case studies (respiratory history, benefits verification) and show how phase-level evidence turns policy into shared, actionable steps. By giving clinicians control over what to check and engineers a clear specification to implement, OIP-SCE provides a single, auditable evaluation surface that aligns AI capability with clinical workflow and supports routine, safe use.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。