arXiv:2608.09443cs.AI2026-08

用图与策略蒸馏实现老年人多病用药安全个性化推荐

Coupled Graph--Policy Distillation for Personalized Medication Safety in Older Adults with Multimorbidity

论文配图:Coupled Graph--Policy Distillation for Personalized Medication Safety in Older Adults with Multimorbidity
图 1 · 摘自论文原文
  • 构建药物安全图,动态生成患者专属冲突图
  • 在欧洲和亚洲多个基准上超越现有系统,无误推荐
  • 适合临床辅助决策,尤其多病共存老年患者

大型语言模型(LLM)代理可辅助临床间药物审查,但老年多病患者的用药安全依赖于疾病、用药及老年风险因素,用户常遗漏关键信息。我们提出ATLAS,一种耦合图-策略蒸馏框架,实现患者自适应的药物安全评估。ATLAS将指南证据结构化为药物安全图,通过针对性提问更新患者状态,并蒸馏出患者特定的药物冲突图(PMCG)。一个以风险优先的多代理策略利用PMCG筛查禁忌、评估警示与监测需求、识别更安全替代方案,并验证最终用药计划。我们还引入GeriMedBench,一个测试安全关键信息获取与基于证据决策修正的交互式基准。在欧洲非交互式多病基准、亚洲交互式多病基准及亚洲非交互式跨指南基准上,ATLAS在完整决策性能上优于对比系统。在欧洲非交互式多病基准中,其严格成功率达53.73分高于最强专有LLM基线,整体安全推理得分高出14.63分,自动化评估下无误推荐。盲评临床医生评价显示,ATLAS在五项指标上均获更高均分,仅在一处案例中发现潜在不安全建议,而Gemini有两处。

原文摘要 · Abstract (English)

Large language model (LLM) agents can support medication review between clinical visits, but safe choices for older adults with multimorbidity depend on conditions, medications, and geriatric risks that users may omit. We introduce ATLAS, a coupled graph--policy distillation framework for patient-adaptive medication safety. ATLAS structures guideline evidence as a medication-safety graph. Targeted questions update the patient state and distill relevant relations into a patient-specific medication conflict graph (PMCG). A risk-first multi-agent policy uses the PMCG to screen contraindications, assess cautions and monitoring needs, identify safer alternatives, and verify the final medication plan. We also introduce GeriMedBench, an interactive benchmark that tests safety-critical information acquisition and evidence-based decision revision. Across a European non-interactive multimorbidity benchmark, an Asian interactive multimorbidity benchmark, and an Asian non-interactive cross-guideline benchmark, ATLAS achieves the strongest complete-decision performance among the compared systems. On the European non-interactive multimorbidity benchmark, it exceeds the strongest proprietary LLM baseline by 53.73 points in Strict Success Rate and 14.63 points in overall safety reasoning score (OSRS), with no unsafe recommendations under the automated evaluator. A blinded clinician evaluation gives ATLAS higher mean ratings across all five criteria and flags potentially unsafe recommendations in one ATLAS case and two Gemini cases.

药物安全多病共存智能诊疗图神经网络

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。