arXiv:2503.10970cs.AIcs.LG2025-03被引 82

TxAgent用211个工具实现精准用药推理,提升个性化治疗决策质量。

TxAgent: An AI Agent for Therapeutic Reasoning Across a Universe of Tools

  • 基于多步推理与实时生物医学知识检索,动态调用211个工具分析药物交互。
  • 在3,168个药物任务中达92.1%准确率,超越GPT-4o和DeepSeek-R1(671B)。
  • 适合临床医生、药学研究者及AI医疗系统开发者使用。

精准治疗需要多模态自适应模型生成个性化治疗建议。我们提出TxAgent,一个能跨211个工具进行多步推理与实时生物医学知识检索的AI代理,用于分析药物相互作用、禁忌症及患者特异性治疗策略。TxAgent在分子、药代动力学和临床层面评估药物交互,依据患者共病与合并用药识别禁忌症,并根据个体特征定制治疗方案。它从多个生物医学来源检索并整合证据,评估药物与患者状况间的相互作用,并通过迭代推理优化推荐结果。根据任务目标选择工具,执行结构化函数调用以解决需临床推理与跨源验证的治疗任务。ToolUniverse整合了211个可信来源工具,涵盖自1939年以来所有美国FDA批准药物及Open Targets的经验证临床洞察。TxAgent在五个新基准(DrugPC、BrandPC、GenericPC、TreatmentPC、DescriptionPC)上表现优于领先大模型、工具使用模型和推理代理,覆盖3,168个药物推理任务与456个个性化治疗场景。其在开放式药物推理任务中达到92.1%准确率,超越GPT-4o,且在结构化多步推理中优于DeepSeek-R1(671B)。TxAgent能泛化处理药物名称变体与描述差异。通过整合多步推断、实时知识锚定与工具辅助决策,确保治疗建议符合临床指南与真实世界证据,降低不良事件风险,改善治疗决策。

原文摘要 · Abstract (English)

Precision therapeutics require multimodal adaptive models that generate personalized treatment recommendations. We introduce TxAgent, an AI agent that leverages multi-step reasoning and real-time biomedical knowledge retrieval across a toolbox of 211 tools to analyze drug interactions, contraindications, and patient-specific treatment strategies. TxAgent evaluates how drugs interact at molecular, pharmacokinetic, and clinical levels, identifies contraindications based on patient comorbidities and concurrent medications, and tailors treatment strategies to individual patient characteristics. It retrieves and synthesizes evidence from multiple biomedical sources, assesses interactions between drugs and patient conditions, and refines treatment recommendations through iterative reasoning. It selects tools based on task objectives and executes structured function calls to solve therapeutic tasks that require clinical reasoning and cross-source validation. The ToolUniverse consolidates 211 tools from trusted sources, including all US FDA-approved drugs since 1939 and validated clinical insights from Open Targets. TxAgent outperforms leading LLMs, tool-use models, and reasoning agents across five new benchmarks: DrugPC, BrandPC, GenericPC, TreatmentPC, and DescriptionPC, covering 3,168 drug reasoning tasks and 456 personalized treatment scenarios. It achieves 92.1% accuracy in open-ended drug reasoning tasks, surpassing GPT-4o and outperforming DeepSeek-R1 (671B) in structured multi-step reasoning. TxAgent generalizes across drug name variants and descriptions. By integrating multi-step inference, real-time knowledge grounding, and tool-assisted decision-making, TxAgent ensures that treatment recommendations align with established clinical guidelines and real-world evidence, reducing the risk of adverse events and improving therapeutic decision-making.

AI医疗药物推理智能诊疗工具调用

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。