arXiv:2508.20047cs.CL2025-08被引 7

首个阿拉伯语医疗问答共享任务,推动多领域健康信息获取

AraHealthQA 2025: The First Shared Task on Arabic Health Question Answering

  • 分精神健康与综合医学两赛道,覆盖焦虑抑郁等议题
  • 构建双赛道评估体系,含真实场景数据与标准化指标
  • 面向多语言文化背景,助力阿拉伯语医疗AI落地

我们推出AraHealthQA 2025——首个全面的阿拉伯语医疗问答共享任务,于阿拉伯语自然语言处理会议(ArabicNLP 2025,与EMNLP 2025同期)举办。该任务旨在填补高质量阿拉伯语医学问答资源匮乏的空白,设立两个互补赛道:MentalQA聚焦精神健康问答(如焦虑、抑郁、污名化消除);MedArabiQ涵盖内科、儿科及临床决策等更广泛的医学领域。每个赛道包含多个子任务、评测数据集与标准化评估指标,支持公平基准测试。任务设计强调在真实、多语言且文化敏感的医疗情境下进行模型开发。本文详述数据集构建、任务设计与评估框架、参赛情况、基线系统,并总结整体成果。最后,分析性能趋势,展望未来阿拉伯语医疗问答的发展前景。

原文摘要 · Abstract (English)

We introduce AraHealthQA 2025, the Comprehensive Arabic Health Question Answering Shared Task, held in conjunction with ArabicNLP 2025 (co-located with EMNLP 2025). This shared task addresses the paucity of high-quality Arabic medical QA resources by offering two complementary tracks: MentalQA, focusing on Arabic mental health Q&A (e.g., anxiety, depression, stigma reduction), and MedArabiQ, covering broader medical domains such as internal medicine, pediatrics, and clinical decision making. Each track comprises multiple subtasks, evaluation datasets, and standardized metrics, facilitating fair benchmarking. The task was structured to promote modeling under realistic, multilingual, and culturally nuanced healthcare contexts. We outline the dataset creation, task design and evaluation framework, participation statistics, baseline systems, and summarize the overall outcomes. We conclude with reflections on the performance trends observed and prospects for future iterations in Arabic health QA.

医疗问答阿拉伯语共享任务

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。