arXiv:2409.15762cs.CL2024-09被引 2

首个多语言可信度评测基准,揭示大模型在低资源语言中的短板

XTRUST: On the Multilingual Trustworthiness of Large Language Models

  • 构建覆盖10种语言的多语言可信度评测集
  • 发现模型在阿拉伯语、俄语等低资源语言上表现显著下降
  • 适合关注AI伦理与跨语言应用的研究者

大语言模型在自然语言处理任务中展现出卓越能力,但其可信度成为关键问题,尤其在医疗、金融等高风险领域。然而,现有研究多局限于单一语言,如英语。为应对全球部署需求,我们提出XTRUST——首个全面的多语言可信度基准,涵盖非法活动、幻觉、分布外鲁棒性、身心健康、毒性、公平性、虚假信息、隐私及机器伦理等10个主题,覆盖10种语言。基于XTRUST,我们对五款主流大模型进行跨语言实证评估,结果表明,多数模型在阿拉伯语、俄语等低资源语言上表现不佳,凸显当前多语言可信度仍有巨大提升空间。代码已开源:https://github.com/LluckyYH/XTRUST。

原文摘要 · Abstract (English)

Large language models (LLMs) have demonstrated remarkable capabilities across a range of natural language processing (NLP) tasks, capturing the attention of both practitioners and the broader public. A key question that now preoccupies the AI community concerns the capabilities and limitations of these models, with trustworthiness emerging as a central issue, particularly as LLMs are increasingly applied in sensitive fields like healthcare and finance, where errors can have serious consequences. However, most previous studies on the trustworthiness of LLMs have been limited to a single language, typically the predominant one in the dataset, such as English. In response to the growing global deployment of LLMs, we introduce XTRUST, the first comprehensive multilingual trustworthiness benchmark. XTRUST encompasses a diverse range of topics, including illegal activities, hallucination, out-of-distribution (OOD) robustness, physical and mental health, toxicity, fairness, misinformation, privacy, and machine ethics, across 10 different languages. Using XTRUST, we conduct an empirical evaluation of the multilingual trustworthiness of five widely used LLMs, offering an in-depth analysis of their performance across languages and tasks. Our results indicate that many LLMs struggle with certain low-resource languages, such as Arabic and Russian, highlighting the considerable room for improvement in the multilingual trustworthiness of current language models. The code is available at https://github.com/LluckyYH/XTRUST.

多语言可信度大模型评测基准

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。