arXiv:2412.08274cs.CL2024-12被引 3

构建首个多语种语音与美式手语理解数据集,支持跨语言评估。

2M-BELEBELE: Highly Multilingual Speech and American Sign Language Comprehension Dataset

  • 扩展BELEBELE构建74种语言语音+美式手语数据集
  • 语音理解准确率比阅读低2-3%,零样本与少样本均有效
  • 适合多语言语音识别与手语理解研究者使用

我们提出首个高度多语言语音与美国手语(ASL)理解数据集2M-BELEBELE,通过扩展BELEBELE构建。该数据集覆盖74种口语语言(位于BELEBELE与FLEURS交集),以及一种手语(ASL)。我们在5-shot和zero-shot设置下对2M-BELEBELE进行评估,结果显示语音理解准确率相比阅读理解平均低2-3%。该数据集为跨语言语音与手语理解提供了基准支持。

原文摘要 · Abstract (English)

We introduce the first highly multilingual speech and American Sign Language (ASL) comprehension dataset by extending BELEBELE. Our dataset covers 74 spoken languages at the intersection of BELEBELE and FLEURS, and one sign language (ASL). We evaluate 2M-BELEBELE dataset for both 5-shot and zero-shot settings and across languages, the speech comprehension accuracy is ~ 2-3% average lower compared to reading comprehension.

多语言语音理解手语数据集

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。