arXiv:2502.10973cs.CL2025-02ACL被引 6

首个非洲语言情感对话数据集,支持多模态情感识别研究

Akan Cinematic Emotions (ACE): A Multimodal Multi-party Dataset for Emotion Recognition in Movie Dialogues

  • 构建Akan语多模态对话数据集,含音频、视频、文本三模态
  • 包含385段带情感标签的对话,6162个语句,含语音重音标注
  • 推动低资源语言情感识别,助力多元文化NLP发展

本文提出Akan Conversation Emotion(ACE)数据集,首个针对非洲语言的多模态情感对话数据集,弥补了低资源语言在情感识别研究中的空白。该数据集基于阿坎语,包含385段带情感标签的对话、6,162个语句,涵盖音频、视觉与文本模态,并提供词级语音重音标注,是首个标注语音重音的非洲语言数据集。通过使用先进情感识别方法进行实验,验证了其质量与实用性,建立了未来研究的基准。期望ACE能推动更具包容性、语言与文化多样性的自然语言处理资源发展。

原文摘要 · Abstract (English)

In this paper, we introduce the Akan Conversation Emotion (ACE) dataset, the first multimodal emotion dialogue dataset for an African language, addressing the significant lack of resources for low-resource languages in emotion recognition research. ACE, developed for the Akan language, contains 385 emotion-labeled dialogues and 6,162 utterances across audio, visual, and textual modalities, along with word-level prosodic prominence annotations. The presence of prosodic labels in this dataset also makes it the first prosodically annotated African language dataset. We demonstrate the quality and utility of ACE through experiments using state-of-the-art emotion recognition methods, establishing solid baselines for future research. We hope ACE inspires further work on inclusive, linguistically and culturally diverse NLP resources.

情感识别多模态低资源语言非洲语言

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。