构建196名跨男性群体的多模态语料库,支持声音健康研究与语音分析。
TMASC: Transmasculine Attitude and Speech Corpus

- 采集问卷与66段音频,涵盖咳嗽、朗读等语音样本
- 通过多源数据发现群体级语音特征与声学规律
- 适合性别研究、语音医疗与语音合成领域学者使用
我们提出了跨男性态度与语音语料库(TMASC),包含196名跨男性个体的多模态数据,包括问卷回答和66段音频记录。问卷涉及跨男性群体的声音健康状况,音频记录包含咳嗽与清嗓样本、朗读片段及特定会话问题。本文详述了该语料库的构建过程与数据采集流程。为展示其应用价值,我们呈现三个案例研究:整合感知与声学数据、识别群体层面特征、校准声学测量方法,证明该众包多模态语料库在支持跨男性群体研究中的实用性。
原文摘要 · Abstract (English)
We introduce the Transmasculine Attitudes and Speech Corpus (TMASC), a multimodal corpus of 196 transmasculine individuals, including questionnaire responses and 66 audio recordings. The questionnaire includes items exploring the vocal health of transmasculine individuals. The audio recordings include cough and throat-clearing samples, a reading passage, and additional session-specific questions. This paper outlines the development of this corpus and the data collection procedures. To illustrate the utility of this corpus, we present three case studies demonstrating how this crowd-sourced multimodal corpus can be used to support transmasculine individuals. These include the integration of perceptual and acoustic data, the identification of group-level characteristics, and the calibration of acoustic measurements.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。