arXiv:2507.12870cs.SDcs.CY2025-07中稿 · the 10th Workshop …被引 1

为儿童语音语料库建设提供跨教育、临床与法医场景的实操指南。

Best Practices and Considerations for Child Speech Corpus Collection and Curation in Educational, Clinical, and Forensic Scenarios

  • 基于WHO、WHAT、WHEN、WHERE框架设计采集流程。
  • 强调隐私保护与伦理审查,确保数据合规可用。
  • 适合从事儿童语音研究的学者与临床/法医从业者参考。

儿童口语能力持续发展至成年期,7-8岁前发音与语言结构变化迅速。这种动态演变及数据隐私问题,使构建可用于技术开发的儿童语音语料库极具挑战。本研究旨在弥合此差距,为研究人员和实践者提供以目标为导向的语料库建设最佳实践。虽主要面向教育场景,但儿童语音数据的应用已扩展至临床与法医领域。基于前期经验与现有实践,本文提出收集语料的四维框架(谁、什么、何时、何地),并提供建立合作、信任及通过人类受试者研究协议的指导。最后,总结语料质量检查、筛选与标注的规范流程。

原文摘要 · Abstract (English)

A child's spoken ability continues to change until their adult age. Until 7-8yrs, their speech sound development and language structure evolve rapidly. This dynamic shift in their spoken communication skills and data privacy make it challenging to curate technology-ready speech corpora for children. This study aims to bridge this gap and provide researchers and practitioners with the best practices and considerations for developing such a corpus based on an intended goal. Although primarily focused on educational goals, applications of child speech data have spread across fields including clinical and forensics fields. Motivated by this goal, we describe the WHO, WHAT, WHEN, and WHERE of data collection inspired by prior collection efforts and our experience/knowledge. We also provide a guide to establish collaboration, trust, and for navigating the human subjects research protocol. This study concludes with guidelines for corpus quality check, triage, and annotation.

儿童语音语料库伦理规范多场景应用

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。