构建车内语音技术研究数据库,支持静止与移动场景下的声学分析。
CAVEMOVE: An Acoustic Database for the Study of Voice-enabled Technologies inside Moving Vehicles
- 采集车辆静止与运动时的语音与噪声数据,含两种麦克风配置。
- 提供脉冲响应与多工况噪声数据,支持语音与车载音频建模。
- 开源配套Python API,助力语音交互系统开发与测试。
本文介绍一个声学数据库,旨在推动车载语音技术的研究。数据采集包含两部分:(i) 静态条件下获取的声学脉冲响应,用于建模语音与车载音频成分;(ii) 静态及动态行驶条件下的噪声记录。采用两种麦克风配置:紧凑型麦克风阵列和分布式麦克风布置。简要说明了录音环境设置,并提供一个专为车载语音技术研究与开发设计的Python API。首个版本的API及部分数据可免费下载使用。
原文摘要 · Abstract (English)
In this paper, we present an acoustic database, designed to drive and support research on voiced enabled technologies inside moving vehicles. The recording process involves (i) recordings of acoustic impulse responses, acquired under static conditions to provide the means for modeling the speech and car-audio components (ii) recordings of acoustic noise at a wide range of static and in-motion conditions. Data are recorded with two different microphone configurations, particularly (i) a compact microphone array and (ii) a distributed microphone setup. We briefly describe the conditions under which the recordings were acquired, and we provide insight into a Python API that we designed to support the research and development of voice-enabled technologies inside moving vehicles. The first version of this Python API and part of the described dataset are available for free download.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。