arXiv:2606.19987cs.SDeess.AS2026-06

构建波兰语音色语义数据集,支持跨文化音乐研究

PolSeT: Polish Semantics of Timbre Dataset

  • 通过自由描述法收集波兰语音色词汇,生成1901条语义描述
  • 用8个二元量表对18种乐器音色进行评分,重复实验提升可靠性
  • 适合从事音乐信息检索、跨文化心理声学与多语言语义建模的研究者

本文介绍PolSeT(波兰语音色语义)数据集,旨在推动波兰语及跨文化背景下的心理声学与音乐信息检索(MIR)研究。该数据集包含两个连续实验:实验1(N=60)为自由描述任务,基于11个音色刺激,共收集1901条语义描述(701个唯一词);实验2(N=105)利用该词汇库开展语义差异研究,参与者对18种乐器音色在8个二元量表上评分,并通过重复试验验证信度。发布的数据包含原始听觉反应、完整人口统计信息(经验、性别、年龄)、音频刺激及使用Python提取的声学特征。该数据集填补了开放音色研究数据的空白,为心理声学研究和多语言语义嵌入模型训练提供质性与量化双重基础。

原文摘要 · Abstract (English)

This data report introduces PolSeT (Polish Semantic Timbre), a dataset designed to facilitate research in psychoacoustics and Music Information Retrieval (MIR) in Polish and cross-cultural contexts. The dataset contains data from two sequential experiments. Experiment 1 (N=60) was a free-verbalization task aimed at creating a lexicon of Polish semantic descriptors. Using 11 stimuli, a total of 1901 descriptors (701 unique) were gathered. Experiment 2 (N=105) utilized this lexicon to conduct a semantic differential study, where participants rated 18 instrument sounds on 8 bipolar scales, with repeated trials for reliability analysis. The released dataset includes raw listener responses, comprehensive demographics (experience, gender, age), audio stimuli, and extracted acoustic features with Python extraction code. This dataset addresses a gap in open timbre research data, providing both the qualitative linguistic groundwork and the quantitative ratings necessary for psychoacoustic research and the training of multilingual semantic embedding models.

音色研究语义标注多语言

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。