arXiv:2409.09988eess.AScs.SD2024-09被引 2

建模歌手间互动,让合声更统一

DNN-based ensemble singing voice synthesis with interactions between singers

  • 用多声部乐谱和交互损失函数建模歌手间影响
  • 实验显示合声统一性显著提升
  • 适合需要自然合声的音乐生成场景

我们提出一种歌唱语音合成(SVS)方法,通过建模歌手间的互动关系,实现更统一的合声效果。现有大多数SVS方法专注于单人演唱合成,未考虑歌手间相互调整声音的互动机制。若仅将单人歌声拼接成合声,会因忽略互动而降低整体统一性。为此,我们设计了一种基于多声部乐谱输入、并引入模拟互动效应的损失函数的架构。实验结果表明,该方法能有效提升合声的统一性。

原文摘要 · Abstract (English)

We propose a singing voice synthesis (SVS) method for a more unified ensemble singing voice by modeling interactions between singers. Most existing SVS methods aim to synthesize a solo voice, and do not consider interactions between singers, i.e., adjusting one's own voice to the others' voices. Since the production of ensemble voices from solo singing voices ignores the interactions, it can degrade the unity of the vocal ensemble. Therefore, we propose a SVS that reproduces the interactions. It is based on an architecture that uses musical scores of multiple voice parts, and loss functions that simulate the interactions' effect to acoustic features. Experimental results show that our methods improve the unity of the vocal ensemble.

语音合成合声生成深度学习

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。