arXiv:2506.10754cs.SDcs.AI2025-06NeurIPS

用用户自定义音乐融合环境噪音,让杂音更难被察觉。

BNMusic: Blending Environmental Noises into Personalized Music

  • 根据文本生成音乐,将噪音融入节奏对齐的旋律中
  • 自适应增强音乐片段,降低噪音感知且保持音质
  • 适合想改善嘈杂环境中听觉体验的用户

当受到环境噪音干扰时,传统声学掩蔽技术通过用主导但不突兀的声音覆盖噪音来减少不适感。然而,若主导声与噪音不匹配(如节拍错位),往往需要大幅提高音量才能有效掩蔽。受跨模态生成进展启发,本文提出一种替代方案:通过基于用户提供的文本提示生成个性化音乐,并将环境噪音融入其中,以降低噪音的可察觉性。我们构建了BNMusic框架,包含两个阶段:第一阶段生成包含噪音音乐本质的完整旋律(以梅尔频谱表示);第二阶段自适应增强生成的音乐段落,进一步降低噪音感知并提升融合效果,同时保证听觉质量。在MusicBench、EPIC-SOUNDS和ESC-50上的综合评估表明,该方法能实现节奏对齐、自适应增强且令人愉悦的音乐融合,显著降低噪音可见度,从而提升整体听觉体验。

原文摘要 · Abstract (English)

While being disturbed by environmental noises, the acoustic masking technique is a conventional way to reduce the annoyance in audio engineering that seeks to cover up the noises with other dominant yet less intrusive sounds. However, misalignment between the dominant sound and the noise-such as mismatched downbeats-often requires an excessive volume increase to achieve effective masking. Motivated by recent advances in cross-modal generation, in this work, we introduce an alternative method to acoustic masking, aiming to reduce the noticeability of environmental noises by blending them into personalized music generated based on user-provided text prompts. Following the paradigm of music generation using mel-spectrogram representations, we propose a Blending Noises into Personalized Music (BNMusic) framework with two key stages. The first stage synthesizes a complete piece of music in a mel-spectrogram representation that encapsulates the musical essence of the noise. In the second stage, we adaptively amplify the generated music segment to further reduce noise perception and enhance the blending effectiveness, while preserving auditory quality. Our experiments with comprehensive evaluations on MusicBench, EPIC-SOUNDS, and ESC-50 demonstrate the effectiveness of our framework, highlighting the ability to blend environmental noise with rhythmically aligned, adaptively amplified, and enjoyable music segments, minimizing the noticeability of the noise, thereby improving overall acoustic experiences. Project page: https://d-fas.github.io/BNMusic_page/.

音乐生成噪音抑制跨模态

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。