arXiv:2508.20273eess.AScs.SD2025-08

从K-pop现场表演中自动分离出实时演唱人声

Live Vocal Extraction from K-pop Performances

  • 结合源分离、互相关与幅度缩放技术
  • 成功去除预录人声和伴奏,保留真实演唱音轨
  • 为粉丝互动与音乐分析提供新工具

K-pop的全球成功得益于其充满活力的现场表演和高度参与的粉丝文化。受K-pop粉丝文化的启发,我们提出一种自动从现场表演中提取实时人声的方法。该方法结合源分离、互相关分析与幅度缩放技术,有效移除预录制的人声和伴奏,仅保留现场演唱的音频。本研究首次系统定义了“现场人声分离”任务,并为后续相关研究奠定了基础。

原文摘要 · Abstract (English)

K-pop's global success is fueled by its dynamic performances and vibrant fan engagement. Inspired by K-pop fan culture, we propose a methodology for automatically extracting live vocals from performances. We use a combination of source separation, cross-correlation, and amplitude scaling to automatically remove pre-recorded vocals and instrumentals from a live performance. Our preliminary work introduces the task of live vocal separation and provides a foundation for future research in this topic.

语音分离K-pop音频处理

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。