让背景音乐随场景和用户状态实时变化,更懂环境与个性。
MetaBGM: Dynamic Soundtrack Transformation For Continuous Multi-Scene Experiences With Ambient Awareness And Personalization
- 两阶段生成:从数据到音乐描述文本的自动转化
- 实测可生成贴合场景、动态变化的背景音乐
- 适合游戏、影视等需要沉浸式音效的交互应用
本文提出 MetaBGM,一种用于生成自适应背景音乐的创新框架,能够响应动态场景和实时用户交互。我们定义多场景为环境上下文的变化,如游戏场景或电影片段的切换。针对将后端数据转化为音频生成模型可用的音乐描述文本这一挑战,MetaBGM采用新颖的两阶段生成方法,将连续的场景与用户状态数据转换为文本描述,并输入音频生成模型以实现实时原声配乐生成。实验结果表明,MetaBGM能有效生成符合语境且动态变化的背景音乐,适用于交互式应用场景。
原文摘要 · Abstract (English)
This paper introduces MetaBGM, a groundbreaking framework for generating background music that adapts to dynamic scenes and real-time user interactions. We define multi-scene as variations in environmental contexts, such as transitions in game settings or movie scenes. To tackle the challenge of converting backend data into music description texts for audio generation models, MetaBGM employs a novel two-stage generation approach that transforms continuous scene and user state data into these texts, which are then fed into an audio generation model for real-time soundtrack creation. Experimental results demonstrate that MetaBGM effectively generates contextually relevant and dynamic background music for interactive applications.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。