Seed-Music可精准控制风格生成高质量音乐并支持歌词旋律编辑
Seed-Music: A Unified Framework for High Quality and Controlled Music Generation
- 融合自回归与扩散模型,实现多模态输入的音乐生成
- 支持从风格描述到乐谱的细粒度控制,生成音质优秀
- 提供交互式编辑工具,可直接修改生成音频的歌词和人声旋律
我们提出Seed-Music,一套能够生成高质量音乐并实现细粒度风格控制的音乐生成系统。该统一框架结合自回归语言建模与扩散方法,支持两种核心音乐创作流程:可控音乐生成与后期编辑。在可控生成中,系统可基于多模态输入(如风格描述、音频参考、乐谱和语音提示)生成带表演控制的人声音乐。在后期编辑中,提供交互式工具,可直接在生成音频中编辑歌词与人声旋律。更多示例请访问 https://team.doubao.com/seed-music。
原文摘要 · Abstract (English)
We introduce Seed-Music, a suite of music generation systems capable of producing high-quality music with fine-grained style control. Our unified framework leverages both auto-regressive language modeling and diffusion approaches to support two key music creation workflows: controlled music generation and post-production editing. For controlled music generation, our system enables vocal music generation with performance controls from multi-modal inputs, including style descriptions, audio references, musical scores, and voice prompts. For post-production editing, it offers interactive tools for editing lyrics and vocal melodies directly in the generated audio. We encourage readers to listen to demo audio examples at https://team.doubao.com/seed-music "https://team.doubao.com/seed-music".
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。