实现分段精准控音节的歌词生成,让歌词结构更自然。
Song Form-aware Full-Song Text-to-Lyrics Generation with Multi-Level Granularity Syllable Count Control
- 分词、短语、句子、段落多层级控制音节数。
- 生成歌词严格匹配指定音节和歌曲结构要求。
- 适合音乐创作与自动作词系统开发者使用。
歌词生成面临独特挑战,尤其在保持音节精确控制的同时遵循歌曲结构(如主歌、副歌)。传统逐行生成方法常导致语义不连贯,亟需更细粒度的音节管理。本文提出一种新框架,支持在词、短语、行、段落四个层级实现多级音节控制,并感知歌曲形式。该方法基于输入文本和歌曲结构生成完整歌词,确保与指定音节约束对齐。生成样例可访问:https://tinyurl.com/lyrics9999
原文摘要 · Abstract (English)
Lyrics generation presents unique challenges, particularly in achieving precise syllable control while adhering to song form structures such as verses and choruses. Conventional line-by-line approaches often lead to unnatural phrasing, underscoring the need for more granular syllable management. We propose a framework for lyrics generation that enables multi-level syllable control at the word, phrase, line, and paragraph levels, aware of song form. Our approach generates complete lyrics conditioned on input text and song form, ensuring alignment with specified syllable constraints. Generated lyrics samples are available at: https://tinyurl.com/lyrics9999
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。