arXiv:2506.07073cs.SDcs.HC2025-06被引 1

AI音乐模型能用单一音色生成多重旋律,突破传统听觉认知。

Insights on Harmonic Tones from a Generative Music Experiment

  • 用单音复合音实现多声部旋律生成,基于单调音序列构建
  • 音乐制作人成功感知并使用复合音传递两个以上音高
  • 为音乐感知与生成的交叉研究提供新视角,适合音乐科技探索者

生成式音乐人工智能的最终目标是服务于音乐创作。本研究通过一次跨学科艺术-科学实验室实验,邀请研究人员、音乐制作人与一款生成低音类音频的AI模型协作。实验中发现,制作人能够利用模型输出的单一谐波复合音传达两个或更多音高,表明该模型已学会通过单调音序列生成结构清晰、连贯的多重同时旋律线。这一现象促使人们重新思考人类是否可将谐波感知为独立音高这一长期争议,并揭示生成式AI不仅能激发音乐创造力,还能深化对音乐本质的理解。

原文摘要 · Abstract (English)

The ultimate purpose of generative music AI is music production. The studio-lab, a social form within the art-science branch of cross-disciplinarity, is a way to advance music production with AI music models. During a studio-lab experiment involving researchers, music producers, and an AI model for music generating bass-like audio, it was observed that the producers used the model's output to convey two or more pitches with a single harmonic complex tone, which in turn revealed that the model had learned to generate structured and coherent simultaneous melodic lines using monophonic sequences of harmonic complex tones. These findings prompt a reconsideration of the long-standing debate on whether humans can perceive harmonics as distinct pitches and highlight how generative AI can not only enhance musical creativity but also contribute to a deeper understanding of music.

生成音乐谐波感知人机协作音频生成

Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。