人机实时钢琴合奏,模型即兴接棒延续乐句。
The Ghost in the Keys: A Disklavier Demo for Human-AI Musical Co-Creativity
- 用钢琴作为共享界面,实现人与生成模型的即时互动演奏。
- 模型能保持风格一致并发展出连贯的乐句结构。
- 适合音乐创作者探索人机协作新范式。
尽管音乐生成模型能力日益增强,但其在音乐家中的应用仍受限于文本提示方式,这种异步工作流与乐器演奏的具身性、响应性脱节。为此,我们提出Aria-Duet系统,通过雅马哈Disklavier钢琴作为共同物理接口,实现人类钢琴家与先进生成模型Aria之间的实时音乐二重奏。该框架支持轮流演奏:用户演奏后发出交接信号,模型即时生成连贯的延续部分,并由钢琴实际演奏出来。除描述支撑低延迟交互的技术架构外,我们从音乐学角度分析系统输出,发现模型能够维持风格语义并发展出连贯的乐句构思,表明此类具身系统可实现音乐上复杂的对话,为人类与AI协同创作开辟了有前景的新路径。
原文摘要 · Abstract (English)
While generative models for music composition are increasingly capable, their adoption by musicians is hindered by text-prompting, an asynchronous workflow disconnected from the embodied, responsive nature of instrumental performance. To address this, we introduce Aria-Duet, an interactive system facilitating a real-time musical duet between a human pianist and Aria, a state-of-the-art generative model, using a Yamaha Disklavier as a shared physical interface. The framework enables a turn-taking collaboration: the user performs, signals a handover, and the model generates a coherent continuation performed acoustically on the piano. Beyond describing the technical architecture enabling this low-latency interaction, we analyze the system's output from a musicological perspective, finding the model can maintain stylistic semantics and develop coherent phrasal ideas, demonstrating that such embodied systems can engage in musically sophisticated dialogue and open a promising new path for human-AI co-creation.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。