构建统一编码体系,训练通用乐谱识别模型
Towards a foundational model for recognising diastematic Gregorian notation

- 设计统一编码标准,整合四个不同格式数据集
- 跨数据集测试下达到最新最佳性能
- 适合音乐信息检索与古籍数字化研究者
近年来,端到端方法被用于识别格里高利圣咏记谱法,已发布四个数据集。但每个数据集采用不同编码方式。本文基于S-GABC提案设计统一编码标准,将四个数据集转换至该格式,并训练一个共享的端到端基础模型,实现对四组数据的全面最优表现,显著提升格里高利记谱法的光学识别水平。
原文摘要 · Abstract (English)
Optical recognition of Gregorian notation has recently been attempted with end-to-end methods, with four datasets introduced. However, each of these datasets is in a different encoding. We design a common encoding based on the S-GABC proposal, convert all four datasets to this common encoding, and train a shared end-to-end foundational model for diastematic Gregorian notation that establishes a new state of the art across all four datasets.
Thank you to arXiv for use of its open access interoperability. PaperDance 不是 arXiv 官方产品;中文卡片由大模型生成,请以原文为准。