Audio-to-Score Transcription using Pre-trained Features, Data Augmentation, and the New SheetSage-A2S Dataset
基于预训练特征、数据增强及新SheetSage-A2S数据集的音频转乐谱转录
机构 * University College Dublin(都柏林大学学院) ; Guangxi Normal University(广西师范大学) ; Great Bay University(大湾区大学) ; Shenzhen University(深圳大学)
AI总结 该研究针对流行音乐音频转乐谱研究不足的问题,构建了SheetSage-A2S数据集,结合预训练模型MuQ与数据增强改进A2S方法,在古典与流行音乐基准上均取得优于现有技术的性能。
Comments Accepted at the 34th ACM International Conference on Multimedia (MM '26)