MIDI 表演的基于节拍的节奏量化
声音
2025-08-28 v1 计算与语言
多媒体
音频与语音处理
摘要
我们提出了一种基于 Transformer 的节奏量化模型,该模型结合节拍和强拍信息,将 MIDI 表演量化为节拍对齐的、人类可读的乐谱。我们提出了一种基于节拍的预处理方法,将乐谱和表演数据转换为统一的 Token 表示。我们优化了模型架构和数据表示,并在钢琴和吉他表演上进行了训练。基于 MUSTER 指标,我们的模型超越了 SOTA 性能。
引用
@article{arxiv.2508.19262,
title = {Beat-Based Rhythm Quantization of MIDI Performances},
author = {Maximilian Wachter and Sebastian Murgul and Michael Heizmann},
journal= {arXiv preprint arXiv:2508.19262},
year = {2025}
}
备注
Accepted to the Late Breaking Demo Papers of the 1st AES International Conference on Artificial Intelligence and Machine Learning for Audio (AIMLA LBDP), 2025