中文

MIDI 表演的基于节拍的节奏量化

声音 2025-08-28 v1 计算与语言 多媒体 音频与语音处理

摘要

我们提出了一种基于 Transformer 的节奏量化模型,该模型结合节拍和强拍信息,将 MIDI 表演量化为节拍对齐的、人类可读的乐谱。我们提出了一种基于节拍的预处理方法,将乐谱和表演数据转换为统一的 Token 表示。我们优化了模型架构和数据表示,并在钢琴和吉他表演上进行了训练。基于 MUSTER 指标,我们的模型超越了 SOTA 性能。

关键词

引用

@article{arxiv.2508.19262,
  title  = {Beat-Based Rhythm Quantization of MIDI Performances},
  author = {Maximilian Wachter and Sebastian Murgul and Michael Heizmann},
  journal= {arXiv preprint arXiv:2508.19262},
  year   = {2025}
}

备注

Accepted to the Late Breaking Demo Papers of the 1st AES International Conference on Artificial Intelligence and Machine Learning for Audio (AIMLA LBDP), 2025