English

Automatic DJ Transitions with Differentiable Audio Effects and Generative Adversarial Networks

Sound 2022-02-18 v2 Machine Learning Audio and Speech Processing

Abstract

A central task of a Disc Jockey (DJ) is to create a mixset of mu-sic with seamless transitions between adjacent tracks. In this paper, we explore a data-driven approach that uses a generative adversarial network to create the song transition by learning from real-world DJ mixes. In particular, the generator of the model uses two differentiable digital signal processing components, an equalizer (EQ) and a fader, to mix two tracks selected by a data generation pipeline. The generator has to set the parameters of the EQs and fader in such away that the resulting mix resembles real mixes created by humanDJ, as judged by the discriminator counterpart. Result of a listening test shows that the model can achieve competitive results compared with a number of baselines.

Cite

@article{arxiv.2110.06525,
  title  = {Automatic DJ Transitions with Differentiable Audio Effects and Generative Adversarial Networks},
  author = {Bo-Yu Chen and Wei-Han Hsu and Wei-Hsiang Liao and Marco A. Martínez Ramírez and Yuki Mitsufuji and Yi-Hsuan Yang},
  journal= {arXiv preprint arXiv:2110.06525},
  year   = {2022}
}

Comments

To be published at ICASSP 2022

R2 v1 2026-06-24T06:51:03.540Z