English

WaveFake: A Data Set to Facilitate Audio Deepfake Detection

Machine Learning 2021-11-05 v1 Cryptography and Security Sound Audio and Speech Processing

Abstract

Deep generative modeling has the potential to cause significant harm to society. Recognizing this threat, a magnitude of research into detecting so-called "Deepfakes" has emerged. This research most often focuses on the image domain, while studies exploring generated audio signals have, so-far, been neglected. In this paper we make three key contributions to narrow this gap. First, we provide researchers with an introduction to common signal processing techniques used for analyzing audio signals. Second, we present a novel data set, for which we collected nine sample sets from five different network architectures, spanning two languages. Finally, we supply practitioners with two baseline models, adopted from the signal processing community, to facilitate further research in this area.

Keywords

Cite

@article{arxiv.2111.02813,
  title  = {WaveFake: A Data Set to Facilitate Audio Deepfake Detection},
  author = {Joel Frank and Lea Schönherr},
  journal= {arXiv preprint arXiv:2111.02813},
  year   = {2021}
}

Comments

Accepted to NeurIPS 2021 (Benchmark and Dataset Track); Code: https://github.com/RUB-SysSec/WaveFake; Data: https://zenodo.org/record/5642694

R2 v1 2026-06-24T07:25:59.877Z