English

An evaluation of data augmentation methods for sound scene geotagging

Audio and Speech Processing 2021-10-12 v1 Sound

Abstract

Sound scene geotagging is a new topic of research which has evolved from acoustic scene classification. It is motivated by the idea of audio surveillance. Not content with only describing a scene in a recording, a machine which can locate where the recording was captured would be of use to many. In this paper we explore a series of common audio data augmentation methods to evaluate which best improves the accuracy of audio geotagging classifiers. Our work improves on the state-of-the-art city geotagging method by 23% in terms of classification accuracy.

Keywords

Cite

@article{arxiv.2110.04585,
  title  = {An evaluation of data augmentation methods for sound scene geotagging},
  author = {Helen L. Bear and Veronica Morfi and Emmanouil Benetos},
  journal= {arXiv preprint arXiv:2110.04585},
  year   = {2021}
}

Comments

Presented at Interspeech 2021

R2 v1 2026-06-24T06:45:43.357Z