English

Towards joint sound scene and polyphonic sound event recognition

Audio and Speech Processing 2019-07-02 v2 Sound

Abstract

Acoustic Scene Classification (ASC) and Sound Event Detection (SED) are two separate tasks in the field of computational sound scene analysis. In this work, we present a new dataset with both sound scene and sound event labels and use this to demonstrate a novel method for jointly classifying sound scenes and recognizing sound events. We show that by taking a joint approach, learning is more efficient and whilst improvements are still needed for sound event detection, SED results are robust in a dataset where the sample distribution is skewed towards sound scenes.

Keywords

Cite

@article{arxiv.1904.10408,
  title  = {Towards joint sound scene and polyphonic sound event recognition},
  author = {Helen L. Bear and Ines Nolasco and Emmanouil Benetos},
  journal= {arXiv preprint arXiv:1904.10408},
  year   = {2019}
}

Comments

Accepted to Interspeech 2019