Towards joint sound scene and polyphonic sound event recognition
Audio and Speech Processing
2019-07-02 v2 Sound
Abstract
Acoustic Scene Classification (ASC) and Sound Event Detection (SED) are two separate tasks in the field of computational sound scene analysis. In this work, we present a new dataset with both sound scene and sound event labels and use this to demonstrate a novel method for jointly classifying sound scenes and recognizing sound events. We show that by taking a joint approach, learning is more efficient and whilst improvements are still needed for sound event detection, SED results are robust in a dataset where the sample distribution is skewed towards sound scenes.
Keywords
Cite
@article{arxiv.1904.10408,
title = {Towards joint sound scene and polyphonic sound event recognition},
author = {Helen L. Bear and Ines Nolasco and Emmanouil Benetos},
journal= {arXiv preprint arXiv:1904.10408},
year = {2019}
}
Comments
Accepted to Interspeech 2019