General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline
Sound
2018-10-09 v3 Machine Learning
Audio and Speech Processing
Machine Learning
Abstract
This paper describes Task 2 of the DCASE 2018 Challenge, titled "General-purpose audio tagging of Freesound content with AudioSet labels". This task was hosted on the Kaggle platform as "Freesound General-Purpose Audio Tagging Challenge". The goal of the task is to build an audio tagging system that can recognize the category of an audio clip from a subset of 41 diverse categories drawn from the AudioSet Ontology. We present the task, the dataset prepared for the competition, and a baseline system.
Cite
@article{arxiv.1807.09902,
title = {General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline},
author = {Eduardo Fonseca and Manoj Plakal and Frederic Font and Daniel P. W. Ellis and Xavier Favory and Jordi Pons and Xavier Serra},
journal= {arXiv preprint arXiv:1807.09902},
year = {2018}
}
Comments
Camera ready for DCASE Workshop 2018