English

General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline

Sound 2018-10-09 v3 Machine Learning Audio and Speech Processing Machine Learning

Abstract

This paper describes Task 2 of the DCASE 2018 Challenge, titled "General-purpose audio tagging of Freesound content with AudioSet labels". This task was hosted on the Kaggle platform as "Freesound General-Purpose Audio Tagging Challenge". The goal of the task is to build an audio tagging system that can recognize the category of an audio clip from a subset of 41 diverse categories drawn from the AudioSet Ontology. We present the task, the dataset prepared for the competition, and a baseline system.

Keywords

Cite

@article{arxiv.1807.09902,
  title  = {General-purpose Tagging of Freesound Audio with AudioSet Labels: Task Description, Dataset, and Baseline},
  author = {Eduardo Fonseca and Manoj Plakal and Frederic Font and Daniel P. W. Ellis and Xavier Favory and Jordi Pons and Xavier Serra},
  journal= {arXiv preprint arXiv:1807.09902},
  year   = {2018}
}

Comments

Camera ready for DCASE Workshop 2018