English

Timers and Such: A Practical Benchmark for Spoken Language Understanding with Numbers

Computation and Language 2021-10-04 v2 Audio and Speech Processing

Abstract

This paper introduces Timers and Such, a new open source dataset of spoken English commands for common voice control use cases involving numbers. We describe the gap in existing spoken language understanding datasets that Timers and Such fills, the design and creation of the dataset, and experiments with a number of ASR-based and end-to-end baseline models, the code for which has been made available as part of the SpeechBrain toolkit.

Keywords

Cite

@article{arxiv.2104.01604,
  title  = {Timers and Such: A Practical Benchmark for Spoken Language Understanding with Numbers},
  author = {Loren Lugosch and Piyush Papreja and Mirco Ravanelli and Abdelwahab Heba and Titouan Parcollet},
  journal= {arXiv preprint arXiv:2104.01604},
  year   = {2021}
}

Comments

Accepted to NeurIPS 2021 - Datasets and Benchmarks Track

R2 v1 2026-06-24T00:50:18.364Z