English

Dyn-ASR: Compact, Multilingual Speech Recognition via Spoken Language and Accent Identification

Computation and Language 2021-08-05 v1 Sound Audio and Speech Processing

Abstract

Running automatic speech recognition (ASR) on edge devices is non-trivial due to resource constraints, especially in scenarios that require supporting multiple languages. We propose a new approach to enable multilingual speech recognition on edge devices. This approach uses both language identification and accent identification to select one of multiple monolingual ASR models on-the-fly, each fine-tuned for a particular accent. Initial results for both recognition performance and resource usage are promising with our approach using less than 1/12th of the memory consumed by other solutions.

Keywords

Cite

@article{arxiv.2108.02034,
  title  = {Dyn-ASR: Compact, Multilingual Speech Recognition via Spoken Language and Accent Identification},
  author = {Sangeeta Ghangam and Daniel Whitenack and Joshua Nemecek},
  journal= {arXiv preprint arXiv:2108.02034},
  year   = {2021}
}

Comments

Accepted to IEEE WF-IOT 2021

R2 v1 2026-06-24T04:49:28.450Z