English

Recognize Foreign Low-Frequency Words with Similar Pairs

Computation and Language 2015-06-17 v1

Abstract

Low-frequency words place a major challenge for automatic speech recognition (ASR). The probabilities of these words, which are often important name entities, are generally under-estimated by the language model (LM) due to their limited occurrences in the training data. Recently, we proposed a word-pair approach to deal with the problem, which borrows information of frequent words to enhance the probabilities of low-frequency words. This paper presents an extension to the word-pair method by involving multiple `predicting words' to produce better estimation for low-frequency words. We also employ this approach to deal with out-of-language words in the task of multi-lingual speech recognition.

Keywords

Cite

@article{arxiv.1506.04940,
  title  = {Recognize Foreign Low-Frequency Words with Similar Pairs},
  author = {Xi Ma and Xiaoxi Wang and Dong Wang and Zhiyong Zhang},
  journal= {arXiv preprint arXiv:1506.04940},
  year   = {2015}
}
R2 v1 2026-06-22T09:54:28.585Z