English

Building Corpora for Single-Channel Speech Separation Across Multiple Domains

Computation and Language 2024-10-30 v1

Abstract

To date, the bulk of research on single-channel speech separation has been conducted using clean, near-field, read speech, which is not representative of many modern applications. In this work, we develop a procedure for constructing high-quality synthetic overlap datasets, necessary for most deep learning-based separation frameworks. We produced datasets that are more representative of realistic applications using the CHiME-5 and Mixer 6 corpora and evaluate standard methods on this data to demonstrate the shortcomings of current source-separation performance. We also demonstrate the value of a wide variety of data in training robust models that generalize well to multiple conditions.

Keywords

Cite

@article{arxiv.1811.02641,
  title  = {Building Corpora for Single-Channel Speech Separation Across Multiple Domains},
  author = {Matthew Maciejewski and Gregory Sell and Leibny Paola Garcia-Perera and Shinji Watanabe and Sanjeev Khudanpur},
  journal= {arXiv preprint arXiv:1811.02641},
  year   = {2024}
}

Comments

This work has been submitted to the IEEE for possible publication

R2 v1 2026-06-23T05:07:02.149Z