Computation and Language · Computer Science
Text Generation with Speech Synthesis for ASR Data Augmentation
Zhuangqun Huang, Gil Keren, Ziran Jiang, Shashank Jain +10
2023-05-29
Computation and Language · Computer Science
Generating Synthetic Audio Data for Attention-Based Speech Recognition Systems
Nick Rossenbach, Albert Zeyer, Ralf Schlüter, Hermann Ney
2020-02-18
Audio and Speech Processing · Electrical Eng. & Systems
Towards Selection of Text-to-speech Data to Augment ASR Training
Shuo Liu, Leda Sarı, Chunyang Wu, Gil Keren +3
2023-06-05
Computation and Language · Computer Science
Evaluating Speech Synthesis by Training Recognizers on Synthetic Speech
Dareen Alharthi, Roshan Sharma, Hira Dhamyal, Soumi Maiti +2
2023-10-03
Audio and Speech Processing · Electrical Eng. & Systems
You Do Not Need More Data: Improving End-To-End Speech Recognition by Text-To-Speech Data Augmentation
Aleksandr Laptev, Roman Korostik, Aleksey Svischev, Andrei Andrusenko +2
2020-12-21
Audio and Speech Processing · Electrical Eng. & Systems
Too Good to Be True: A Study on Modern Automatic Speech Recognition for the Evaluation of Speech Enhancement
Danilo de Oliveira, Tal Peer, Timo Gerkmann
2026-05-13
Computation and Language · Computer Science
Towards Pretraining Robust ASR Foundation Model with Acoustic-Aware Data Augmentation
Dancheng Liu, Amir Nassereldine, Chenhui Xu, Jinjun Xiong
2025-05-28
Audio and Speech Processing · Electrical Eng. & Systems
Synt++: Utilizing Imperfect Synthetic Data to Improve Speech Recognition
Ting-Yao Hu, Mohammadreza Armandpour, Ashish Shrivastava, Jen-Hao Rick Chang +2
2021-10-25
Audio and Speech Processing · Electrical Eng. & Systems
Spectral Modification Based Data Augmentation For Improving End-to-End ASR For Children's Speech
Vishwanath Pratap Singh, Hardik Sailor, Supratik Bhattacharya, Abhishek Pandey
2022-03-15
Audio and Speech Processing · Electrical Eng. & Systems
Using Synthetic Audio to Improve The Recognition of Out-Of-Vocabulary Words in End-To-End ASR Systems
Xianrui Zheng, Yulan Liu, Deniz Gunceler, Daniel Willett
2021-02-11
Audio and Speech Processing · Electrical Eng. & Systems
A Neural Acoustic Echo Canceller Optimized Using An Automatic Speech Recognizer And Large Scale Synthetic Data
Nathan Howard, Alex Park, Turaj Zakizadeh Shabestary, Alexander Gruenstein +1
2021-06-03
Audio and Speech Processing · Electrical Eng. & Systems
ASR data augmentation in low-resource settings using cross-lingual multi-speaker TTS and cross-lingual voice conversion
Edresson Casanova, Christopher Shulby, Alexander Korolev, Arnaldo Candido Junior +3
2023-05-23
Machine Learning · Computer Science
SynthASR: Unlocking Synthetic Data for Speech Recognition
Amin Fazel, Wei Yang, Yulan Liu, Roberto Barra-Chicote +3
2021-06-16
Audio and Speech Processing · Electrical Eng. & Systems
Zero Shot Text to Speech Augmentation for Automatic Speech Recognition on Low-Resource Accented Speech Corpora
Francesco Nespoli, Daniel Barreda, Patrick A. Naylor
2024-09-18
Computation and Language · Computer Science
MixSpeech: Data Augmentation for Low-resource Automatic Speech Recognition
Linghui Meng, Jin Xu, Xu Tan, Jindong Wang +2
2021-02-26
Computation and Language · Computer Science
WER we are and WER we think we are
Piotr Szymański, Piotr Żelasko, Mikolaj Morzy, Adrian Szymczak +5
2020-10-08
Computation and Language · Computer Science
Elderly-Contextual Data Augmentation via Speech Synthesis for Elderly ASR
Minsik Lee, Seoi Hong, Chongmin Lee, Sieun Choi +3
2026-04-29
Audio and Speech Processing · Electrical Eng. & Systems
The Potential of Neural Speech Synthesis-based Data Augmentation for Personalized Speech Enhancement
Anastasia Kuznetsova, Aswin Sivaraman, Minje Kim
2022-11-15