Computation and Language · Computer Science
Sample, Translate, Recombine: Leveraging Audio Alignments for Data Augmentation in End-to-end Speech Translation
Tsz Kin Lam, Shigehiko Schamoni, Stefan Riezler
2023-06-12
Computation and Language · Computer Science
On Using SpecAugment for End-to-End Speech Translation
Parnia Bahar, Albert Zeyer, Ralf Schlüter, Hermann Ney
2019-11-21
Audio and Speech Processing · Electrical Eng. & Systems
ASR data augmentation in low-resource settings using cross-lingual multi-speaker TTS and cross-lingual voice conversion
Edresson Casanova, Christopher Shulby, Alexander Korolev, Arnaldo Candido Junior +3
2023-05-23
Computation and Language · Computer Science
Speech Synthesis as Augmentation for Low-Resource ASR
Deblin Bagchi, Shannon Wotherspoon, Zhuolin Jiang, Prasanna Muthukumar
2020-12-25
Audio and Speech Processing · Electrical Eng. & Systems
Using Speech Synthesis to Train End-to-End Spoken Language Understanding Models
Loren Lugosch, Brett Meyer, Derek Nowrouzezahrai, Mirco Ravanelli
2019-10-22
Audio and Speech Processing · Electrical Eng. & Systems
Low-Resource Text-to-Speech Synthesis Using Noise-Augmented Training of ForwardTacotron
Kishor Kayyar Lakshminarayana, Frank Zalkow, Christian Dittmar, Nicola Pia +1
2025-06-03
Audio and Speech Processing · Electrical Eng. & Systems
You Do Not Need More Data: Improving End-To-End Speech Recognition by Text-To-Speech Data Augmentation
Aleksandr Laptev, Roman Korostik, Aleksey Svischev, Andrei Andrusenko +2
2020-12-21
Computation and Language · Computer Science
Leveraging Weakly Supervised Data to Improve End-to-End Speech-to-Text Translation
Ye Jia, Melvin Johnson, Wolfgang Macherey, Ron J. Weiss +5
2019-02-12
Audio and Speech Processing · Electrical Eng. & Systems
Can Speaker Augmentation Improve Multi-Speaker End-to-End TTS?
Erica Cooper, Cheng-I Lai, Yusuke Yasuda, Junichi Yamagishi
2020-08-10
Audio and Speech Processing · Electrical Eng. & Systems
Noise Robust TTS for Low Resource Speakers using Pre-trained Model and Speech Enhancement
Dongyang Dai, Li Chen, Yuping Wang, Mu Wang +4
2020-10-23
Audio and Speech Processing · Electrical Eng. & Systems
Generating Data with Text-to-Speech and Large-Language Models for Conversational Speech Recognition
Samuele Cornell, Jordan Darefsky, Zhiyao Duan, Shinji Watanabe
2024-08-20
Audio and Speech Processing · Electrical Eng. & Systems
Cross-speaker style transfer for text-to-speech using data augmentation
Manuel Sam Ribeiro, Julian Roth, Giulia Comini, Goeric Huybrechts +2
2022-02-11
Computation and Language · Computer Science
Improving End-to-End Speech Processing by Efficient Text Data Utilization with Latent Synthesis
Jianqiao Lu, Wenyong Huang, Nianzu Zheng, Xingshan Zeng +2
2023-10-25
Computation and Language · Computer Science
Listen and Translate: A Proof of Concept for End-to-End Speech-to-Text Translation
Alexandre Berard, Olivier Pietquin, Christophe Servan, Laurent Besacier
2016-12-07
Audio and Speech Processing · Electrical Eng. & Systems
Low-resource expressive text-to-speech using data augmentation
Goeric Huybrechts, Thomas Merritt, Giulia Comini, Bartek Perz +2
2021-06-03
Computation and Language · Computer Science
Elderly-Contextual Data Augmentation via Speech Synthesis for Elderly ASR
Minsik Lee, Seoi Hong, Chongmin Lee, Sieun Choi +3
2026-04-29
Audio and Speech Processing · Electrical Eng. & Systems
Auditory-Based Data Augmentation for End-to-End Automatic Speech Recognition
Zehai Tu, Jack Deadman, Ning Ma, Jon Barker
2022-04-12
Audio and Speech Processing · Electrical Eng. & Systems
Speechless: Speech Instruction Training Without Speech for Low Resource Languages
Alan Dao, Dinh Bach Vu, Huy Hoang Ha, Tuan Le Duc Anh +5
2025-08-26
Computation and Language · Computer Science
From Start to Finish: Latency Reduction Strategies for Incremental Speech Synthesis in Simultaneous Speech-to-Speech Translation
Danni Liu, Changhan Wang, Hongyu Gong, Xutai Ma +2
2022-07-18