Audio and Speech Processing · Electrical Eng. & Systems
HiFiTTS-2: A Large-Scale High Bandwidth Speech Dataset
Ryan Langman, Xuesong Yang, Paarth Neekhara, Shehzeen Hussain +3
2025-09-23
Audio and Speech Processing · Electrical Eng. & Systems
LibriTTS-R: A Restored Multi-Speaker Text-to-Speech Corpus
Yuma Koizumi, Heiga Zen, Shigeki Karita, Yifan Ding +6
2023-05-31
Audio and Speech Processing · Electrical Eng. & Systems
MLS: A Large-Scale Multilingual Dataset for Speech Research
Vineel Pratap, Qiantong Xu, Anuroop Sriram, Gabriel Synnaeve +1
2020-12-22
Computation and Language · Computer Science
Combining speakers of multiple languages to improve quality of neural voices
Javier Latorre, Charlotte Bailleul, Tuuli Morrill, Alistair Conkie +1
2021-08-18
Audio and Speech Processing · Electrical Eng. & Systems
Training Multi-Speaker Neural Text-to-Speech Systems using Speaker-Imbalanced Speech Corpora
Hieu-Thi Luong, Xin Wang, Junichi Yamagishi, Nobuyuki Nishizawa
2019-04-09
Audio and Speech Processing · Electrical Eng. & Systems
Low-resource expressive text-to-speech using data augmentation
Goeric Huybrechts, Thomas Merritt, Giulia Comini, Bartek Perz +2
2021-06-03
Audio and Speech Processing · Electrical Eng. & Systems
Improving the quality of neural TTS using long-form content and multi-speaker multi-style modeling
Tuomo Raitio, Javier Latorre, Andrea Davis, Tuuli Morrill +1
2023-06-29
Audio and Speech Processing · Electrical Eng. & Systems
Text-To-Speech Synthesis In The Wild
Jee-weon Jung, Wangyou Zhang, Soumi Maiti, Yihan Wu +10
2025-06-03
Audio and Speech Processing · Electrical Eng. & Systems
CML-TTS A Multilingual Dataset for Speech Synthesis in Low-Resource Languages
Frederico S. Oliveira, Edresson Casanova, Arnaldo Cândido Júnior, Anderson S. Soares +1
2023-06-21
Audio and Speech Processing · Electrical Eng. & Systems
KazakhTTS: An Open-Source Kazakh Text-to-Speech Synthesis Dataset
Saida Mussakhojayeva, Aigerim Janaliyeva, Almas Mirzakhmetov, Yerbolat Khassanov +1
2021-09-09
Audio and Speech Processing · Electrical Eng. & Systems
MultiSpeech: Multi-Speaker Text to Speech with Transformer
Mingjian Chen, Xu Tan, Yi Ren, Jin Xu +4
2020-08-04
Audio and Speech Processing · Electrical Eng. & Systems
Extending Multilingual Speech Synthesis to 100+ Languages without Transcribed Data
Takaaki Saeki, Gary Wang, Nobuyuki Morioka, Isaac Elias +7
2024-07-17
Machine Learning · Computer Science
Sample Efficient Adaptive Text-to-Speech
Yutian Chen, Yannis Assael, Brendan Shillingford, David Budden +10
2019-01-18
Computation and Language · Computer Science
RyanSpeech: A Corpus for Conversational Text-to-Speech Synthesis
Rohola Zandie, Mohammad H. Mahoor, Julia Madsen, Eshrat S. Emamian
2021-06-17
Audio and Speech Processing · Electrical Eng. & Systems
DualSpeech: Enhancing Speaker-Fidelity and Text-Intelligibility Through Dual Classifier-Free Guidance
Jinhyeok Yang, Junhyeok Lee, Hyeong-Seok Choi, Seunghun Ji +2
2024-08-28
Audio and Speech Processing · Electrical Eng. & Systems
Cross-speaker style transfer for text-to-speech using data augmentation
Manuel Sam Ribeiro, Julian Roth, Giulia Comini, Goeric Huybrechts +2
2022-02-11
Audio and Speech Processing · Electrical Eng. & Systems
Transfer Learning Framework for Low-Resource Text-to-Speech using a Large-Scale Unlabeled Speech Corpus
Minchan Kim, Myeonghun Jeong, Byoung Jin Choi, Sunghwan Ahn +2
2022-10-07