Audio and Speech Processing · Electrical Eng. & Systems
Investigating Stochastic Methods for Prosody Modeling in Speech Synthesis
Paul Mayer, Florian Lux, Alejandro Pérez-González-de-Martos, Angelina Elizarova +3
2025-07-02
Audio and Speech Processing · Electrical Eng. & Systems
Controllable Prosody Generation With Partial Inputs
Dan Andrei Iliescu, Devang Savita Ram Mohan, Tian Huey Teh, Zack Hodari
2024-04-17
Audio and Speech Processing · Electrical Eng. & Systems
Ctrl-P: Temporal Control of Prosodic Variation for Speech Synthesis
Devang S Ram Mohan, Vivian Hu, Tian Huey Teh, Alexandra Torresquintero +5
2021-06-17
Audio and Speech Processing · Electrical Eng. & Systems
CAMP: a Two-Stage Approach to Modelling Prosody in Context
Zack Hodari, Alexis Moinet, Sri Karlapati, Jaime Lorenzo-Trueba +5
2021-02-15
Audio and Speech Processing · Electrical Eng. & Systems
Controllable Neural Prosody Synthesis
Max Morrison, Zeyu Jin, Justin Salamon, Nicholas J. Bryan +1
2020-08-13
Computation and Language · Computer Science
ProsodyLM: Uncovering the Emerging Prosody Processing Capabilities in Speech Language Models
Kaizhi Qian, Xulin Fan, Junrui Ni, Slava Shechtman +3
2025-08-11
Audio and Speech Processing · Electrical Eng. & Systems
Prosody Learning Mechanism for Speech Synthesis System Without Text Length Limit
Zhen Zeng, Jianzong Wang, Ning Cheng, Jing Xiao
2020-08-14
Audio and Speech Processing · Electrical Eng. & Systems
Conditional Diffusion Probabilistic Model for Speech Enhancement
Yen-Ju Lu, Zhong-Qiu Wang, Shinji Watanabe, Alexander Richard +2
2022-02-11
Audio and Speech Processing · Electrical Eng. & Systems
An Optimized Signal Processing Pipeline for Syllable Detection and Speech Rate Estimation
Kamini Sabu, Syomantak Chaudhuri, Preeti Rao, Mahesh Patil
2021-03-09
Computation and Language · Computer Science
Prosody-Based Automatic Segmentation of Speech into Sentences and Topics
E. Shriberg, A. Stolcke, D. Hakkani-Tur, G. Tur
2022-02-28
Audio and Speech Processing · Electrical Eng. & Systems
ProMode: A Speech Prosody Model Conditioned on Acoustic and Textual Inputs
Eray Eren, Qingju Liu, Hyeongwoo Kim, Pablo Garrido +1
2025-08-14
Computation and Language · Computer Science
Alternate Endings: Improving Prosody for Incremental Neural TTS with Predicted Future Text Input
Brooke Stephenson, Thomas Hueber, Laurent Girin, Laurent Besacier
2021-06-16
Computation and Language · Computer Science
Learning De-identified Representations of Prosody from Raw Audio
Jack Weston, Raphael Lenain, Udeepa Meepegama, Emil Fristed
2021-07-20
Audio and Speech Processing · Electrical Eng. & Systems
Context-Aware Prosody Correction for Text-Based Speech Editing
Max Morrison, Lucas Rencker, Zeyu Jin, Nicholas J. Bryan +2
2021-02-17