English
Related papers

Related papers: Glottal source estimation robustness: A comparison…

200 papers

Source-tract decomposition (or glottal flow estimation) is one of the basic problems of speech processing. For this, several techniques have been proposed in the literature. However studies comparing different approaches are almost…

Sound · Computer Science 2020-01-06 Thomas Drugman , Baris Bozkurt , Thierry Dutoit

Some glottal analysis approaches based upon linear prediction or complex cepstrum approaches have been proved to be effective to estimate glottal source from real speech utterances. We propose a new approach employing both an all-pole…

Sound · Computer Science 2016-12-16 Yiqiao Chen , John N. Gowdy

In a previous work, we showed that the glottal source can be estimated from speech signals by computing the Zeros of the Z-Transform (ZZT). Decomposition was achieved by separating the roots inside (causal contribution) and outside…

Sound · Computer Science 2020-05-19 Thomas Drugman , Baris Bozkurt , Thierry Dutoit

Homomorphic analysis is a well-known method for the separation of non-linearly combined signals. More particularly, the use of complex cepstrum for source-tract deconvolution has been discussed in various articles. However there exists no…

Sound · Computer Science 2020-01-01 Thomas Drugman , Baris Bozkurt , Thierry Dutoit

Complex cepstrum is known in the literature for linearly separating causal and anticausal components. Relying on advances achieved by the Zeros of the Z-Transform (ZZT) technique, we here investigate the possibility of using complex…

Sound · Computer Science 2020-01-01 Thomas Drugman , Baris Bozkurt , Thierry Dutoit

The estimation of glottal flow from a speech waveform is a key method for speech analysis and parameterization. Significant research effort has been made to dissociate the first vocal tract resonance from the glottal formant (the…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-09 Olivier Perrotin , Ian Vince McLoughlin

The pseudo-periodicity of voiced speech can be exploited in several speech processing applications. This requires however that the precise locations of the Glottal Closure Instants (GCIs) are available. The focus of this paper is the…

Sound · Computer Science 2020-01-03 Thomas Drugman , Mark Thomas , Jon Gudnason , Patrick Naylor , Thierry Dutoit

Automatic detection of voice pathology enables objective assessment and earlier intervention for the diagnosis. This study provides a systematic analysis of glottal source features and investigates their effectiveness in voice pathology…

Audio and Speech Processing · Electrical Eng. & Systems 2023-10-18 Sudarsana Reddy Kadiri , Paavo Alku

This paper addresses the problem of automatic detection of voice pathologies directly from the speech signal. For this, we investigate the use of the glottal source estimation as a means to detect voice disorders. Three sets of features are…

Sound · Computer Science 2020-01-06 Thomas Drugman , Thomas Dubuisson , Thierry Dutoit

In this paper, we propose a classification based glottal closure instants (GCI) detection from pathological acoustic speech signal, which finds many applications in vocal disorder analysis. Till date, GCI for pathological disorder is…

Sound · Computer Science 2018-11-28 Gurunath Reddy M , Tanumay Mandal , Krothapalli Sreenivasa Rao

This paper proposes a new procedure to detect Glottal Closure and Opening Instants (GCIs and GOIs) directly from speech waveforms. The procedure is divided into two successive steps. First a mean-based signal is computed, and intervals…

Sound · Computer Science 2020-01-06 Thomas Drugman , Thierry Dutoit

A sound source was proposed for acoustic measurements of physical models of the human vocal tract. The physical models are produced by Fast Prototyping, based on Magnetic Resonance Imaging during prolonged vowel production. The sound…

Instrumentation and Detectors · Physics 2017-11-22 Antti Hannukainen , Juha Kuortti , Jarmo Malinen , Antti Ojalammi

Glottal Closure Instants (GCI) detection consists in automatically detecting temporal locations of most significant excitation of the vocal tract from the speech signal. It is used in many speech analysis and processing applications, and…

Audio and Speech Processing · Electrical Eng. & Systems 2020-02-21 Luc Ardaillon , Axel Roebel

Glottal Closure Instants (GCIs) correspond to the temporal locations of significant excitation to the vocal tract occurring during the production of voiced speech. GCI detection from speech signals is a well-studied problem given its…

Sound · Computer Science 2019-07-11 Mohit Goyal , Varun Srivastava , Prathosh A. P

The Glottal Source is an important component of voice as it can be considered as the excitation signal to the voice apparatus. Nowadays, new techniques of speech processing such as speech recognition and speech synthesis use the glottal…

Other Computer Science · Computer Science 2010-04-20 Nikhil Raj , R. K. Sharma

It was recently shown that complex cepstrum can be effectively used for glottal flow estimation by separating the causal and anticausal components of speech. In order to guarantee a correct estimation, some constraints on the window have…

Sound · Computer Science 2020-05-12 Thomas Drugman , Thierry Dutoit

A dictionary learning based audio source classification algorithm is proposed to classify a sample audio signal as one amongst a finite set of different audio sources. Cosine similarity measure is used to select the atoms during dictionary…

Sound · Computer Science 2015-10-28 K V Vijay Girish , T V Ananthapadmanabha , A G Ramakrishnan

Recent speech technology research has seen a growing interest in using WaveNets as statistical vocoders, i.e., generating speech waveforms from acoustic features. These models have been shown to improve the generated speech quality over…

Audio and Speech Processing · Electrical Eng. & Systems 2018-04-26 Lauri Juvela , Vassilis Tsiaras , Bajibabu Bollepalli , Manu Airaksinen , Junichi Yamagishi , Paavo Alku

End-to-end speech recognition generally uses hand-engineered acoustic features as input and excludes the feature extraction module from its joint optimization. To extract learnable and adaptive features and mitigate information loss, we…

Sound · Computer Science 2021-06-09 Max W. Y. Lam , Jun Wang , Chao Weng , Dan Su , Dong Yu

Graph-based Transform (GT) has been recently leveraged successfully in the signal processing domain, specifically for compression purposes. In this paper, we employ the GBT, as well as the Singular Value Decomposition (SVD) with the goal to…

Audio and Speech Processing · Electrical Eng. & Systems 2020-03-19 Majid Farzaneh , Rahil Mahdian Toroghi
‹ Prev 1 2 3 10 Next ›