中文
相关论文

相关论文: Complex Cepstrum-based Decomposition of Speech for…

200 篇论文

Complex cepstrum is known in the literature for linearly separating causal and anticausal components. Relying on advances achieved by the Zeros of the Z-Transform (ZZT) technique, we here investigate the possibility of using complex…

声音 · 计算机科学 2020-01-01 Thomas Drugman , Baris Bozkurt , Thierry Dutoit

It was recently shown that complex cepstrum can be effectively used for glottal flow estimation by separating the causal and anticausal components of speech. In order to guarantee a correct estimation, some constraints on the window have…

声音 · 计算机科学 2020-05-12 Thomas Drugman , Thierry Dutoit

Some glottal analysis approaches based upon linear prediction or complex cepstrum approaches have been proved to be effective to estimate glottal source from real speech utterances. We propose a new approach employing both an all-pole…

声音 · 计算机科学 2016-12-16 Yiqiao Chen , John N. Gowdy

In a previous work, we showed that the glottal source can be estimated from speech signals by computing the Zeros of the Z-Transform (ZZT). Decomposition was achieved by separating the roots inside (causal contribution) and outside…

声音 · 计算机科学 2020-05-19 Thomas Drugman , Baris Bozkurt , Thierry Dutoit

Source-tract decomposition (or glottal flow estimation) is one of the basic problems of speech processing. For this, several techniques have been proposed in the literature. However studies comparing different approaches are almost…

声音 · 计算机科学 2020-01-06 Thomas Drugman , Baris Bozkurt , Thierry Dutoit

This paper addresses the problem of estimating the voice source directly from speech waveforms. A novel principle based on Anticausality Dominated Regions (ACDR) is used to estimate the glottal open phase. This technique is compared to two…

音频与语音处理 · 电气工程与系统科学 2020-05-26 Thomas Drugman , Thomas Dubuisson , Alexis Moinet , Nicolas D'Alessandro , Thierry Dutoit

Automatic detection of voice pathology enables objective assessment and earlier intervention for the diagnosis. This study provides a systematic analysis of glottal source features and investigates their effectiveness in voice pathology…

音频与语音处理 · 电气工程与系统科学 2023-10-18 Sudarsana Reddy Kadiri , Paavo Alku

In this paper, we propose an effective and robust method of spatial feature extraction for acoustic scene analysis utilizing partially synchronized and/or closely located distributed microphones. In the proposed method, a new cepstrum…

音频与语音处理 · 电气工程与系统科学 2020-04-22 Keisuke Imoto

In this paper, we propose an effective and robust method for acoustic scene analysis based on spatial information extracted from partially synchronized and/or closely located distributed microphones. In the proposed method, to extract…

声音 · 计算机科学 2018-07-10 Keisuke Imoto

The estimation of glottal flow from a speech waveform is a key method for speech analysis and parameterization. Significant research effort has been made to dissociate the first vocal tract resonance from the glottal formant (the…

音频与语音处理 · 电气工程与系统科学 2021-06-09 Olivier Perrotin , Ian Vince McLoughlin

The great majority of current voice technology applications relies on acoustic features characterizing the vocal tract response, such as the widely used MFCC of LPC parameters. Nonetheless, the airflow passing through the vocal folds, and…

声音 · 计算机科学 2020-01-01 Thomas Drugman , Paavo Alku , Abeer Alwan , Bayya Yegnanarayana

The pseudo-periodicity of voiced speech can be exploited in several speech processing applications. This requires however that the precise locations of the Glottal Closure Instants (GCIs) are available. The focus of this paper is the…

声音 · 计算机科学 2020-01-03 Thomas Drugman , Mark Thomas , Jon Gudnason , Patrick Naylor , Thierry Dutoit

Underwater acoustic monitoring systems record many hours of audio data for marine research, making fast and reliable non-causal signal detection paramount. Such detectors assist in reducing the amount of labor required for signal…

信号处理 · 电气工程与系统科学 2023-02-07 Marco W. Rademan , Daniel J. Versfeld , Johan A. du Preez

Source wavelet estimation is the key in seismic signal processing for resolving subsurface structural properties. Homomorphic deconvolution using cepstrum analysis has been an effective method for wavelet estimation for decades. In general,…

信息论 · 计算机科学 2012-06-06 K. H. Miah , R. H. Herrera , M. van der Baan , M. D. Sacchi

In this paper, we propose a novel family of windowing technique to compute Mel Frequency Cepstral Coefficient (MFCC) for automatic speaker recognition from speech. The proposed method is based on fundamental property of discrete time…

计算机视觉与模式识别 · 计算机科学 2015-06-05 Md. Sahidullah , Goutam Saha

The Z Transform is a mathematical operation in signal processing, which gives a tractable way to solve linear, constant-coefficient difference equations. Based on the classical Z transform and inspired by the thought of sliding DFT, a new…

信号处理 · 电气工程与系统科学 2018-08-21 Peng-fei Xu , Yin-jie Jia , Zhi-jian Wang

We propose to combine cepstrum and nonlinear time-frequency (TF) analysis to study mutiple component oscillatory signals with time-varying frequency and amplitude and with time-varying non-sinusoidal oscillatory pattern. The concept of…

数据分析、统计与概率 · 物理学 2016-11-23 Chen-Yun Lin , Li Su , Hau-tieng Wu

This paper introduces GlOttal-flow LPC Filter (GOLF), a novel method for singing voice synthesis (SVS) that exploits the physical characteristics of the human voice using differentiable digital signal processing. GOLF employs a glottal…

音频与语音处理 · 电气工程与系统科学 2024-10-21 Chin-Yun Yu , György Fazekas

In this paper, we propose a classification based glottal closure instants (GCI) detection from pathological acoustic speech signal, which finds many applications in vocal disorder analysis. Till date, GCI for pathological disorder is…

声音 · 计算机科学 2018-11-28 Gurunath Reddy M , Tanumay Mandal , Krothapalli Sreenivasa Rao

Conventional Frequency Domain Linear Prediction (FDLP) technique models the squared Hilbert envelope of speech with varied degrees of approximation which can be sampled at the required frame rate and used as features for Automatic Speech…

声音 · 计算机科学 2022-04-04 Samik Sadhu , Hynek Hermansky
‹ 上一页 1 2 3 10 下一页 ›