中文
相关论文

相关论文: Chirp Complex Cepstrum-based Decomposition for Asy…

200 篇论文

In this paper, we propose a classification based glottal closure instants (GCI) detection from pathological acoustic speech signal, which finds many applications in vocal disorder analysis. Till date, GCI for pathological disorder is…

声音 · 计算机科学 2018-11-28 Gurunath Reddy M , Tanumay Mandal , Krothapalli Sreenivasa Rao

Automatic detection of voice pathology enables objective assessment and earlier intervention for the diagnosis. This study provides a systematic analysis of glottal source features and investigates their effectiveness in voice pathology…

音频与语音处理 · 电气工程与系统科学 2023-10-18 Sudarsana Reddy Kadiri , Paavo Alku

Visual speech recognition (VSR), also known as lip reading, is the task of recognizing speech from silent video. Despite significant advancements in VSR over recent decades, most existing methods pay limited attention to real-world visual…

计算机视觉与模式识别 · 计算机科学 2025-09-29 Tianyue Wang , Shuang Yang , Shiguang Shan , Xilin Chen

When a motor-powered vessel travels past a fixed hydrophone in a multipath environment, a Lloyd's mirror constructive/destructive interference pattern is observed in the output spectrogram. The power cepstrum detects the periodic structure…

声音 · 计算机科学 2018-10-30 Eric L. Ferguson , Stefan B. Williams , Craig T. Jin

Digital technology has made possible unimaginable applications come true. It seems exciting to have a handful of tools for easy editing and manipulation, but it raises alarming concerns that can propagate as speech clones, duplicates, or…

机器学习 · 计算机科学 2021-04-13 Arun Kumar Singh , Priyanka Singh

This article introduces a novel and computationally fast model to study the association between covariates and power spectra of replicated time series. A random covariate-dependent Cram\'{e}r spectral representation and a semiparametric…

统计方法学 · 统计学 2024-07-03 Zeda Li , Yuexiao Dong

In this paper, we propose a novel family of windowing technique to compute Mel Frequency Cepstral Coefficient (MFCC) for automatic speaker recognition from speech. The proposed method is based on fundamental property of discrete time…

计算机视觉与模式识别 · 计算机科学 2015-06-05 Md. Sahidullah , Goutam Saha

Flow cytometry mainly used for detecting the characteristics of a number of biochemical substances based on the expression of specific markers in cells. It is particularly useful for detecting membrane surface receptors, antigens, ions, or…

机器学习 · 计算机科学 2023-03-17 Yanhua Xu

This paper presents a fully automated approach for identifying speech anomalies from voice recordings to aid in the assessment of speech impairments. By combining Connectionist Temporal Classification (CTC) and encoder-decoder-based…

声音 · 计算机科学 2023-08-04 Laurin Wagner , Mario Zusag , Theresa Bloder

Articulatory acoustic inversion aims to reconstruct the complete geometry of the vocal tract from the speech signal. In this paper, we present a comparative study of several levels of phonetic segmentation accuracy, together with a…

音频与语音处理 · 电气工程与系统科学 2026-03-13 Sofiane Azzouz , Pierre-André Vuissoz , Yves Laprie

This article investigates into recently emerging approaches that use deep neural networks for the estimation of glottal closure instants (GCI). We build upon our previous approach that used synthetic speech exclusively to create perfectly…

音频与语音处理 · 电气工程与系统科学 2020-03-04 Frederik Bous , Luc Ardaillon , Axel Roebel

Background: Infant cry acoustics provide a promising window into early neurodevelopment and may serve as scalable biomarkers for neurodevelopmental disorders. However, conventional microphone-based recordings are highly susceptible to…

声音 · 计算机科学 2026-05-28 Winko W. An , Saketh Sundar , Lisa Yankowitz , Daryush D. Mehta , Carol L. Wilkinson

In this paper, we propose a novel separation system for extracting two speech signals from two microphone recordings. Our system combines the blind source separation technique with cepstral smoothing of binary time-frequency masks. The last…

声音 · 计算机科学 2026-03-17 Ibrahim Missaoui , Zied Lachiri

We propose to combine cepstrum and nonlinear time-frequency (TF) analysis to study mutiple component oscillatory signals with time-varying frequency and amplitude and with time-varying non-sinusoidal oscillatory pattern. The concept of…

数据分析、统计与概率 · 物理学 2016-11-23 Chen-Yun Lin , Li Su , Hau-tieng Wu

Hundreds of applications utilize frequency response characterization of a system. Identification of frequency response requires long experimentation time, use of transformation techniques and other difficulties associated with isolating the…

系统与控制 · 电气工程与系统科学 2021-03-22 Resmi Suresh , Raghunathan Rengaswamy

Most speech enhancement algorithms make use of the short-time Fourier transform (STFT), which is a simple and flexible time-frequency decomposition that estimates the short-time spectrum of a signal. However, the duration of short STFT…

声音 · 计算机科学 2015-09-03 Scott Wisdom , Thomas Powers , Les Atlas , James Pitton

The analysis of chaotic signals with time-frequency methods is considered. For this purpose, two new transformations are presented which consist in the decomposition of a signal onto an orthogonal set of respectively linear and hyperbolic…

混沌动力学 · 物理学 2010-07-28 Benjamin Ricaud , Francoise Briolle , F. Clairet

The electroencephalography (EEG) signals recorded in parallel with speech are used to perform isolated and continuous speech recognition. During speaking process, one also hears his or her own speech and this speech perception is also…

音频与语音处理 · 电气工程与系统科学 2020-06-03 Gautam Krishna , Co Tran , Mason Carnahan , Ahmed Tewfik

In this paper we introduce attention-regression model to demonstrate predicting acoustic features from electroencephalography (EEG) features recorded in parallel with spoken sentences. First we demonstrate predicting acoustic features…

音频与语音处理 · 电气工程与系统科学 2020-05-05 Gautam Krishna , Co Tran , Mason Carnahan , Ahmed Tewfik

Synthesized speech is common today due to the prevalence of virtual assistants, easy-to-use tools for generating and modifying speech signals, and remote work practices. Synthesized speech can also be used for nefarious purposes, including…

声音 · 计算机科学 2022-05-05 Emily R. Bartusiak , Edward J. Delp