中文
相关论文

相关论文: Pathological Voice Classification Using Mel-Cepstr…

200 篇论文

AI-based voice analysis shows promise for disease diagnostics, but existing classifiers often fail to accurately identify specific pathologies because of gender-related acoustic variations and the scarcity of data for rare diseases. We…

声音 · 计算机科学 2025-08-05 Fan Wu , Kaicheng Zhao , Elgar Fleisch , Filipe Barata

Voice disorders significantly impact patient quality of life, yet non-invasive automated diagnosis remains under-explored due to both the scarcity of pathological voice data, and the variability in recording sources. This work introduces…

Speaker Verification (SV) systems involve mainly two individual stages: feature extraction and classification. In this paper, we explore these two modules with the aim of improving the performance of a speaker verification system under…

音频与语音处理 · 电气工程与系统科学 2024-02-06 Kerlos Atia Abdalmalak , Ascensión Gallardo-Antol'in

An important step in speaker verification is extracting features that best characterize the speaker voice. This paper investigates a front-end processing that aims at improving the performance of speaker verification based on the SVMs…

机器学习 · 计算机科学 2013-06-13 Kawthar Yasmine Zergat , Abderrahmane Amrouche

Voice disorders affect patients profoundly, and acoustic tools can potentially measure voice function objectively. Nonetheless, existing tools are limited to analysing voices displaying near periodicity, and do not account for inherent…

元胞自动机与格子气 · 物理学 2019-10-23 Max A Little , Patrick E McSharry , Stephen J Roberts , Declan AE Costello , Irene M Moroz

We discuss how vocal disorders can be post-corrected via a simple nonlinear noise reduction scheme. This work is motivated by the need of a better understanding of voice dysfunctions. This would entail a twofold advantage for affected…

凝聚态物理 · 物理学 2016-08-31 Lorenzo Matassini , Claudia Manfredi

Recent years has witnessed an increase in technologies that use speech for the sensing of the health of the talker. This survey paper proposes a general taxonomy of the technologies and a broad overview of current progress and challenges.…

This study investigates explainable machine learning algorithms for identifying depression from speech. Grounded in evidence from speech production that depression affects motor control and vowel generation, pre-trained vowel-based…

机器学习 · 计算机科学 2024-10-25 Kexin Feng , Theodora Chaspari

Depression, as a typical mental disorder, has become a prevalent issue significantly impacting public health. However, the prevention and treatment of depression still face multiple challenges, including complex diagnostic procedures,…

计算机视觉与模式识别 · 计算机科学 2025-10-28 Yu Luo , Nan Huang , Sophie Yu , Hendry Xu , Jerry Wang , Colin Wang , Zhichao Liu , Chen Zeng

Recognizing human non-speech vocalizations is an important task and has broad applications such as automatic sound transcription and health condition monitoring. However, existing datasets have a relatively small number of vocal sound…

声音 · 计算机科学 2022-06-22 Yuan Gong , Jin Yu , James Glass

We conducted a comprehensive analysis of an Automatic Voice Disorders Detection (AVDD) system using existing voice disorder datasets with available demographic metadata. The study involved analysing system performance across various…

音频与语音处理 · 电气工程与系统科学 2025-04-15 Mariel Estevez , Cyntia Bonomi , Dayana Ribas , Alfonso Ortega , Luciana Ferrer

More than 90% of the Parkinson Disease (PD) patients suffer from vocal disorders. Speech impairment is already indicator of PD. This study focuses on PD diagnosis through voiceprint features. In this paper, a method based on Deep Neural…

声音 · 计算机科学 2018-12-18 Zhijing Xu , Juan Wang , Ying Zhang , Xiangjian He

Psychiatric illnesses are often associated with multiple symptoms, whose severity must be graded for accurate diagnosis and treatment. This grading is usually done by trained clinicians based on human observations and judgments made within…

声音 · 计算机科学 2017-03-17 Rita Singh , Justin Baker , Luciana Pennant , Louis-Philippe Morency

The physical understanding of a method of detecting mammalian cancer via vocalization during a normal echo-Doppler test is provided. The backscattered ultrasound frequency in the case of a vocal humming resonating in the chest wall is…

Voice disorders affect an estimated 14 million working-aged Americans, and many more worldwide. We present the first large scale study of vocal misuse based on long-term ambulatory data collected by an accelerometer placed on the neck. We…

This paper presents a new approach for classification of dysfluent and fluent speech using Mel-Frequency Cepstral Coefficient (MFCC). The speech is fluent when person's speech flows easily and smoothly. Sounds combine into syllable,…

声音 · 计算机科学 2013-01-10 P. Mahesha , D. S. Vinod

Voice signals originating from the respiratory tract are utilized as valuable acoustic biomarkers for the diagnosis and assessment of respiratory diseases. Among the employed acoustic features, Mel Frequency Cepstral Coefficients (MFCC) is…

Objective speech disorder classification for speakers with communication difficulty is desirable for diagnosis and administering therapy. With the current state of speech technology, it is evident to propose neural networks for this…

音频与语音处理 · 电气工程与系统科学 2021-06-15 Jinzi Qi , Hugo Van hamme

Clinical characterization and interpretation of respiratory sound symptoms have remained a challenge due to the similarities in the audio properties that manifest during auscultation in medical diagnosis. The misinterpretation and…

系统与控制 · 电气工程与系统科学 2021-10-18 Chinazunwa Uwaoma , Gunjan Mansingh

While the use of deep neural networks has significantly boosted speaker recognition performance, it is still challenging to separate speakers in poor acoustic environments. Here speech enhancement methods have traditionally allowed improved…

音频与语音处理 · 电气工程与系统科学 2020-08-28 Yanpei Shi , Qiang Huang , Thomas Hain