中文
相关论文

相关论文: On the Mutual Information between Source and Filte…

200 篇论文

Many voice disorders induce subharmonic phonation, but voice signal analysis is currently lacking a technique to detect the presence of subharmonics reliably. Distinguishing subharmonic phonation from normal phonation is a challenging task…

音频与语音处理 · 电气工程与系统科学 2025-01-17 Takeshi Ikuma , Melda Kunduk , Brad Story , Andrew J. McWhorter

Dysarthria is a speech disorder characterized by impaired intelligibility and reduced communicative effectiveness. Automatic dysarthria assessment provides a scalable, cost-effective approach for supporting the diagnosis and treatment of…

音频与语音处理 · 电气工程与系统科学 2025-11-05 Kaimeng Jia , Minzhu Tu , Zengrui Jin , Siyin Wang , Chao Zhang

Automatic conflict detection has grown in relevance with the advent of body-worn technology, but existing metrics such as turn-taking and overlap are poor indicators of conflict in police-public interactions. Moreover, standard techniques…

音频与语音处理 · 电气工程与系统科学 2018-02-15 Alistair Letcher , Jelena Trišović , Collin Cademartori , Xi Chen , Jason Xu

Speech-based clinical tools are increasingly deployed in multilingual settings, yet whether pathological speech markers remain geometrically separable from accent variation remains unclear. Systems may misclassify healthy non-native…

声音 · 计算机科学 2026-02-25 Bipasha Kashyap , Pubudu N. Pathirana

In healthy-to-pathological voice conversion (H2P-VC), healthy speech is converted into pathological while preserving the identity. The paper improves on previous two-stage approach to H2P-VC where (1) speech is created first with the…

This paper presents a novel approach towards identification of human beings from the statistical analysis of their lip prints. Lip features are extracted by studying the spatial orientations of the grooves present in lip prints of…

计算机视觉与模式识别 · 计算机科学 2013-12-04 Saptarshi Bhattacharjee , S Arunkumar , Samir Kumar Bandyopadhyay

Current computational-emotion research has focused on applying acoustic properties to analyze how emotions are perceived mathematically or used in natural language processing machine learning models. While recent interest has focused on…

声音 · 计算机科学 2021-07-06 Daniel Szelogowski

Changes in speech and language are among the first signs of Parkinson's disease (PD). Thus, clinicians have tried to identify individuals with PD from their voices for years. Doctors can leverage AI-based speech assessments to spot PD…

声音 · 计算机科学 2023-12-06 Mahboobeh Parsapoor

On average the lack of biological markers causes a one year diagnostic delay to detect amyotrophic lateral sclerosis (ALS). To improve the diagnostic process an automatic voice assessment based on acoustic analysis can be used. The purpose…

声音 · 计算机科学 2020-03-25 Maxim Vashkevich , Alexander Petrovsky , Yuliya Rushkevich

The pseudo-periodicity of voiced speech can be exploited in several speech processing applications. This requires however that the precise locations of the Glottal Closure Instants (GCIs) are available. The focus of this paper is the…

声音 · 计算机科学 2020-01-03 Thomas Drugman , Mark Thomas , Jon Gudnason , Patrick Naylor , Thierry Dutoit

Prior studies in the automatic classification of voice quality have mainly studied the use of the acoustic speech signal as input. Recently, a few studies have been carried out by jointly using both speech and neck surface accelerometer…

音频与语音处理 · 电气工程与系统科学 2023-08-08 Sudarsana Reddy Kadiri , Farhad Javanmardi , Paavo Alku

The modeling of speech production often relies on a source-filter approach. Although methods parameterizing the filter have nowadays reached a certain maturity, there is still a lot to be gained for several speech processing applications in…

声音 · 计算机科学 2020-01-07 Thomas Drugman , Thierry Dutoit

The time domain waveform of a speech signal carries all of the auditory information. From the phonological point of view, it little can be said on the basis of the waveform itself. However, past research in mathematics, acoustics, and…

声音 · 计算机科学 2013-05-07 Urmila Shrawankar , V M Thakare

Voiced sounds involve self-sustained vocal folds oscillations due to the interaction between the airflow and the vocal folds. Common vocal folds pathologies like polyps and anatomical asymmetry degrade the mechanical vocal fold properties…

经典物理 · 物理学 2007-10-24 Nicolas Ruty , Claire Brutel , Xavier Pelorson , Annemie Van Hirtum

We conducted a comprehensive analysis of an Automatic Voice Disorders Detection (AVDD) system using existing voice disorder datasets with available demographic metadata. The study involved analysing system performance across various…

音频与语音处理 · 电气工程与系统科学 2025-04-15 Mariel Estevez , Cyntia Bonomi , Dayana Ribas , Alfonso Ortega , Luciana Ferrer

Acoustic-to-articulatory speech inversion could enhance automated clinical mispronunciation detection to provide detailed articulatory feedback unattainable by formant-based mispronunciation detection algorithms; however, it is unclear the…

音频与语音处理 · 电气工程与系统科学 2023-08-22 Nina R Benway , Yashish M Siriwardena , Jonathan L Preston , Elaine Hitchcock , Tara McAllister , Carol Espy-Wilson

The physical understanding of a method of detecting mammalian cancer via vocalization during a normal echo-Doppler test is provided. The backscattered ultrasound frequency in the case of a vocal humming resonating in the chest wall is…

Intent is defined for understanding spoken language in existing works. Both textual features and acoustic features involved in medical speech contain intent, which is important for symptomatic diagnosis. In this paper, we propose a medical…

人工智能 · 计算机科学 2024-05-01 Jianzong Wang , Pengcheng Li , Xulong Zhang , Ning Cheng , Jing Xiao

Automatic assessment of dysarthria remains a highly challenging task due to high variability in acoustic signals and the limited data. Currently, research on the automatic assessment of dysarthria primarily focuses on two approaches: one…

音频与语音处理 · 电气工程与系统科学 2024-05-08 Xiaokang Liu , Xiaoxia Du , Juan Liu , Rongfeng Su , Manwa Lawrence Ng , Yumei Zhang , Yudong Yang , Shaofeng Zhao , Lan Wang , Nan Yan

The analysis, processing, and extraction of meaningful information from sounds all around us is the subject of the broader area of audio analytics. Audio captioning is a recent addition to the domain of audio analytics, a cross-modal…

音频与语音处理 · 电气工程与系统科学 2023-05-04 Sandeep Kothinti , Dimitra Emmanouilidou