中文
相关论文

相关论文: Discriminating between Nasal and Mouth Breathing

200 篇论文

Acoustics-to-word models are end-to-end speech recognizers that use words as targets without relying on pronunciation dictionaries or graphemes. These models are notoriously difficult to train due to the lack of linguistic knowledge. It is…

音频与语音处理 · 电气工程与系统科学 2018-11-14 Hao Tang , James Glass

Articulatory-to-acoustic (forward) mapping is a technique to predict speech using various articulatory acquisition techniques as input (e.g. ultrasound tongue imaging, MRI, lip video). The advantage of lip video is that it is easily…

计算机视觉与模式识别 · 计算机科学 2021-04-30 Frigyes Viktor Arthur , Tamás Gábor Csapó

Topical intra-nasal sprays are amongst the most commonly prescribed therapeutic options for sinonasal diseases in humans. However, inconsistency and ambiguity in instructions show a lack of definitive knowledge on best spray use techniques.…

This paper presents and explores a robust deep learning framework for auscultation analysis. This aims to classify anomalies in respiratory cycles and detect disease, from respiratory sound recordings. The framework begins with front-end…

音频与语音处理 · 电气工程与系统科学 2020-06-04 Lam Pham , Huy Phan , Ramaswamy Palaniappan , Alfred Mertins , Ian McLoughlin

Respiratory sound classification (RSC) is challenging due to varied acoustic signatures, primarily influenced by patient demographics and recording environments. To address this issue, we introduce a text-audio multimodal model that…

声音 · 计算机科学 2024-06-17 June-Woo Kim , Miika Toikkanen , Yera Choi , Seoung-Eun Moon , Ho-Young Jung

When human listeners try to guess the spatial position of a speech source, they are influenced by the speaker's production level, regardless of the intensity level reaching their ears. Because the perception of distance is a very difficult…

机器人学 · 计算机科学 2023-05-23 Ambre Davat , Véronique Aubergé , Gang Feng

Despite the remarkable advances in deep learning technology, achieving satisfactory performance in lung sound classification remains a challenge due to the scarcity of available data. Moreover, the respiratory sound samples are collected…

声音 · 计算机科学 2023-12-18 June-Woo Kim , Sangmin Bae , Won-Yang Cho , Byungjo Lee , Ho-Young Jung

The velopharyngeal (VP) valve regulates the opening between the nasal and oral cavities. This valve opens and closes through a coordinated motion of the velum and pharyngeal walls. Nasalance is an objective measure derived from the oral and…

音频与语音处理 · 电气工程与系统科学 2023-06-02 Yashish M. Siriwardena , Carol Espy-Wilson , Suzanne Boyce , Mark K. Tiede , Liran Oren

Our sense of smell relies on sensitive, selective atomic-scale processes that are initiated when a scent molecule meets specific receptors in the nose. However, the physical mechanisms of detection are not clear. While odorant shape and…

生物物理 · 物理学 2015-06-26 Jennifer C. Brookes , Filio Hartoutsiou , A. P. Horsfield , A. M. Stoneham

Interpersonal spoken communication is central to human interaction and the exchange of information. Such interactive processes involve not only speech and spoken language but also non-verbal cues such as hand gestures, facial expressions,…

声音 · 计算机科学 2022-12-20 Tiantian Feng , Shrikanth Narayanan

This paper is focused on nonlinear prediction coding, which consists on the prediction of a speech sample based on a nonlinear combination of previous samples. It is known that in the generation of the glottal pulse, the wave equation does…

声音 · 计算机科学 2022-04-01 Marcos Faundez-Zanuy , Enric Monte , Francesc Vallverdú

Prosody transfer is well-studied in the context of expressive speech synthesis. Cross-lingual prosody transfer, however, is challenging and has been under-explored to date. In this paper, we present a novel solution to learn prosody…

音频与语音处理 · 电气工程与系统科学 2023-06-21 Jakub Swiatkowski , Duo Wang , Mikolaj Babianski , Patrick Lumban Tobing , Ravichander Vipperla , Vincent Pollet

We propose a Perceiver-based sequence classifier to detect abnormalities in speech reflective of several neurological disorders. We combine this classifier with a Universal Speech Model (USM) that is trained (unsupervised) on 12 million…

Nasalization of vowels is a phenomenon where oral and nasal tracts participate simultaneously for the production of speech. Acoustic coupling of oral and nasal tracts results in a complex production system, which is subjected to a…

声音 · 计算机科学 2020-09-15 RaviShankar Prasad , B. Yegnanarayana

Automated respiratory sound classification faces practical challenges from background noise and insufficient denoising in existing systems. We propose Adaptive Differential Denoising network, that integrates noise suppression and…

音频与语音处理 · 电气工程与系统科学 2025-06-04 Gaoyang Dong , Zhicheng Zhang , Ping Sun , Minghui Zhang

In the healthcare industry, researchers have been developing machine learning models to automate diagnosing patients with respiratory illnesses based on their breathing patterns. However, these models do not consider the demographic biases,…

机器学习 · 计算机科学 2025-01-10 Rachel Pfeifer , Sudip Vhaduri , James Eric Dietz

Multivariate data analysis and machine-learning classification become popular tools to extract features without physical models for complex environments recognition. For electronic noses, time sampling over multiple sensors must be a fair…

软凝聚态物质 · 物理学 2023-08-25 Wiem Haj Ammar , Aicha Boujnah , Aimen Boubaker , Adel Kalboussi , Kamal Lmimouni , Sébastien Pecqueur

Communicative gestures and speech acoustic are tightly linked. Our objective is to predict the timing of gestures according to the acoustic. That is, we want to predict when a certain gesture occurs. We develop a model based on a recurrent…

人机交互 · 计算机科学 2021-04-27 Fajrian Yunus , Chloé Clavel , Catherine Pelachaud

Our native language influences the way we perceive speech sounds, affecting our ability to discriminate non-native sounds. We compare two ideas about the influence of the native language on speech perception: the Perceptual Assimilation…

计算与语言 · 计算机科学 2022-06-01 Juliette Millet , Ioana Chitoran , Ewan Dunbar

Body sounds provide rich information about the state of the human body and can be useful in many medical applications. Auscultation, the practice of listening to body sounds, has been used for centuries in respiratory and cardiac medicine…

人机交互 · 计算机科学 2020-08-13 Shyam A. Tailor , Jagmohan Chauhan , Cecilia Mascolo