English
Related papers

Related papers: Fused Audio Instance and Representation for Respir…

200 papers

Robust strategies for Alzheimer's disease (AD) detection are important, given the high prevalence of AD. In this paper, we study the performance and generalizability of three approaches for AD detection from speech on the recent ADReSSo…

Computation and Language · Computer Science 2022-09-16 Aparna Balagopalan , Jekaterina Novikova

Traditional echocardiographic parameters such as ejection fraction (EF) and global longitudinal strain (GLS) have limitations in the early detection of cardiac dysfunction. EF often remains normal despite underlying pathology, and GLS is…

Machine Learning · Computer Science 2025-07-21 Beka Begiashvili , Carlos J. Fernandez-Candel , Matías Pérez Paredes

Cardiovascular diseases are the leading cause of deaths and severely threaten human health in daily life. On the one hand, there have been dramatically increasing demands from both the clinical practice and the smart home application for…

Sound · Computer Science 2021-01-14 Zhao Ren , Kun Qian , Fengquan Dong , Zhenyu Dai , Yoshiharu Yamamoto , Björn W. Schuller

As deepfake audio becomes more realistic and diverse, developing generalizable countermeasure systems has become crucial. Existing detection methods primarily depend on XLS-R front-end features to improve generalization. Nonetheless, their…

Sound · Computer Science 2026-02-17 Zhe Ye , Xiangui Kang , Jiayi He , Chengxin Chen , Wei Zhu , Kai Wu , Yin Yang , Jiwu Huang

Rapidly scaling screening, testing and quarantine has shown to be an effective strategy to combat the COVID-19 pandemic. We consider the application of deep learning techniques to distinguish individuals with COVID from non-COVID by using…

Unsupervised anomalous sound detection aims to detect unknown abnormal sounds of machines from normal sounds. However, the state-of-the-art approaches are not always stable and perform dramatically differently even for machines of the same…

Sound · Computer Science 2022-05-02 Youde Liu , Jian Guan , Qiaoxi Zhu , Wenwu Wang

Tuberculosis (TB) is an infectious disease caused by the bacterium Mycobacterium tuberculosis and primarily affects the lungs, as well as other body parts. TB is spread through the air when an infected person coughs, sneezes, or talks.…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-04 George P. Kafentzis , Stephane Tetsing , Joe Brew , Lola Jover , Mindaugas Galvosas , Carlos Chaccour , Peter M. Small

AI-driven respiratory sound classification (RSC) is promising for automated pulmonary disease detection, yet multi-site deployment is hindered by inter-stethoscope variability. We introduce a federated domain generalization (FedDG)…

Audio and Speech Processing · Electrical Eng. & Systems 2026-05-29 Heejoon Koo , Yoon Tae Kim , Miika Toikkanen , June-Woo Kim

Lung auscultation is the most effective and indispensable method for diagnosing various respiratory disorders by using the sounds from the airways during inspirium and exhalation using a stethoscope. In this study, the statistical features…

Sound · Computer Science 2021-01-22 Gökhan Altan , Yakup Kutlu , Adnan Özhan Pekmezci , Serkan Nural

Accurate and efficient auscultation-based diagnostics are vital for early disease detection, especially in resource-limited settings where specialized clinical expertise is scarce. Traditional auscultation, which heavily depends on…

Sound · Computer Science 2025-03-26 Pingjie Wang , Liudan Zhao , Zihan Zhao , Miao He , Xin Sun , Ya Zhang , Kun Sun , Yanfeng Wang , Yu Wang

Auscultation is the most efficient way to diagnose cardiovascular and respiratory diseases. To reach accurate diagnoses, a device must be able to recognize heart and lung sounds from various clinical situations. However, the recorded chest…

Audio and Speech Processing · Electrical Eng. & Systems 2020-12-14 Kun-Hsi Tsai , Wei-Chien Wang , Chui-Hsuan Cheng , Chan-Yen Tsai , Jou-Kou Wang , Tzu-Hao Lin , Shih-Hau Fang , Li-Chin Chen , Yu Tsao

This paper presents a macroscopic approach to automatic detection of speech sound disorder (SSD) in child speech. Typically, SSD is manifested by persistent articulation and phonological errors on specific phonemes in the language. The…

Audio and Speech Processing · Electrical Eng. & Systems 2022-06-30 Si-Ioi Ng , Cymie Wing-Yee Ng , Jiarui Wang , Tan Lee

The SARS-CoV-2 coronavirus emerged in 2019, causing a COVID-19 pandemic that resulted in 7 million deaths out of 770 million reported cases over the next four years. The global health emergency called for unprecedented efforts to monitor…

There have been several documented outbreaks of COVID-19 associated with vocalization, either by speech or by singing, in indoor confined spaces. Here, we model the risk of in-room airborne disease transmission via expiratory particle…

Quantitative Methods · Quantitative Biology 2020-09-10 Santiago Barreda , Sima Asadi , Christopher D. Cappa , Anthony S. Wexler , Nicole M. Bouvier , William D. Ristenpart

Humans use context to assess the veracity of information. However, current audio deepfake detectors only analyze the audio file without considering either context or transcripts. We create and analyze a Journalist-provided Deepfake Dataset…

Breathing is an essential part of human survival, which carries information about a person's physiological and psychological state. Generally, breath boundaries are marked by experts before using for any task. An unsupervised algorithm for…

Audio and Speech Processing · Electrical Eng. & Systems 2023-04-10 Shivani Yadav , Dipanjan Gope , Uma Maheswari K. , Prasanta Kumar Ghosh

The process of human speech production involves coordinated respiratory action to elicit acoustic speech signals. Typically, speech is produced when air is forced from the lungs and is modulated by the vocal tract, where such actions are…

Depression is a common mental disorder which has been affecting millions of people around the world and becoming more severe with the arrival of COVID-19. Nevertheless proper diagnosis is not accessible in many regions due to a severe…

Speech-based depression detection tools could aid early screening. Here, we propose an interpretable speech foundation model approach to enhance the clinical applicability of such tools. We introduce a speech-level Audio Spectrogram…

Sound · Computer Science 2026-03-26 Qingkun Deng , Saturnino Luz , Sofia de la Fuente Garcia

Audio deepfakes are increasingly in-differentiable from organic speech, often fooling both authentication systems and human listeners. While many techniques use low-level audio features or optimization black-box model training, focusing on…

Sound · Computer Science 2025-02-21 Kevin Warren , Daniel Olszewski , Seth Layton , Kevin Butler , Carrie Gates , Patrick Traynor
‹ Prev 1 8 9 10 Next ›