English
Related papers

Related papers: Domain-Agnostic Causal-Aware Audio Transformer for…

200 papers

The issue of domain shift remains a problematic phenomenon in most real-world datasets and clinical audio is no exception. In this work, we study the nature of domain shift in a clinical database of infant cry sounds acquired across…

Audio and Speech Processing · Electrical Eng. & Systems 2023-12-04 Charles C. Onu , Hemanth K. Sheetha , Arsenii Gorin , Doina Precup

This paper addresses a major challenge in acoustic event detection, in particular infant cry detection in the presence of other sounds and background noises: the lack of precise annotated data. We present two contributions for supervised…

Infant cry detection is a crucial component of baby care system. In this paper, we propose a lightweight and robust method for infant cry detection. The method leverages blueprint separable convolutions to reduce computational complexity,…

Sound · Computer Science 2025-08-28 Haolin Yu , Yanxiong Li

Despite continuing medical advances, the rate of newborn morbidity and mortality globally remains high, with over 6 million casualties every year. The prediction of pathologies affecting newborns based on their cry is thus of significant…

Machine Learning · Computer Science 2020-03-20 Charles C. Onu , Jonathan Lebensold , William L. Hamilton , Doina Precup

This thesis addresses the technical challenges of applying machine learning to understand and interpret medical audio signals. The sounds of our lungs, heart, and voice convey vital information about our health. Yet, in contemporary…

Sound · Computer Science 2025-06-18 Charles C Onu

A promising approach for steering auditory attention in complex listening environments relies on Auditory Attention Decoding (AAD), which aim to identify the attended speech stream in a multiple speaker scenario from neural recordings.…

Unsupervised Domain Adaptation (UDA) seeks to transfer knowledge from a labeled source domain to an unlabeled target domain but often suffers from severe domain and scale gaps that degrade performance. Existing cross-attention-based…

Computer Vision and Pattern Recognition · Computer Science 2026-03-19 Zelin Zang , Yehui Yang , Fei Wang , Liangyu Li , Baigui Sun

Domain shift is considered a challenge in machine learning as it causes significant degradation of model performance. In the Acoustic Scene Classification task (ASC), domain shift is mainly caused by different recording devices. Several…

Sound · Computer Science 2023-06-16 Shahed Masoudian , Khaled Koutini , Markus Schedl , Gerhard Widmer , Navid Rekabsaz

EEG-based tinnitus classification is a valuable tool for tinnitus diagnosis, research, and treatments. Most current works are limited to a single dataset where data patterns are similar. But EEG signals are highly non-stationary, resulting…

Signal Processing · Electrical Eng. & Systems 2022-11-08 Yun Li , Zhe Liu , Lina Yao , Jessica J. M. Monaghan , David McAlpine

In a recent acoustic scene classification (ASC) research field, training and test device channel mismatch have become an issue for the real world implementation. To address the issue, this paper proposes a channel domain conversion using…

Sound · Computer Science 2018-12-06 Seongkyu Mun , Suwon Shon

Diagnostic procedures for ASD (autism spectrum disorder) involve semi-naturalistic interactions between the child and a clinician. Computational methods to analyze these sessions require an end-to-end speech and language processing pipeline…

Audio and Speech Processing · Electrical Eng. & Systems 2019-10-28 Rimita Lahiri , Manoj Kumar , Somer Bishop , Shrikanth Narayanan

Human Activity Recognition (HAR) plays a crucial role in various applications such as human-computer interaction and healthcare monitoring. However, challenges persist in HAR models due to the data distribution differences between training…

Machine Learning · Computer Science 2024-09-04 Xiaozhou Ye , Kevin I-Kai Wang

Despite recent success, most contrastive self-supervised learning methods are domain-specific, relying heavily on data augmentation techniques that require knowledge about a particular domain, such as image cropping and rotation. To…

Machine Learning · Computer Science 2021-07-21 Vikas Verma , Minh-Thang Luong , Kenji Kawaguchi , Hieu Pham , Quoc V. Le

Domain mismatch is a noteworthy issue in acoustic event detection tasks, as the target domain data is difficult to access in most real applications. In this study, we propose a novel CNN-based discriminative training framework as a domain…

Audio and Speech Processing · Electrical Eng. & Systems 2021-03-29 Tiantian Tang , Xinyuan Zhou , Yanhua Long , Yijie Li , Jiaen Liang

Understanding the meaning of infant cries is a significant challenge for young parents in caring for their newborns. The presence of background noise and the lack of labeled data present practical challenges in developing systems that can…

Sound · Computer Science 2025-02-05 Mengze Hong , Chen Jason Zhang , Lingxiao Yang , Yuanfeng Song , Di Jiang

Neonatal pain assessment in clinical environments is challenging as it is discontinuous and biased. Facial/body occlusion can occur in such settings due to clinical condition, developmental delays, prone position, or other external factors.…

Computer Vision and Pattern Recognition · Computer Science 2020-03-25 Md Sirajus Salekin , Ghada Zamzmi , Rahul Paul , Dmitry Goldgof , Rangachar Kasturi , Thao Ho , Yu Sun

In this study, we propose a novel noise adaptive speech enhancement (SE) system, which employs a domain adversarial training (DAT) approach to tackle the issue of a noise type mismatch between the training and testing conditions. Such a…

Sound · Computer Science 2019-07-02 Chien-Feng Liao , Yu Tsao , Hung-Yi Lee , Hsin-Min Wang

In acoustic signal processing, the target signals usually carry semantic information, which is encoded in a hierarchal structure of short and long-term contexts. However, the background noise distorts these structures in a nonuniform way.…

Audio and Speech Processing · Electrical Eng. & Systems 2022-01-26 Tassadaq Hussain , Wei-Chien Wang , Mandar Gogate , Kia Dashtipour , Yu Tsao , Xugang Lu , Adeel Ahsan , Amir Hussain

Infant cry emotion recognition is crucial for parenting and medical applications. It faces many challenges, such as subtle emotional variations, noise interference, and limited data. The existing methods lack the ability to effectively…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-24 Junyu Zhou , Yanxiong Li , Haolin Yu

Speech recognition systems have improved dramatically over the last few years, however, their performance is significantly degraded for the cases of accented or impaired speech. This work explores domain adversarial neural networks (DANN)…

Sound · Computer Science 2020-10-09 Dominika Woszczyk , Stavros Petridis , David Millard
‹ Prev 1 2 3 10 Next ›