English
Related papers

Related papers: ICSD: An Open-source Dataset for Infant Cry and Sn…

200 papers

Understanding the meaning of infant cries is a significant challenge for young parents in caring for their newborns. The presence of background noise and the lack of labeled data present practical challenges in developing systems that can…

Sound · Computer Science 2025-02-05 Mengze Hong , Chen Jason Zhang , Lingxiao Yang , Yuanfeng Song , Di Jiang

Infant cry detection is a crucial component of baby care system. In this paper, we propose a lightweight and robust method for infant cry detection. The method leverages blueprint separable convolutions to reduce computational complexity,…

Sound · Computer Science 2025-08-28 Haolin Yu , Yanxiong Li

This paper addresses a major challenge in acoustic event detection, in particular infant cry detection in the presence of other sounds and background noises: the lack of precise annotated data. We present two contributions for supervised…

Most existing cry detection models have been tested with data collected in controlled settings. Thus, the extent to which they generalize to noisy and lived environments is unclear. In this paper, we evaluate several established machine…

Audio and Speech Processing · Electrical Eng. & Systems 2022-02-18 Xuewen Yao , Megan Micheletti , Mckensey Johnson , Edison Thomaz , Kaya de Barbaro

This paper describes the Ubenwa CryCeleb dataset - a labeled collection of infant cries - and the accompanying CryCeleb 2023 task, which is a public speaker verification challenge based on cry sounds. We released more than 6 hours of…

Sound · Computer Science 2024-03-22 David Budaghyan , Charles C. Onu , Arsenii Gorin , Cem Subakan , Doina Precup

Infant crying can serve as a crucial indicator of various physiological and emotional states. This paper introduces a comprehensive approach detecting infant cries within audio data. We integrate Wav2Vec with traditional audio features and…

Detection of baby cries is an important part of baby monitoring and health care. Almost all existing methods use supervised SVM, CNN, or their varieties. In this work, we propose to use weakly supervised anomaly detection to detect a baby…

Computer Vision and Pattern Recognition · Computer Science 2023-11-28 Weijun Tan , Qi Yao , Jingfeng Liu

Background: Infant cry acoustics provide a promising window into early neurodevelopment and may serve as scalable biomarkers for neurodevelopmental disorders. However, conventional microphone-based recordings are highly susceptible to…

Sound · Computer Science 2026-05-28 Winko W. An , Saketh Sundar , Lisa Yankowitz , Daryush D. Mehta , Carol L. Wilkinson

This thesis addresses the technical challenges of applying machine learning to understand and interpret medical audio signals. The sounds of our lungs, heart, and voice convey vital information about our health. Yet, in contemporary…

Sound · Computer Science 2025-06-18 Charles C Onu

Since the 1960s, neonatal clinicians have known that newborns suffering from certain neurological conditions exhibit altered crying patterns such as the high-pitched cry in birth asphyxia. Despite an annual burden of over 1.5 million infant…

Infant cry emotion recognition is crucial for parenting and medical applications. It faces many challenges, such as subtle emotional variations, noise interference, and limited data. The existing methods lack the ability to effectively…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-24 Junyu Zhou , Yanxiong Li , Haolin Yu

In this paper, we explore self-supervised learning (SSL) for analyzing a first-of-its-kind database of cry recordings containing clinical indications of more than a thousand newborns. Specifically, we target cry-based detection of…

Sound · Computer Science 2023-05-03 Arsenii Gorin , Cem Subakan , Sajjad Abdoli , Junhao Wang , Samantha Latremouille , Charles Onu

Perinatal Asphyxia is one of the top three causes of infant mortality in developing countries, resulting to the death of about 1.2 million newborns every year. At its early stages, the presence of asphyxia cannot be conclusively determined…

Applications · Statistics 2018-08-28 Charles C. Onu

The detection of shouted speech is crucial in audio surveillance and monitoring. Although it is desirable for a security system to be able to identify emergencies, existing corpora provide only a binary label (i.e., shouted or normal) for…

Sound · Computer Science 2024-10-22 Takahiro Fukumori , Taito Ishida , Yoichi Yamashita

The rapid proliferation of drones across various industries has introduced significant challenges related to privacy, security, and noise pollution. Current drone detection systems, primarily based on visual and radar technologies, face…

Sound · Computer Science 2025-09-08 Mia Y. Wang , Mackenzie Linn , Andrew P. Berg , Qian Zhang

The performance of sound event detection methods can significantly degrade when they are used in unseen conditions (e.g. recording devices, ambient noise). Domain adaptation is a promising way to tackle this problem. In this paper, we…

Sound · Computer Science 2019-11-26 Shayan Gharib , Konstantinos Drossos , Eemi Fagerlund , Tuomas Virtanen

Environmental sound scene and sound event recognition is important for the recognition of suspicious events in indoor and outdoor environments (such as nurseries, smart homes, nursing homes, etc.) and is a fundamental task involved in many…

Sound · Computer Science 2023-08-31 Nan Che , Chenrui Liu , Fei Yu

Acoustic analyses of infant vocalizations are valuable for research on speech development as well as applications in sound classification. Previous studies have focused on measures of acoustic features based on theories of speech…

Sound · Computer Science 2020-05-27 Mohammad K. Ebrahimpour , Sara Schneider , David C. Noelle , Christopher T. Kello

The effectiveness of pain management relies on the choice and the correct use of suitable pain assessment tools. In the case of newborns, some of the most common tools are human-based and observational, thus affected by subjectivity and…

Applications · Statistics 2018-12-24 Davide Ricossa , Enrico Baccaglini , Elvira Di Nardo , Emilia Parodi , Riccardo Scopigno

Automatic speech recognition (ASR) has been significantly advanced with the use of deep learning and big data. However improving robustness, including achieving equally good performance on diverse speakers and accents, is still a…

Sound · Computer Science 2020-11-17 Fan Yu , Zhuoyuan Yao , Xiong Wang , Keyu An , Lei Xie , Zhijian Ou , Bo Liu , Xiulin Li , Guanqiong Miao
‹ Prev 1 2 3 10 Next ›