English
Related papers

Related papers: Joint Scattering for Automatic Chick Call Recognit…

200 papers

In nature and engineering world, the acquired signals are usually affected by multiple complicated factors and appear as multicomponent nonstationary modes. In such and many other situations, it is necessary to separate these signals into a…

Signal Processing · Electrical Eng. & Systems 2021-10-14 Lin Li , Ningning Han , Qingtang Jiang , Charles K. Chui

Background: Infant cry acoustics provide a promising window into early neurodevelopment and may serve as scalable biomarkers for neurodevelopmental disorders. However, conventional microphone-based recordings are highly susceptible to…

Sound · Computer Science 2026-05-28 Winko W. An , Saketh Sundar , Lisa Yankowitz , Daryush D. Mehta , Carol L. Wilkinson

In young animals like poultry chicks (Gallus gallus), vocalisations convey information about affective and behavioural states. Traditional approaches to vocalisation analysis, relying on manual annotation and predefined categories,…

WARNING: This paper contains content that maybe upsetting or offensive to some readers. Dog whistles are coded expressions with dual meanings: one intended for the general public (outgroup) and another that conveys a specific message to an…

Computation and Language · Computer Science 2025-02-18 Kuleen Sasse , Carlos Aguirre , Isabel Cachola , Sharon Levy , Mark Dredze

Anomaly detection in multivariate time series is challenging as heterogeneous subsequence anomalies may occur. Reconstruction-based methods, which focus on learning normal patterns in the frequency domain to detect diverse abnormal…

Machine Learning · Computer Science 2025-05-09 Xingjian Wu , Xiangfei Qiu , Zhengyu Li , Yihang Wang , Jilin Hu , Chenjuan Guo , Hui Xiong , Bin Yang

To improve the performance of speaker identification systems, an effective and robust method is proposed to extract speech features, capable of operating in noisy environment. Based on the time-frequency multi-resolution property of wavelet…

Sound · Computer Science 2010-03-31 Mahmoud I. Abdalla , Hanaa S. Ali

Coughing is a typical symptom of COVID-19. To detect and localize coughing sounds remotely, a convolutional neural network (CNN) based deep learning model was developed in this work and integrated with a sound camera for the visualization…

Audio and Speech Processing · Electrical Eng. & Systems 2022-06-16 Gyeong-Tae Lee , Hyeonuk Nam , Seong-Hu Kim , Sang-Min Choi , Youngkey Kim , Yong-Hwa Park

In many application settings involving networks, such as messages between users of an on-line social network or transactions between traders in financial markets, the observed data consist of timestamped relational events, which form a…

Social and Information Networks · Computer Science 2020-11-11 Makan Arastuie , Subhadeep Paul , Kevin S. Xu

This paper introduces the T23 team's system submitted to the Singing Voice Conversion Challenge 2023. Following the recognition-synthesis framework, our singing conversion model is based on VITS, incorporating four key modules: a prior…

Audio and Speech Processing · Electrical Eng. & Systems 2023-10-05 Ziqian Ning , Yuepeng Jiang , Zhichao Wang , Bin Zhang , Lei Xie

Speech super-resolution (SSR) enhances low-resolution speech by increasing the sampling rate. While most SSR methods focus on magnitude reconstruction, recent research highlights the importance of phase reconstruction for improved…

Whispered speech as an acceptable form of human-computer interaction is gaining traction. Systems that address multiple modes of speech require a robust front-end speech classifier. Performance of whispered vs normal speech classification…

Audio and Speech Processing · Electrical Eng. & Systems 2024-08-28 S. Johanan Joysingh , P. Vijayalakshmi , T. Nagarajan

Phonetic speech transcription is crucial for fine-grained linguistic analysis and downstream speech applications. While Connectionist Temporal Classification (CTC) is a widely used approach for such tasks due to its efficiency, it often…

The Complete Vocal Technique (CVT) is a school of singing developed in the past decades by Cathrin Sadolin et al.. CVT groups the use of the voice into so called vocal modes, namely Neutral, Curbing, Overdrive and Edge. Knowledge of the…

Sound · Computer Science 2026-04-30 Reemt Hinrichs , Sonja Stephan , Alexander Lange , Jörn Ostermann

This paper describes the data acquisition and trigger system of the Thin Time-of-flight PET (TT-PET) scanner. The system is designed to read out in the order of 1000 pixel sensors used in the scanner and to provide a reference timing signal…

Instrumentation and Detectors · Physics 2018-12-11 Y. Bandi , Y. Favre , D. Ferrere , D. Forshaw , R. Hanni , D. Hayakawa , G. Iacobucci , P. Lutz , A. Miucci , L. Paolozzi , E. Ripiccini , C. Tognina , P. Valerio , M. Weber

DeepFake Audio, unlike DeepFake images and videos, has been relatively less explored from detection perspective, and the solutions which exist for the synthetic speech classification either use complex networks or dont generalize to…

Sound · Computer Science 2022-10-24 Vardhan Dongre , Abhinav Thimma Reddy , Nikhitha Reddeddy

The primary purpose of the collective Thomson scattering (CTS) diagnostic at ITER is to measure the properties of fast-ion populations, in particular those of fusion-born $\alpha$-particles. Based on the present design of the diagnostic, we…

Diagnosing language disorders associated with autism is a complex challenge, often hampered by the subjective nature and variability of traditional assessment methods. Traditional diagnostic methods not only require intensive human effort…

Computation and Language · Computer Science 2024-12-02 Chuanbo Hu , Wenqi Li , Mindi Ruan , Xiangxu Yu , Shalaka Deshpande , Lynn K. Paul , Shuo Wang , Xin Li

With the emergence of GAN-based vocoders, the discriminator, as a crucial component, has been developed recently. In our work, we focus on improving the time-frequency based discriminator. Particularly, Short-Time Fourier Transform (STFT)…

Audio and Speech Processing · Electrical Eng. & Systems 2025-12-04 Nan Xu , Zhaolong Huang , Xiao Zeng

Identification of bird species from audio records is one of the challenging tasks due to the existence of multiple species in the same recording, noise in the background, and long-term recording. Besides, choosing a proper acoustic feature…

Sound · Computer Science 2022-01-04 Nahian Ibn Hasan

When measuring a range of different genomic, epigenomic, transcriptomic and other variables, an integrative approach to analysis can strengthen inference and give new insights. This is also the case when clustering patient samples, and…

Methodology · Statistics 2014-11-03 Kristoffer Hellton , Magne Thoresen