English
Related papers

Related papers: Accommodating false positives within acoustic spat…

200 papers

We collect novel data in the public service domain to evaluate the capability of the state-of-the-art automatic speech recognition (ASR) models in capturing regional differences in accents in the United Kingdom (UK), specifically focusing…

Computation and Language · Computer Science 2025-01-16 Melissa Torgbi , Andrew Clayman , Jordan J. Speight , Harish Tayyar Madabushi

The Fearless Steps APOLLO Community Resource provides unparalleled opportunities to explore the potential of multi-speaker team communications from NASA Apollo missions. This study focuses on discovering the characteristics that make Apollo…

Audio and Speech Processing · Electrical Eng. & Systems 2025-12-19 Alkis Koudounas , Flavio Giobergia

Advancements in AI-synthesized human voices have created a growing threat of impersonation and disinformation, making it crucial to develop methods to detect synthetic human voices. This study proposes a new approach to identifying…

Sound · Computer Science 2023-04-28 Chengzhe Sun , Shan Jia , Shuwei Hou , Siwei Lyu

The rapid advancement of audio generation technologies has escalated the risks of malicious deepfake audio across speech, sound, singing voice, and music, threatening multimedia security and trust. While existing countermeasures (CMs)…

Sound · Computer Science 2026-01-12 Yuankun Xie , Ruibo Fu , Zhiyong Wang , Xiaopeng Wang , Songjun Cao , Long Ma , Haonan Cheng , Long Ye

Passive acoustic sensing is a cost-effective solution for monitoring moving targets such as vessels and aircraft, but its performance is hindered by complex propagation effects like multi-path reflections and motion-induced artefacts.…

Sound · Computer Science 2026-01-23 Lucas C. F. Domingos , Russell S. A. Brinkworth , Paulo E. Santos , Karl Sammut

We consider the scenario where important signals are not strong enough to be separable from a large amount of noise. Such weak signals commonly exist in large-scale data analysis and play vital roles in many biomedical applications.…

Methodology · Statistics 2022-01-26 X. Jessie Jeng , Yifei Hu

Single-word Automatic Speech Recognition (ASR) is a challenging task due to the lack of linguistic context and sensitivity to noise, pronunciation variation, and channel artifacts, especially in low-resource, communication-critical domains…

Sound · Computer Science 2026-01-30 Manali Sharma , Riya Naik , Buvaneshwari G

Joint optimization of multi-channel front-end and automatic speech recognition (ASR) has attracted much interest. While promising results have been reported for various tasks, past studies on its meeting transcription application were…

Audio and Speech Processing · Electrical Eng. & Systems 2020-11-30 Xiaofei Wang , Naoyuki Kanda , Yashesh Gaur , Zhuo Chen , Zhong Meng , Takuya Yoshioka

Bird sounds possess distinctive spectral structure which may exhibit small shifts in spectrum depending on the bird species and environmental conditions. In this paper, we propose using convolutional recurrent neural networks on the task of…

Automatic speaker verification (ASV) technology is recently finding its way to end-user applications for secure access to personal data, smart services or physical facilities. Similar to other biometric technologies, speaker verification is…

Sound · Computer Science 2016-09-16 Cemal Hanilci , Tomi Kinnunen , Md Sahidullah , Aleksandr Sizov

One-bit compressive sensing (CS) is an advanced version of sparse recovery in which the sparse signal of interest can be recovered from extremely quantized measurements. Namely, only the sign of each measurement is available to us. In many…

Information Theory · Computer Science 2019-10-22 Hossein Beheshti , Sajad Daei , Farzan Haddadi

We present results from fitting the baryon acoustic oscillation (BAO) signal in the correlation function obtained from the first application of reconstruction to a galaxy redshift survey, namely, the Sloan Digital Sky Survey (SDSS) Data…

Cosmology and Nongalactic Astrophysics · Physics 2015-06-04 Xiaoying Xu , Nikhil Padmanabhan , Daniel J. Eisenstein , Kushal T. Mehta , Antonio J. Cuesta

A common question being raised in automatic speech recognition (ASR) evaluations is how reliable is an observed word error rate (WER) improvement comparing two ASR systems, where statistical hypothesis testing and confidence interval (CI)…

Machine Learning · Statistics 2020-05-22 Zhe Liu , Fuchun Peng

This project proposes the development of a comprehensive real-time biodiversity monitoring system that harnesses sound data through a network of acoustic sensors and advanced artificial intelligence algorithms. The system analyzes sound…

Audio and Speech Processing · Electrical Eng. & Systems 2024-10-18 Kumar Srinivas Bobba , Kartheeban K , Vamsi Krishna Sai , Dinesh Bugga , Vijaya Mani Surendra Bolla

In clinical voice signal analysis, mishandling of subharmonic voicing may cause an acoustic parameter to signal false negatives. As such, the ability of a fundamental frequency estimator to identify speaking fundamental frequency is…

Audio and Speech Processing · Electrical Eng. & Systems 2025-01-10 Takeshi Ikuma , Melda Kunduk , Andrew J. McWhorter

Stochastic resonance (SR), a phenomenon originally introduced in climate modeling, enhances signal detection by leveraging optimal noise levels within non-linear systems. Traditional SR techniques, mainly based on single-threshold…

Signal Processing · Electrical Eng. & Systems 2025-10-27 Dixon Vimalajeewa , Ursula U. Muller , Brani Vidakovic

The baryon acoustic oscillations are a promising route to the precision measure of the cosmological distance scale and hence the measurement of the time evolution of dark energy. We show that the non-linear degradation of the acoustic…

Astrophysics · Physics 2008-11-26 Daniel J. Eisenstein , Hee-jong Seo , Edwin Sirko , David Spergel

Baryonic Acoustic Oscillations (BAO) and their effects on the matter power spectrum can be studied by using the Lyman-alpha absorption signature of the matter density field along quasar (QSO) lines of sight. A measurement sufficiently…

Cosmology and Nongalactic Astrophysics · Physics 2009-10-21 Ch. Yeche , P. Petitjean , J. Rich , E. Aubourg , N. Busca , J. -Ch. Hamilton , J. -M. Le Goff , I. Paris , S. Peirani , Ch. Pichon , E. Rollinde , M. Vargas-Magana

In speech quality estimation for speech enhancement (SE) systems, subjective listening tests so far are considered as the gold standard. This should be even more true considering the large influx of new generative or hybrid methods into the…

Audio and Speech Processing · Electrical Eng. & Systems 2025-07-28 Marvin Sach , Yihui Fu , Kohei Saijo , Wangyou Zhang , Samuele Cornell , Robin Scheibler , Chenda Li , Anurag Kumar , Wei Wang , Yanmin Qian , Shinji Watanabe , Tim Fingscheidt

Non-interactive and linear experiences like cinema film offer high quality surround sound audio to enhance immersion, however the listener's experience is usually fixed to a single acoustic perspective. With the rise of virtual reality,…

Audio and Speech Processing · Electrical Eng. & Systems 2024-10-30 Lachlan Birnie , Thushara Abhayapala , Vladimir Tourbabin , Prasanga Samarasinghe