English
Related papers

Related papers: Estimation of binary time-frequency masks from amb…

200 papers

The reverberation time is one of the most important parameters used to characterize the acoustic property of an enclosure. In real-world scenarios, it is much more convenient to estimate the reverberation time blindly from recorded speech…

Sound · Computer Science 2021-12-10 Kaitong Zheng , Chengshi Zheng , Jinqiu Sang , Yulong Zhang , Xiaodong Li

Cochlear implant (CI) users have considerable difficulty in understanding speech in reverberant listening environments. Time-frequency (T-F) masking is a common technique that aims to improve speech intelligibility by multiplying…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-01 Kevin M. Chu , Leslie M. Collins , Boyla O. Mainsah

We propose a general white noise test for functional time series based on estimating a distance between the spectral density operator of a weakly stationary time series and the constant spectral density operator of an uncorrelated time…

Statistics Theory · Mathematics 2020-07-07 Vaidotas Characiejus , Gregory Rice

Phase retrieval with pre-defined optical masks can provide extra constraint and thus achieve improved performance. The recent progress in optimization theory demonstrates the superiority of random masks in phase retrieval algorithms.…

Image and Video Processing · Electrical Eng. & Systems 2022-07-27 Qiuliang Ye , Yuk-Hee Chan , Michael G. Somekh , Daniel P. K. Lun

People often listen to music in noisy environments, seeking to isolate themselves from ambient sounds. Indeed, a music signal can mask some of the noise's frequency components due to the effect of simultaneous masking. In this article, we…

Sound · Computer Science 2025-02-26 Clémentine Berger , Roland Badeau , Slim Essid

Sound reflections and late reverberation alter energetic and binaural cues of a target source, thereby affecting it's detection in noise. Two experiments investigated detection of harmonic complex tones, centered around 500 Hz, in noise in…

Audio and Speech Processing · Electrical Eng. & Systems 2021-07-01 Norbert Kolotzek , Pierre G. Aublin , Bernhard U. Seeber

We aim to develop a technology that makes the sound from earphones and headphones easier to hear without increasing the sound pressure or eliminating ambient noise. To this end, we focus on harnessing the phenomenon of binaural unmasking…

Audio and Speech Processing · Electrical Eng. & Systems 2026-02-23 Rina Kotani , Chiaki Miyazaki , Shiro Suzuki

Differences in interaural phase configuration between a target and a masker can lead to substantial binaural unmasking. This effect is decreased for masking noises with an interaural time difference (ITD). Adding a second noise with an…

Audio and Speech Processing · Electrical Eng. & Systems 2022-06-13 Bernhard Eurich , Jörg Encke , Stephan D. Ewert , Mathias Dietz

We propose and analyze the use of an explicit time-context window for neural network-based spectral masking speech enhancement to leverage signal context dependencies between neighboring frames. In particular, we concentrate on soft masking…

Audio and Speech Processing · Electrical Eng. & Systems 2024-08-29 Luan Vinícius Fiorio , Boris Karanov , Bruno Defraene , Johan David , Wim van Houtum , Frans Widdershoven , Ronald M. Aarts

Cochlear implant users struggle to understand speech in reverberant environments. To restore speech perception, artifacts dominated by reverberant reflections can be removed from the cochlear implant stimulus. Artifacts can be identified…

Sound · Computer Science 2021-08-16 Lidea K. Shahidi , Leslie M. Collins , Boyla O. Mainsah

This paper proposes simple moment based spectrum sensing algorithm for cognitive radio networks in a flat fading channel. It is assumed that the transmitted signal samples are binary (quadrature) phase-shift keying BPSK (QPSK), Mary…

Applications · Statistics 2013-11-26 Tadilo Endeshaw Bogale , Luc Vandendorpe

This letter proposes a novel blind acoustic mask (BAM) designed to adaptively detect noise components and preserve target speech segments in time-domain. A robust standard deviation estimator is applied to the non-stationary noisy speech to…

Audio and Speech Processing · Electrical Eng. & Systems 2021-07-07 F. Farias , R. Coelho

Practitioners deploying time series forecasting models face a dilemma: exhaustively validating dozens of models is computationally prohibitive, yet choosing the wrong model risks poor performance. We show that spectral…

Machine Learning · Computer Science 2025-11-13 Oliver Wang , Pengrui Quan , Kang Yang , Mani Srivastava

The short-time Fourier transform (STFT) provides the foundation of binary-mask based audio source separation approaches. In computing a spectrogram, the STFT window size parameterizes the trade-off between time and frequency resolution.…

Sound · Computer Science 2015-04-29 Andrew J. R. Simpson

We propose a new procedure for white noise testing of a functional time series. Our approach is based on an explicit representation of the $L^2$-distance between the spectral density operator and its best ($L^2$-)approximation by a spectral…

Statistics Theory · Mathematics 2017-09-06 Pramita Bagchi , Vaidotas Characiejus , Holger Dette

Differences between the interaural phase of a noise and a target tone improve detection thresholds. The maximum masking release is obtained for detecting an antiphasic tone (S$\pi$) in diotic noise (N0). It has been shown in several studies…

Audio and Speech Processing · Electrical Eng. & Systems 2022-07-13 Mathias Dietz , Jörg Encke , Kristin I. Bracklo , Stephan D. Ewert

The state-of-art methods for acoustic beamforming in multi-channel ASR are based on a neural mask estimator that predicts the presence of speech and noise. These models are trained using a paired corpus of clean and noisy recordings…

Audio and Speech Processing · Electrical Eng. & Systems 2019-12-02 Rohit Kumar , Anirudh Sreeram , Anurenjan Purushothaman , Sriram Ganapathy

In a typical multi-standard military communication receiver, fast and reliable spectrum sensing unit is required to extract the information of multiple channels (frequency bands) present in a wideband input signal. In this paper, an energy…

Information Theory · Computer Science 2016-08-16 S. J. Darak , A. P. Vinod , E. M-K. Lai

With the advancement of audio generation, generative models can produce highly realistic audios. However, the proliferation of deepfake general audio can pose negative consequences. Therefore, we propose a new task, deepfake general audio…

Sound · Computer Science 2024-06-13 Zeyu Xie , Baihan Li , Xuenan Xu , Zheng Liang , Kai Yu , Mengyue Wu

Due to their robustness and flexibility, neural-driven beamformers are a popular choice for speech separation in challenging environments with a varying amount of simultaneous speakers alongside noise and reverberation. Time-frequency masks…

Audio and Speech Processing · Electrical Eng. & Systems 2025-01-10 Jakob Kienegger , Alina Mannanova , Timo Gerkmann
‹ Prev 1 2 3 10 Next ›