English
Related papers

Related papers: PSD Estimation and Source Separation in a Noisy Re…

200 papers

Complex architectures for wireless communications, digital electronics and space-based navigation interlink several oscillator-based devices such as clocks, transponders and synthesizers. Estimators characterizing their stability are…

Data Analysis, Statistics and Probability · Physics 2023-11-02 Fabrizio De Marchi , Michael K. Plumaris , Eric A. Burt , Luciano Iess

Due to the high variation in the application requirements of sound event detection (SED) systems, it is not sufficient to evaluate systems only in a single operating mode. Therefore, the community recently adopted the polyphonic sound…

Audio and Speech Processing · Electrical Eng. & Systems 2023-06-28 Janek Ebbers , Reinhold Haeb-Umbach , Romain Serizel

There is a growing interest in new sensing technologies and processing algorithms to increase the level of driving automation towards self-driving vehicles. The challenge for autonomy is especially difficult for the negotiation of uncharted…

Systems and Control · Electrical Eng. & Systems 2019-10-15 Giulio Reina , Antonio Leanza , Annalisa Milella , Arcangelo Messina

We consider the problem of separating speech sources captured by multiple spatially separated devices, each of which has multiple microphones and samples its signals at a slightly different rate. Most asynchronous array processing methods…

Audio and Speech Processing · Electrical Eng. & Systems 2019-12-12 Ryan M. Corey , Andrew C. Singer

Multi-channel speech enhancement aims to recover clean speech from noisy multi-channel recordings. Most deep learning methods employ discriminative training, which can lead to non-linear distortions from regression-based objectives,…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-26 Zhongweiyang Xu , Ashutosh Pandey , Juan Azcarreta , Zhaoheng Ni , Sanjeel Parekh , Buye Xu

A two-step enhancement method based on spectral subtraction and phase spectrum compensation is presented in this paper for noisy speeches in adverse environments involving non-stationary noise and medium to low levels of SNR. The magnitude…

Audio and Speech Processing · Electrical Eng. & Systems 2018-03-02 Md Tauhidul Islam , Asaduzzaman , Celia Shahnaz , Wei-Ping Zhu , M. Omair Ahmad

The estimation of the frequencies of multiple superimposed exponentials in noise is an important research problem due to its various applications from engineering to chemistry. In this paper, we propose an efficient and accurate algorithm…

Numerical Analysis · Mathematics 2016-05-05 Shanglin Ye , Elias Aboutanios

Speech separation with several speakers is a challenging task because of the non-stationarity of the speech and the strong signal similarity between interferent sources. Current state-of-the-art solutions can separate well the different…

Signal Processing · Electrical Eng. & Systems 2021-02-09 Nicolas Furnon , Romain Serizel , Irina Illina , Slim Essid

Super-resolution theory aims to estimate the discrete components lying in a continuous space that constitute a sparse signal with optimal precision. This work investigates the potential of recent super-resolution techniques for spectral…

Information Theory · Computer Science 2016-11-24 M. Ferreira Da Costa , W. Dai

It is often required to extract the sound of an objective instrument played in concert with other instruments. Microphone array is one of the effective ways to enhance a sound from a specific direction. However it is not effective in an…

Sound · Computer Science 2016-11-11 Naoya Takahashi , Mitsuharu Matsumoto , Shuji Hashimoto

Audio processing methods based on deep neural networks are typically trained at a single sampling frequency (SF). To handle untrained SFs, signal resampling is commonly employed, but it can degrade performance, particularly when the input…

Sound · Computer Science 2026-01-22 Kanami Imamura , Tomohiko Nakamura , Kohei Yatabe , Hiroshi Saruwatari

We propose a data-driven sparse recovery framework for hybrid spherical linear microphone arrays using singular value decomposition (SVD) of the transfer operator. The SVD yields orthogonal microphone and field modes, reducing to spherical…

Audio and Speech Processing · Electrical Eng. & Systems 2026-02-04 Shunxi Xu , Thushara Abhayapala , Craig T. Jin

The direction of arrival (DOA) estimation of sound sources has been a popular signal processing research topic due to its widespread applications. Using spherical microphone arrays (SMA), DOA estimation can be applied in the spherical…

Audio and Speech Processing · Electrical Eng. & Systems 2017-11-07 Hossein Lolaee , Mohammad Ali Akhaee

Random pulse width modulation techniques are used in AC motors powered by two-level three-phase inverters, which cause a broadband spectrum of voltage, current, and electromagnetic force. The voltage distribution across a wide range of…

Systems and Control · Electrical Eng. & Systems 2024-06-07 Jian Wen , Xiaobin Cheng , Peifeng Ji , Jun Yang , Feng Zhao

We propose a benchmark of state-of-the-art sound event detection systems (SED). We designed synthetic evaluation sets to focus on specific sound event detection challenges. We analyze the performance of the submissions to DCASE 2021 task 4…

Spherical microphone arrays are convenient tools for capturing the spatial characteristics of a sound field. However, achieving superior spatial resolution requires arrays with numerous capsules, consequently leading to expensive devices.…

Audio and Speech Processing · Electrical Eng. & Systems 2024-07-29 Federico Miotello , Ferdinando Terminiello , Mirco Pezzoli , Alberto Bernardini , Fabio Antonacci , Augusto Sarti

Real-time single-channel speech separation aims to unmix an audio stream captured from a single microphone that contains multiple people talking at once, environmental noise, and reverberation into multiple de-reverberated and noise-free…

Audio and Speech Processing · Electrical Eng. & Systems 2023-04-18 Julian Neri , Sebastian Braun

Blind Speech Separation (BSS) aims to separate multiple speech sources from audio mixtures recorded by a microphone array. The problem is challenging because it is a blind inverse problem, i.e., the microphone array geometry, the room…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-18 Zhongweiyang Xu , Xulin Fan , Zhong-Qiu Wang , Xilin Jiang , Romit Roy Choudhury

Multi-channel acoustic signal processing is a well-established and powerful tool to exploit the spatial diversity between a target signal and non-target or noise sources for signal enhancement. However, the textbook solutions for optimal…

Audio and Speech Processing · Electrical Eng. & Systems 2025-01-14 Reinhold Haeb-Umbach , Tomohiro Nakatani , Marc Delcroix , Christoph Boeddeker , Tsubasa Ochiai

This article addresses the modeling of reverberant recording environments in the context of under-determined convolutive blind source separation. We model the contribution of each source to all mixture channels in the time-frequency domain…

Machine Learning · Statistics 2009-12-14 Ngoc Duong , Emmanuel Vincent , Remi Gribonval