English
Related papers

Related papers: RTF-steered binaural MVDR beamforming incorporatin…

200 papers

This paper proposes a digital amplitude-phase weighting array based a minimum variance multi-frequency distortionless restriction (MVMFDR) to aviod the frequency band signal distortion in digital beamformer and too short time delay line…

Information Theory · Computer Science 2011-06-21 Yipeng Liu , Jia Xu , Qun Wan , Yingning Peng

This paper analyzes the statistical properties of the signal-to-noise ratio (SNR) at the output of the Capon's minimum variance distortionless response (MVDR) beamformers when operating over impulsive noises. Particularly, we consider the…

Information Theory · Computer Science 2016-05-17 Khalil Elkhalil , Abla Kammoun , Tareq Y. Al-Naffouri , Mohamed-Slim Alouini

Applying a sparse constraint on the beam pattern has been suggested to suppress the sidelobe of the minimum variance distortionless response (MVDR) beamformer recently. To further improve the performance, we add a mixed norm constraint on…

Information Theory · Computer Science 2015-03-17 Yipeng Liu , Qun Wan

We propose a deep beamforming framework for enhancing target speaker(s) in multi-speaker environments. A deep neural network (DNN) is trained to estimate beamforming weights directly from noisy multichannel inputs while satisfying linear…

Audio and Speech Processing · Electrical Eng. & Systems 2026-05-21 Ilai Zaidel , Ori Engel , Bar Engel , Sharon Gannot

Time-domain audio separation network (TasNet) has achieved remarkable performance in blind source separation (BSS). Classic multi-channel speech processing framework employs signal estimation and beamforming. For example, Beam-TasNet links…

Audio and Speech Processing · Electrical Eng. & Systems 2022-04-13 Hangting Chen , Yang Yi , Dang Feng , Pengyuan Zhang

For multichannel speech enhancement, this letter derives a robust maximum likelihood distortionless response beamformer by modeling speech sparse priors with a complex generalized Gaussian distribution, where we refer to as the CGGD-MLDR…

Audio and Speech Processing · Electrical Eng. & Systems 2021-02-22 Weixin Meng , Chengshi Zheng , Xiaodong Li

This paper proposes a novel bidirectional neural vocoder, named BiVocoder, capable both of feature extraction and reverse waveform generation within the short-time Fourier transform (STFT) domain. For feature extraction, the BiVocoder takes…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-05 Hui-Peng Du , Ye-Xin Lu , Yang Ai , Zhen-Hua Ling

Time-domain training criteria have proven to be very effective for the separation of single-channel non-reverberant speech mixtures. Likewise, mask-based beamforming has shown impressive performance in multi-channel reverberant speech…

Invariance to microphone array configuration is a rare attribute in neural beamformers. Filter-and-sum (FS) methods in this class define the target signal with respect to a reference channel. However, this not only complicates formulation…

Audio and Speech Processing · Electrical Eng. & Systems 2023-02-28 Anton Kovalyov , Kashyap Patel , Issa Panahi

Extracting a target source from underdetermined mixtures is challenging for beamforming approaches. Recently proposed time-frequency-bin-wise switching (TFS) and linear combination (TFLC) strategies mitigate this by combining multiple…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-17 Changda Chen , Yichen Yang , Wei Liu , Shoji Makino

Many spatial filtering algorithms used for voice capture in, e.g., teleconferencing applications, can benefit from or even rely on knowledge of Relative Transfer Functions (RTFs). Accordingly, many RTF estimators have been proposed which,…

Audio and Speech Processing · Electrical Eng. & Systems 2021-10-06 Andreas Brendel , Johannes Zeitler , Walter Kellermann

This paper addresses the problem of estimating the shape of objects that exhibit spatially-varying reflectance. We assume that multiple images of the object are obtained under a fixed view-point and varying illumination, i.e., the setting…

Computer Vision and Pattern Recognition · Computer Science 2016-09-22 Zhuo Hui , Aswin C Sankaranarayanan

The increasing popularity of spatial audio in applications such as teleconferencing, entertainment, and virtual reality has led to the recent developments of binaural reproduction methods. However, only a few of these methods are…

Audio and Speech Processing · Electrical Eng. & Systems 2025-02-17 Ami Berger , Vladimir Tourbabin , Jacob Donley , Zamir Ben-Hur , Boaz Rafaely

Distant speech processing is a challenging task, especially when dealing with the cocktail party effect. Sound source separation is thus often required as a preprocessing step prior to speech recognition to improve the signal to distortion…

Audio and Speech Processing · Electrical Eng. & Systems 2020-08-06 Francois Grondin , Jean-Samuel Lauzon , Jonathan Vincent , Francois Michaud

The multichannel Wiener filter (MWF) and its variations have been extensively applied to binaural hearing aids. However, its major drawback is the distortion of the binaural cues of the residual noise, changing the original acoustic…

Audio and Speech Processing · Electrical Eng. & Systems 2019-11-12 Johnny Werner , Marcio H. Costa

This paper presents a Head-Related Transfer Function (HRTF)-guided framework for binaural Target Speaker Extraction (TSE) from mixtures of concurrent sources. Unlike conventional TSE methods based on Direction of Arrival (DOA) estimation or…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-18 Yoav Ellinson , Sharon Gannot

Relative impulse responses between microphones are usually long and dense due to the reverberant acoustic environment. Estimating them from short and noisy recordings poses a long-standing challenge of audio signal processing. In this paper…

Sound · Computer Science 2016-11-17 Zbynek Koldovsky , Jiri Malek , Sharon Gannot

A full performance analysis of the widely linear (WL) minimum variance distortionless response (MVDR) beamformer is introduced. While the WL MVDR is known to outperform its strictly linear counterpart, the Capon beamformer, for noncircular…

Information Theory · Computer Science 2021-12-01 Zhe Li , Rui Pu , Yili Xia , Wenjiang Pei , Danilo P. Mandic

Dereverberation of a moving speech source in the presence of other directional interferers, is a harder problem than that of stationary source and interference cancellation. We explore joint multi channel linear prediction (MCLP) and…

Audio and Speech Processing · Electrical Eng. & Systems 2019-10-23 Srikanth Raj Chetupalli , Thippur V. Sreenivas

The robust adaptive beamforming design problem based on estimation of the signal of interest steering vector is considered in the paper. In this case, the optimal beamformer is obtained by computing the sample matrix inverse and an optimal…

Signal Processing · Electrical Eng. & Systems 2019-06-26 Yongwei Huang , Mingkang Zhou , Sergiy A. Vorobyov