English
Related papers

Related papers: RTF-steered binaural MVDR beamforming incorporatin…

200 papers

Recently, many deep learning based beamformers have been proposed for multi-channel speech separation. Nevertheless, most of them rely on extra cues known in advance, such as speaker feature, face image or directional information. In this…

Audio and Speech Processing · Electrical Eng. & Systems 2022-12-08 Yanjie Fu , Haoran Yin , Meng Ge , Longbiao Wang , Gaoyan Zhang , Jianwu Dang , Chengyun Deng , Fei Wang

Line differential microphone arrays have attracted attention for their ability to achieve frequency-invariant beampatterns and high directivity. Recently, the Jacobi-Anger expansion-based approach has enabled the design of fully…

Signal Processing · Electrical Eng. & Systems 2025-08-26 Yankai Zhang , Jiafeng Ding , Jingjing Ning , Qiaoxi Zhu

The aim of this study is to implement a method to remove ambient noise in biomedical sounds captured in auscultation. We propose an incremental approach based on multichannel non-negative matrix partial co-factorization (NMPCF) for ambient…

This paper presents Rec-RIR for monaural blind room impulse response (RIR) identification. Rec-RIR is developed based on the convolutive transfer function (CTF) approximation, which models reverberation effect within narrow-band filter…

Audio and Speech Processing · Electrical Eng. & Systems 2026-01-22 Pengyu Wang , Xiaofei Li

Blind speech separation (BSS) aims to recover multiple speech sources from multi-channel, multi-speaker mixtures under unknown array geometry and room impulse responses. In unsupervised setup where clean target speech is not available for…

Sound · Computer Science 2025-10-13 Shulin He , Zhong-Qiu Wang

Full Duplex (FD) radio has emerged as a promising solution to increase the data rates by up to a factor of two via simultaneous transmission and reception in the same frequency band. This paper studies a novel hybrid beamforming (HYBF)…

Information Theory · Computer Science 2022-01-04 Chandan Kumar Sheemar , Christo Kurisummoottil Thomas , Dirk Slock

In this paper, we present the Blind Speech Separation and Dereverberation (BSSD) network, which performs simultaneous speaker separation, dereverberation and speaker identification in a single neural network. Speaker separation is guided by…

Sound · Computer Science 2021-11-08 Lukas Pfeifenberger , Franz Pernkopf

Audio-visual speech separation methods aim to integrate different modalities to generate high-quality separated speech, thereby enhancing the performance of downstream tasks such as speech recognition. Most existing state-of-the-art (SOTA)…

Sound · Computer Science 2024-03-22 Samuel Pegg , Kai Li , Xiaolin Hu

This paper focuses on multiuser MIMO channel estimation and data transmission at millimeter wave (mmWave) frequencies. The proposed approach relies on the time-division-duplex (TDD) protocol and is based on two distinct phases. First of…

Information Theory · Computer Science 2017-04-25 Stefano Buzzi , Carmen D'Andrea

In this letter, we propose enhanced factored three way restricted Boltzmann machines (EFTW-RBMs) for speech detection. The proposed model incorporates conditional feature learning by multiplying the dynamical state of the third unit, which…

Sound · Computer Science 2017-04-24 Pengfei Sun , Jun Qin

Usually, hearing impaired people use hearing aids which are implemented with speech enhancement algorithms. Estimation of speech and estimation of nose are the components in single channel speech enhancement system. The main objective of…

Sound · Computer Science 2014-11-10 M. Ravichandra Kumar , B. Ravi Teja

We investigate a speech enhancement method based on the binaural coherence-to-diffuse power ratio (CDR), which preserves auditory spatial cues for maskers and a broadside target. Conventional CDR estimators typically rely on a mathematical…

Audio and Speech Processing · Electrical Eng. & Systems 2022-07-19 Reza Ghanavi , Craig Jin

Support vector regression (SVR) is one of the most popular machine learning algorithms aiming to generate the optimal regression curve through maximizing the minimal margin of selected training samples, i.e., support vectors. Recent…

Machine Learning · Computer Science 2019-05-07 Gaoyang Li , Jinyu Yang , Chunguo Wu , Qin Ma

Delay-and-Sum (DAS) is the most common algorithm used in photoacoustic (PA) image formation. However, this algorithm results in a reconstructed image with a wide mainlobe and high level of sidelobes. Minimum variance (MV), as an adaptive…

Signal Processing · Electrical Eng. & Systems 2018-05-11 Roya Paridar , Moein Mozaffarzadeh , Mohammad Mehrmohammadi , Mahdi Orooji

The contrast transfer function (CTF) is widely used to evaluate phase retrieval methods in scanning transmission electron microscopy (STEM), including center-of-mass imaging, parallax imaging, direct ptychography, and iterative…

This paper considers the problem of jointly designing the transmit waveforms and weights for a frequency diverse array (FDA) in a spectrally congested environment in which unintentional spectral interferences exist. Exploiting the…

Signal Processing · Electrical Eng. & Systems 2022-12-21 Wenkai Jia , Andreas Jakobsson , Wen-Qin Wang

In millimeter-wave (mmWave) dual-function radar-communication (DFRC) systems, hybrid beamforming (HBF) is recognized as a promising technique utilizing a limited number of radio frequency chains. In this work, in the presence of extended…

Signal Processing · Electrical Eng. & Systems 2022-11-07 Ziyang Cheng , Linlong Wu , Bowen Wang , Bhavani Shankar M. R. , Björn Ottersten

Complexity reduction of optimal linear receiver is considered in a scenario where both the number of single-antenna user equipments (UEs) $K$ and base station (BS) antennas $N$ are large. Two-stage beamforming (TSB) greatly alleviates the…

Signal Processing · Electrical Eng. & Systems 2019-12-03 Hossein Asgharimoghaddam , Antti Tölli

We propose a method of separating a desired sound source from a single-channel mixture, based on either a textual description or a short audio sample of the target source. This is achieved by combining two distinct models. The first model,…

Audio and Speech Processing · Electrical Eng. & Systems 2022-04-13 Kevin Kilgour , Beat Gfeller , Qingqing Huang , Aren Jansen , Scott Wisdom , Marco Tagliasacchi

Hybrid precoders and combiners are designed for cooperative cell-free multi-user millimeter wave (mmWave) multiple-input multiple-output (MIMO) cellular networks for low complexity interference mitigation. Initially, we derive an optimal…

Signal Processing · Electrical Eng. & Systems 2022-12-15 Meesam Jafri , Suraj Srivastava , Naveen K. D. Venkategowda , Aditya K. Jagannatham , Lajos Hanzo
‹ Prev 1 8 9 10 Next ›