English
Related papers

Related papers: HRTF-based Robust Least-Squares Frequency-Invarian…

200 papers

Speaker localization for binaural microphone arrays has been widely studied for applications such as speech communication, video conferencing, and robot audition. Many methods developed for this task, including the direct path dominance…

Audio and Speech Processing · Electrical Eng. & Systems 2023-11-01 Yanir Maymon , Israel Nelken , Boaz Rafaely

As spatial audio is enjoying a surge in popularity, data-driven machine learning techniques that have been proven successful in other domains are increasingly used to process head-related transfer function measurements. However, these…

Audio and Speech Processing · Electrical Eng. & Systems 2022-12-09 Johan Pauwels , Lorenzo Picinali

Head-related transfer functions (HRTFs) are essential for virtual acoustic realities, as they contain all cues for localizing sound sources in three-dimensional space. Acoustic measurements are one way to obtain high-quality HRTFs. To…

Audio and Speech Processing · Electrical Eng. & Systems 2023-10-25 Johannes M. Arend , Christoph Pörschmann , Stefan Weinzierl , Fabian Brinkmann

Hybrid beamforming (HBF) design is a crucial stage in millimeter wave (mmWave) multi-user multi-input multi-output (MU-MIMO) systems. However, conventional HBF methods are still with high complexity and strongly rely on the quality of…

Signal Processing · Electrical Eng. & Systems 2020-04-28 Shaocheng Huang , Yu Ye , Ming Xiao

Millimeter Wave (mmWave) communications with full-duplex (FD) have the potential of increasing the spectral efficiency, relative to those with half-duplex. However, the residual self-interference (SI) from FD and high pathloss inherent to…

Signal Processing · Electrical Eng. & Systems 2020-04-20 Shaocheng Huang , Yu Ye , Ming Xiao

We propose BeamTransformer, an efficient architecture to leverage beamformer's edge in spatial filtering and transformer's capability in context sequence modeling. BeamTransformer seeks to optimize modeling of sequential relationship among…

Sound · Computer Science 2021-09-10 Siqi Zheng , Shiliang Zhang , Weilong Huang , Qian Chen , Hongbin Suo , Ming Lei , Jinwei Feng , Zhijie Yan

In many multi-microphone algorithms, an estimate of the relative transfer functions (RTFs) of the desired speaker is required. Recently, a computationally efficient RTF vector estimation method was proposed for acoustic sensor networks,…

Audio and Speech Processing · Electrical Eng. & Systems 2023-04-07 Wiebke Middelberg , Simon Doclo

Head-related transfer functions (HRTFs) are important for immersive audio, and their spatial interpolation has been studied to upsample finite measurements. Recently, neural fields (NFs) which map from sound source direction to HRTF have…

Audio and Speech Processing · Electrical Eng. & Systems 2024-02-29 Yoshiki Masuyama , Gordon Wichern , François G. Germain , Zexu Pan , Sameer Khurana , Chiori Hori , Jonathan Le Roux

This paper investigates the issues of the hybrid beamforming design for the orthogonal frequency division multiplexing dual-function radar-communication (DFRC) system in multiple task scenarios involving the radar scanning and detection…

Signal Processing · Electrical Eng. & Systems 2025-04-29 Lingyun Xu , Bowen Wang , Ziyang Cheng

Conventional hybrid beamforming architectures are often compared with one another and with the fully-digital architecture under the same \emph{radiated} antenna power. However, the physically relevant budget is the power injected by the…

Signal Processing · Electrical Eng. & Systems 2026-03-19 Nikola Zlatanov , Damir Salakhov

This paper proposes a Transformer-based hybrid beamforming framework for reconfigurable pixel antenna (RPA)-equipped massive multiple-input multiple-output (MIMO) in high-altitude platform station (HAPS) communications. The proposed pattern…

Information Theory · Computer Science 2026-05-19 Ruiqi Wang , Ziwei Wan , Keke Ying , Zhen Gao

From a machine learning perspective, the human ability localize sounds can be modeled as a non-parametric and non-linear regression problem between binaural spectral features of sound received at the ears (input) and their sound-source…

Sound · Computer Science 2015-02-12 Yuancheng Luo , Dmitry N. Zotkin , Ramani Duraiswami

This paper addresses the problem of under-determinded speech source separation from multichannel microphone singals, i.e. the convolutive mixtures of multiple sources. The time-domain signals are first transformed to the short-time Fourier…

Sound · Computer Science 2019-04-11 Xiaofei Li , Laurent Girin , Radu Horaud

Far-field speech recognition in noisy and reverberant conditions remains a challenging problem despite recent deep learning breakthroughs. This problem is commonly addressed by acquiring a speech signal from multiple microphones and…

Audio and Speech Processing · Electrical Eng. & Systems 2018-10-17 Zhong Meng , Shinji Watanabe , John R. Hershey , Hakan Erdogan

This paper investigates dynamic hybrid beamforming (HBF) for a dual-function radar-communication (DFRC) system, where the DFRC base station (BS) simultaneously serves multiple single-antenna users and senses a target in the presence of…

Signal Processing · Electrical Eng. & Systems 2023-09-15 Bowen Wang , Hongyu Li , Ziyang Cheng

Reconfigurable antennas possess the capability to dynamically adjust their fundamental operating characteristics, thereby enhancing system adaptability and performance. To fully exploit this flexibility in modern wireless communication…

Signal Processing · Electrical Eng. & Systems 2025-12-19 Mengzhen Liu , Ming Li , Rang Liu , Qian Liu

This paper introduces an explainable DNN-based beamformer with a postfilter (ExNet-BF+PF) for multichannel signal processing. Our approach combines the U-Net network with a beamformer structure to address this problem. The method involves a…

Audio and Speech Processing · Electrical Eng. & Systems 2024-11-19 Adi Cohen , Daniel Wong , Jung-Suk Lee , Sharon Gannot

Speech enhancement is a fundamental challenge in signal processing, particularly when robustness is required across diverse acoustic conditions and microphone setups. Deep learning methods have been successful for speech enhancement, but…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-28 Yuval Bar Ilan , Boaz Rafaely , Vladimir Tourbabin

Far-field speech processing is an important and challenging problem. In this paper, we propose \textit{deep ad-hoc beamforming}, a deep-learning-based multichannel speech enhancement framework based on ad-hoc microphone arrays, to address…

Sound · Computer Science 2021-02-10 Xiao-Lei Zhang

The objective of Audio Augmented Reality (AAR) applications are to seamlessly integrate virtual sound sources within a real environment. It is critical for these applications that virtual sources are localised precisely at the intended…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-13 Vincent Martin , Lorenzo Picinali