English
Related papers

Related papers: Direction-Preserving MIMO Speech Enhancement Using…

200 papers

Spatial wideband effects are known to affect channel estimation and localization performance in millimeter wave (mmWave) massive multiple-input multiple-output (MIMO) systems. Based on perturbation analysis, we show that the spatial…

Signal Processing · Electrical Eng. & Systems 2022-09-20 Shudi Weng , Fan Jiang , Henk Wymeersch

Audio and omni-modal large language models exhibit impressive cross-modal reasoning capabilities. However, applying standard reinforcement learning post-training algorithms to these models exposes a critical structural vulnerability:…

Computation and Language · Computer Science 2026-05-28 Cihan Xiao , Yiwen Shao , Chenxing Li , Xiang He , Zhenwen Liang , Steve Yves , Sanjeev Khudanpur , Liefeng Bo

In multi-user millimeter wave (mmWave) multiple-input-multiple-output (MIMO) systems, hybrid precoding is a crucial task to lower the complexity and cost while achieving a sufficient sum-rate. Previous works on hybrid precoding were usually…

Signal Processing · Electrical Eng. & Systems 2020-04-28 Ahmet M. Elbir , Anastasios Papazafeiropoulos

To glean the benefits offered by massive multi-input multi-output (MIMO) systems, channel state information must be accurately acquired. Despite the high accuracy, the computational complexity of classical linear minimum mean squared error…

Information Theory · Computer Science 2024-04-23 Bin Li , Ziping Wei , Shaoshi Yang , Yang Zhang , Jun Zhang , Chenglin Zhao , Sheng Chen

Multimodal representation learning has shown promising improvements on various vision-language tasks. Most existing methods excel at building global-level alignment between vision and language while lacking effective fine-grained image-text…

Computer Vision and Pattern Recognition · Computer Science 2023-06-16 Zijia Zhao , Longteng Guo , Xingjian He , Shuai Shao , Zehuan Yuan , Jing Liu

In this paper we present a single-microphone speech enhancement algorithm. A hybrid approach is proposed merging the generative mixture of Gaussians (MoG) model and the discriminative neural network (NN). The proposed algorithm is executed…

Sound · Computer Science 2015-10-27 Shlomo E. Chazan , Jacob Goldberger , Sharon Gannot

This paper considers pilot-based channel estimation in large-scale multiple-input multiple-output (MIMO) communication systems, also known as "massive MIMO". Unlike previous works on this topic, which mainly considered the impact of…

Information Theory · Computer Science 2013-07-18 Nafiseh Shariati , Emil Björnson , Mats Bengtsson , Mérouane Debbah

Audio-visual speech enhancement system is regarded as one of promising solutions for isolating and enhancing speech of desired speaker. Typical methods focus on predicting clean speech spectrum via a naive convolution neural network based…

Audio and Speech Processing · Electrical Eng. & Systems 2022-07-01 Xinmeng Xu , Yang Wang , Jie Jia , Binbin Chen , Dejun Li

Goal: Numerous studies had successfully differentiated normal and abnormal voice samples. Nevertheless, further classification had rarely been attempted. This study proposes a novel approach, using continuous Mandarin speech instead of a…

Audio and Speech Processing · Electrical Eng. & Systems 2022-02-23 Syu-Siang Wang , Chi-Te Wang , Chih-Chung Lai , Yu Tsao , Shih-Hau Fang

The majority of multichannel speech enhancement algorithms are two-step procedures that first apply a linear spatial filter, a so-called beamformer, and combine it with a single-channel approach for postprocessing. However, the serial…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-16 Kristina Tesch , Timo Gerkmann

Interference during the uplink training phase significantly deteriorates the performance of a massive MIMO system. The impact of the interference can be reduced by exploiting second order statistics of the channel vectors, e.g., to obtain…

Information Theory · Computer Science 2018-05-23 David Neumann , Michael Joham , Wolfgang Utschick

In this paper, we present a novel multi-channel speech extraction system to simultaneously extract multiple clean individual sources from a mixture in noisy and reverberant environments. The proposed method is built on an improved…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-17 Jisi Zhang , Catalin Zorila , Rama Doddipatla , Jon Barker

Orthogonal delay-Doppler division multiplexing~(ODDM) modulation has recently been regarded as a promising technology to provide reliable communications in high-mobility situations. Accurate and low-complexity channel estimation is one of…

Signal Processing · Electrical Eng. & Systems 2025-07-29 Dezhi Wang , Chongwen Huang , Xiaojun Yuan , Sami Muhaidat , Lei Liu , Xiaoming Chen , Zhaoyang Zhang , Chau Yuen , Mérouane Debbah

Holographic multiple-input multiple-output (MIMO) systems represent a spatially constrained MIMO architecture with a massive number of antennas with small antenna spacing as a close approximation of a spatially continuous electromagnetic…

Signal Processing · Electrical Eng. & Systems 2024-12-19 Nikolaos Kolomvakis , Emil Björnson

Audio-visual speech enhancement aims to extract clean speech from a noisy environment by leveraging not only the audio itself but also the target speaker's lip movements. This approach has been shown to yield improvements over audio-only…

The aim of speech enhancement is to improve speech signal quality and intelligibility from a noisy microphone signal. In many applications, it is crucial to enable processing with small computational complexity and minimal requirements…

Audio and Speech Processing · Electrical Eng. & Systems 2023-09-08 Julitta Bartolewska , Stanisław Kacprzak , Konrad Kowalczyk

Channel estimation is a critical task in digital communications that greatly impacts end-to-end system performance. In this work, we introduce a novel approach for multiple-input multiple-output (MIMO) channel estimation using score-based…

Signal Processing · Electrical Eng. & Systems 2022-02-16 Marius Arvinte , Jonathan I Tamir

In this paper, we present a method that allows to further improve speech enhancement obtained with recently introduced Deep Neural Network (DNN) models. We propose a multi-channel refinement method of time-frequency masks obtained with…

Audio and Speech Processing · Electrical Eng. & Systems 2023-09-19 Julitta Bartolewska , Stanisław Kacprzak , Konrad Kowalczyk

We introduce a novel online convex optimization (OCO) framework to estimate the user's signal-to-interference-plus-noise ratio (SINR) from ACK/NACK feedback, channel quality indicator (CQI) reports, and previously selected modulation and…

Information Theory · Computer Science 2026-03-16 Lorenzo Maggi , Boris Bonev , Reinhard Wiesmayr , Sebastian Cammerer , Alexander Keller

Robust spatial audio control relies on accurate acoustic propagation models, yet environmental variations, especially changes in the speed of sound, cause systematic mismatches that degrade performance. Existing methods either assume known…

Audio and Speech Processing · Electrical Eng. & Systems 2026-05-13 Andreas Jonas Fuglsig , Mads Græsbøll Christensen , Jesper Rindom Jensen
‹ Prev 1 8 9 10 Next ›