中文
相关论文

相关论文: Direction-Preserving MIMO Speech Enhancement Using…

200 篇论文

Spatial wideband effects are known to affect channel estimation and localization performance in millimeter wave (mmWave) massive multiple-input multiple-output (MIMO) systems. Based on perturbation analysis, we show that the spatial…

信号处理 · 电气工程与系统科学 2022-09-20 Shudi Weng , Fan Jiang , Henk Wymeersch

Audio and omni-modal large language models exhibit impressive cross-modal reasoning capabilities. However, applying standard reinforcement learning post-training algorithms to these models exposes a critical structural vulnerability:…

计算与语言 · 计算机科学 2026-05-28 Cihan Xiao , Yiwen Shao , Chenxing Li , Xiang He , Zhenwen Liang , Steve Yves , Sanjeev Khudanpur , Liefeng Bo

In multi-user millimeter wave (mmWave) multiple-input-multiple-output (MIMO) systems, hybrid precoding is a crucial task to lower the complexity and cost while achieving a sufficient sum-rate. Previous works on hybrid precoding were usually…

信号处理 · 电气工程与系统科学 2020-04-28 Ahmet M. Elbir , Anastasios Papazafeiropoulos

To glean the benefits offered by massive multi-input multi-output (MIMO) systems, channel state information must be accurately acquired. Despite the high accuracy, the computational complexity of classical linear minimum mean squared error…

信息论 · 计算机科学 2024-04-23 Bin Li , Ziping Wei , Shaoshi Yang , Yang Zhang , Jun Zhang , Chenglin Zhao , Sheng Chen

Multimodal representation learning has shown promising improvements on various vision-language tasks. Most existing methods excel at building global-level alignment between vision and language while lacking effective fine-grained image-text…

计算机视觉与模式识别 · 计算机科学 2023-06-16 Zijia Zhao , Longteng Guo , Xingjian He , Shuai Shao , Zehuan Yuan , Jing Liu

In this paper we present a single-microphone speech enhancement algorithm. A hybrid approach is proposed merging the generative mixture of Gaussians (MoG) model and the discriminative neural network (NN). The proposed algorithm is executed…

声音 · 计算机科学 2015-10-27 Shlomo E. Chazan , Jacob Goldberger , Sharon Gannot

This paper considers pilot-based channel estimation in large-scale multiple-input multiple-output (MIMO) communication systems, also known as "massive MIMO". Unlike previous works on this topic, which mainly considered the impact of…

信息论 · 计算机科学 2013-07-18 Nafiseh Shariati , Emil Björnson , Mats Bengtsson , Mérouane Debbah

Audio-visual speech enhancement system is regarded as one of promising solutions for isolating and enhancing speech of desired speaker. Typical methods focus on predicting clean speech spectrum via a naive convolution neural network based…

音频与语音处理 · 电气工程与系统科学 2022-07-01 Xinmeng Xu , Yang Wang , Jie Jia , Binbin Chen , Dejun Li

Goal: Numerous studies had successfully differentiated normal and abnormal voice samples. Nevertheless, further classification had rarely been attempted. This study proposes a novel approach, using continuous Mandarin speech instead of a…

音频与语音处理 · 电气工程与系统科学 2022-02-23 Syu-Siang Wang , Chi-Te Wang , Chih-Chung Lai , Yu Tsao , Shih-Hau Fang

The majority of multichannel speech enhancement algorithms are two-step procedures that first apply a linear spatial filter, a so-called beamformer, and combine it with a single-channel approach for postprocessing. However, the serial…

音频与语音处理 · 电气工程与系统科学 2021-06-16 Kristina Tesch , Timo Gerkmann

Interference during the uplink training phase significantly deteriorates the performance of a massive MIMO system. The impact of the interference can be reduced by exploiting second order statistics of the channel vectors, e.g., to obtain…

信息论 · 计算机科学 2018-05-23 David Neumann , Michael Joham , Wolfgang Utschick

In this paper, we present a novel multi-channel speech extraction system to simultaneously extract multiple clean individual sources from a mixture in noisy and reverberant environments. The proposed method is built on an improved…

音频与语音处理 · 电气工程与系统科学 2021-06-17 Jisi Zhang , Catalin Zorila , Rama Doddipatla , Jon Barker

Orthogonal delay-Doppler division multiplexing~(ODDM) modulation has recently been regarded as a promising technology to provide reliable communications in high-mobility situations. Accurate and low-complexity channel estimation is one of…

信号处理 · 电气工程与系统科学 2025-07-29 Dezhi Wang , Chongwen Huang , Xiaojun Yuan , Sami Muhaidat , Lei Liu , Xiaoming Chen , Zhaoyang Zhang , Chau Yuen , Mérouane Debbah

Holographic multiple-input multiple-output (MIMO) systems represent a spatially constrained MIMO architecture with a massive number of antennas with small antenna spacing as a close approximation of a spatially continuous electromagnetic…

信号处理 · 电气工程与系统科学 2024-12-19 Nikolaos Kolomvakis , Emil Björnson

Audio-visual speech enhancement aims to extract clean speech from a noisy environment by leveraging not only the audio itself but also the target speaker's lip movements. This approach has been shown to yield improvements over audio-only…

The aim of speech enhancement is to improve speech signal quality and intelligibility from a noisy microphone signal. In many applications, it is crucial to enable processing with small computational complexity and minimal requirements…

音频与语音处理 · 电气工程与系统科学 2023-09-08 Julitta Bartolewska , Stanisław Kacprzak , Konrad Kowalczyk

Channel estimation is a critical task in digital communications that greatly impacts end-to-end system performance. In this work, we introduce a novel approach for multiple-input multiple-output (MIMO) channel estimation using score-based…

信号处理 · 电气工程与系统科学 2022-02-16 Marius Arvinte , Jonathan I Tamir

In this paper, we present a method that allows to further improve speech enhancement obtained with recently introduced Deep Neural Network (DNN) models. We propose a multi-channel refinement method of time-frequency masks obtained with…

音频与语音处理 · 电气工程与系统科学 2023-09-19 Julitta Bartolewska , Stanisław Kacprzak , Konrad Kowalczyk

We introduce a novel online convex optimization (OCO) framework to estimate the user's signal-to-interference-plus-noise ratio (SINR) from ACK/NACK feedback, channel quality indicator (CQI) reports, and previously selected modulation and…

信息论 · 计算机科学 2026-03-16 Lorenzo Maggi , Boris Bonev , Reinhard Wiesmayr , Sebastian Cammerer , Alexander Keller

Robust spatial audio control relies on accurate acoustic propagation models, yet environmental variations, especially changes in the speed of sound, cause systematic mismatches that degrade performance. Existing methods either assume known…

音频与语音处理 · 电气工程与系统科学 2026-05-13 Andreas Jonas Fuglsig , Mads Græsbøll Christensen , Jesper Rindom Jensen
‹ 上一页 1 8 9 10 下一页 ›