中文
相关论文

相关论文: Direction-Preserving MIMO Speech Enhancement Using…

200 篇论文

We present a hybrid framework that leverages the trade-off between temporal and frequency precision in audio representations to improve the performance of speech enhancement task. We first show that conventional approaches using specific…

音频与语音处理 · 电气工程与系统科学 2018-12-24 Jang-Hyun Kim , Jaejun Yoo , Sanghyuk Chun , Adrian Kim , Jung-Woo Ha

This paper addresses the problem of multi-channel multi-speech separation based on deep learning techniques. In the short time Fourier transform domain, we propose an end-to-end narrow-band network that directly takes as input the…

声音 · 计算机科学 2022-04-13 Changsheng Quan , Xiaofei Li

Channel estimation is a critical task in multiple-input multiple-output (MIMO) digital communications that substantially effects end-to-end system performance. In this work, we introduce a novel approach for channel estimation using deep…

信号处理 · 电气工程与系统科学 2022-11-09 Marius Arvinte , Jonathan I Tamir

Recent progress on end-to-end neural diarization (EEND) has enabled overlap-aware speaker diarization with a single neural network. This paper proposes to enhance EEND by using multi-channel signals from distributed microphones. We replace…

音频与语音处理 · 电气工程与系统科学 2022-03-29 Shota Horiguchi , Yuki Takashima , Paola Garcia , Shinji Watanabe , Yohei Kawaguchi

Audio-visual speech enhancement (AVSE) has been found to be particularly useful at low signal-to-noise (SNR) ratios due to the immunity of the visual features to acoustic noise. However, a significant gap exists in AVSE methods tailored to…

音频与语音处理 · 电气工程与系统科学 2025-10-21 Danielle Yaffe , Ferdinand Campe , Prachi Sharma , Dorothea Kolossa , Boaz Rafaely

Multichannel speech enhancement algorithms are essential for improving the intelligibility of speech signals in noisy environments. These algorithms are usually evaluated at the utterance level, but this approach overlooks the disparities…

声音 · 计算机科学 2025-06-24 Nasser-Eddine Monir , Paul Magron , Romain Serizel

Intelligent reflecting surfaces (IRSs) are regarded as promising enablers for future millimeter wave (mmWave) wireless communication, due to their ability to create favorable line-of-sight (LoS) propagation environments. In this paper, we…

信息论 · 计算机科学 2020-05-12 Tian Lin , Xianghao Yu , Yu Zhu , Robert Schober

This paper investigates the robust wideband channel estimation problem in the millimeter-wave (mmWave) massive multiple-input multiple-output (MIMO) systems. In such a scenario, the beam squint effect that the array response vectors vary…

信号处理 · 电气工程与系统科学 2024-09-25 Li Ge , Lin Chen , Xue Jiang , Weifeng Zhu , Qibo Qin , Xingzhao Liu

Multiple input multiple output (MIMO) system transmission is a popular diversity technique to improve the reliability of a communication system where transmitter, communication channel and receiver are the important elements. Data…

信号处理 · 电气工程与系统科学 2020-05-06 Subrato Bharati , Prajoy Podder , Niketa Gandhi , Ajith Abraham

Speech recognition in adverse real-world environments is highly affected by reverberation and nonstationary background noise. A well-known strategy to reduce such undesired signal components in multi-microphone scenarios is spatial…

声音 · 计算机科学 2017-08-08 Hendrik Barfuss , Christian Huemmer , Andreas Schwarz , Walter Kellermann

We investigate a general channel estimation problem in the massive multiple-input multiple-output (MIMO) system which employs the hybrid analog/digital precoding structure with limited radio-frequency (RF) chains. By properly designing RF…

信息论 · 计算机科学 2017-12-27 Leyuan Pan , Le Liang , Wei Xu , Xiaodai Dong

With the development of deep learning, speech enhancement has been greatly optimized in terms of speech quality. Previous methods typically focus on the discriminative supervised learning or generative modeling, which tends to introduce…

音频与语音处理 · 电气工程与系统科学 2025-10-31 Nan Xu , Zhaolong Huang , Xiaonan Zhi

In this work, we extend our previously proposed offline SpatialNet for long-term streaming multichannel speech enhancement in both static and moving speaker scenarios. SpatialNet exploits spatial information, such as the spatial/steering…

声音 · 计算机科学 2024-06-21 Changsheng Quan , Xiaofei Li

This paper addresses channel estimation and data equalization on frequency-selective 1-bit quantized Multiple Input-Multiple Output (MIMO) systems. No joint processing or Channel State Information is assumed at the transmitter, and…

信息论 · 计算机科学 2021-03-09 Javier García , Jawad Munir , Kilian Roth , Josef A. Nossek

Millimeter Wave (mmWave) massive Multiple Input Multiple Output (MIMO) systems realizing directive beamforming require reliable estimation of the wireless propagation channel. However, mmWave channels are characterized by high variability…

信息论 · 计算机科学 2019-06-07 Evangelos Vlachos , George C. Alexandropoulos , John Thompson

Casual conversations involving multiple speakers and noises from surrounding devices are common in everyday environments, which degrades the performances of automatic speech recognition systems. These challenging characteristics of…

音频与语音处理 · 电气工程与系统科学 2019-06-24 Nelson Yalta , Shinji Watanabe , Takaaki Hori , Kazuhiro Nakadai , Tetsuya Ogata

The state-of-art methods for acoustic beamforming in multi-channel ASR are based on a neural mask estimator that predicts the presence of speech and noise. These models are trained using a paired corpus of clean and noisy recordings…

音频与语音处理 · 电气工程与系统科学 2019-12-02 Rohit Kumar , Anirudh Sreeram , Anurenjan Purushothaman , Sriram Ganapathy

In this paper, we investigate the blind channel estimation problem for MIMO systems under Rayleigh fading channel. Conventional MIMO communication techniques require transmitting a considerable amount of training symbols as pilots in each…

信号处理 · 电气工程与系统科学 2021-11-17 Jiancheng Tang , Qianqian Yang , Zhaoyang Zhang

Acoustic echo cancellation (AEC) is an important speech signal processing technology that can remove echoes from microphone signals to enable natural-sounding full-duplex speech communication. While single-channel AEC is widely adopted,…

声音 · 计算机科学 2025-06-09 Fei Zhao , Xueliang Zhang , Zhong-Qiu Wang

Single-channel speech enhancement approaches do not always improve automatic recognition rates in the presence of noise, because they can introduce distortions unhelpful for recognition. Following a trend towards end-to-end training of…

声音 · 计算机科学 2021-12-14 Peter Plantinga , Deblin Bagchi , Eric Fosler-Lussier