中文
相关论文

相关论文: Binaural coherent-to-diffuse-ratio estimation for …

200 篇论文

Deep learning technology has been widely applied to speech enhancement. While testing the effectiveness of various network structures, researchers are also exploring the improvement of the loss function used in network training. Although…

音频与语音处理 · 电气工程与系统科学 2023-04-25 Tianrui Wang , Weibin Zhu

We investigate the effects of four strategies for improving the ecological validity of synthetic room impulse response (RIR) datasets for monoaural Speech Enhancement (SE). We implement three features on top of the traditional image source…

声音 · 计算机科学 2025-07-15 Enric Gusó , Joanna Luberadzka , Umut Sayin , Xavier Serra

This work proposes a new learning target based on reverberation time shortening (RTS) for speech dereverberation. The learning target for dereverberation is usually set as the direct-path speech or optionally with some early reflections.…

音频与语音处理 · 电气工程与系统科学 2022-11-22 Rui Zhou , Wenye Zhu , Xiaofei Li

Self-supervised speech representation learning aims to extract meaningful factors from the speech signal that can later be used across different downstream tasks, such as speech and/or emotion recognition. Existing models, such as HuBERT,…

Diffusion probabilistic models have demonstrated an outstanding capability to model natural images and raw audio waveforms through a paired diffusion and reverse processes. The unique property of the reverse process (namely, eliminating…

音频与语音处理 · 电气工程与系统科学 2021-11-23 Yen-Ju Lu , Yu Tsao , Shinji Watanabe

Prior studies on Intelligent Reflecting Surface (IRS) have mostly assumed perfect channel state information (CSI) available for designing the IRS passive beamforming as well as the continuously adjustable phase shift at each of its…

信息论 · 计算机科学 2020-03-25 Changsheng You , Beixiong Zheng , Rui Zhang

State-of-the-art deep-learning-based voice activity detectors (VADs) are often trained with anechoic data. However, real acoustic environments are generally reverberant, which causes the performance to significantly deteriorate. To mitigate…

声音 · 计算机科学 2021-06-28 Amir Ivry , Israel Cohen , Baruch Berdugo

Utilizing intelligent reflecting surface (IRS) was proven to be efficient in improving the energy efficiency for wireless networks. In this paper, we investigate the passive beamforming and channel estimation for IRS assisted wireless…

信息论 · 计算机科学 2021-03-02 Jingnan Li , Rui Wang , Erwu Liu

A Cramer-Rao bound (CRB) for semi-blind channel estimators in redundant block transmission systems is derived. The derived CRB is valid for any system adopting a full-rank linear redundant precoder, including the popular cyclic-prefixed…

信息论 · 计算机科学 2012-09-20 Yen-Huan Li , Borching Su , Ping-Cheng Yeh

Conditional diffusion models are powerful generative models that can leverage various types of conditional information, such as class labels, segmentation masks, or text captions. However, in many real-world scenarios, conditional…

计算机视觉与模式识别 · 计算机科学 2025-02-19 Nicolas Dufour , Victor Besnier , Vicky Kalogeiton , David Picard

Hand-crafted spatial features, such as inter-channel intensity difference (IID) and inter-channel phase difference (IPD), play a fundamental role in recent deep learning based dual-microphone speech enhancement (DMSE) systems. However,…

音频与语音处理 · 电气工程与系统科学 2022-05-04 Xinmeng Xu , Rongzhi Gu , Yuexian Zou

While deep neural networks (NN) significantly advance image compressed sensing (CS) by improving reconstruction quality, the necessity of training current CS NNs from scratch constrains their effectiveness and hampers rapid deployment.…

计算机视觉与模式识别 · 计算机科学 2025-02-04 Bin Chen , Zhenyu Zhang , Weiqi Li , Chen Zhao , Jiwen Yu , Shijie Zhao , Jie Chen , Jian Zhang

In this paper, we propose a multi-channel network for simultaneous speech dereverberation, enhancement and separation (DESNet). To enable gradient propagation and joint optimization, we adopt the attentional selection mechanism of the…

声音 · 计算机科学 2020-11-17 Yihui Fu , Jian Wu , Yanxin Hu , Mengtao Xing , Lei Xie

Speech quality and intelligibility are significantly degraded in noisy environments. This paper presents a novel transformer-based learning framework to address the single-channel noise suppression problem for real-time applications.…

声音 · 计算机科学 2025-11-18 Behnaz Bahmei , Siamak Arzanpour , Elina Birmingham

This work presents a method for designing the weighting parameter required by Wiener-based binaural noise reduction methods. This parameter establishes the desired tradeoff between noise reduction and binaural cue preservation in hearing…

音频与语音处理 · 电气工程与系统科学 2021-04-21 Diego M. Carmo , Ricardo Borsoi , Márcio H. Costa

We suggest double/debiased machine learning estimators of direct and indirect quantile treatment effects under a selection-on-observables assumption. This permits disentangling the causal effect of a binary treatment at a specific outcome…

计量经济学 · 经济学 2023-07-04 Yu-Chin Hsu , Martin Huber , Yu-Min Yen

Estimating the position of a speech source based on time-differences-of-arrival (TDOAs) is often adversely affected by background noise and reverberation. A popular method to estimate the TDOA between a microphone pair involves maximizing a…

音频与语音处理 · 电气工程与系统科学 2025-07-28 Klaus Brümann , Kouei Yamaoka , Nobutaka Ono , Simon Doclo

This note introduces a doubly robust (DR) estimator for regression discontinuity (RD) designs. RD designs provide a quasi-experimental framework for estimating treatment effects, where treatment assignment depends on whether a running…

计量经济学 · 经济学 2025-01-28 Masahiro Kato

Distributed integrated sensing and communication (D-ISAC) enables multiple spatially distributed nodes to cooperatively perform sensing and communication. However, achieving coherent cooperation across distributed nodes is challenging due…

系统与控制 · 电气工程与系统科学 2026-04-06 Seonghoon Yoo , Seulhyun Kwon , Kawon Han , Elaheh Ataeebojd , Mehdi Rasti , Joonhyuk Kang

In this paper we consider a binaural hearing aid setup, where in addition to the head-mounted microphones an external microphone is available. For this setup, we investigate the performance of several relative transfer function (RTF) vector…

音频与语音处理 · 电气工程与系统科学 2022-05-19 Daniel Fejgin , Simon Doclo
‹ 上一页 1 8 9 10 下一页 ›