English
Related papers

Related papers: Time-Domain Wideband Image Source Method for Spher…

200 papers

Diffusion Probabilistic Models have demonstrated remarkable performance across a wide range of generative tasks. However, we have observed that these models often suffer from a Signal-to-Noise Ratio-timestep (SNR-t) bias. This bias refers…

Computer Vision and Pattern Recognition · Computer Science 2026-04-20 Meng Yu , Lei Sun , Jianhao Zeng , Xiangxiang Chu , Kun Zhan

Since space-domain information can be utilized, microphone array beamforming is often used to enhance the quality of the speech by suppressing directional disturbance. However, with the increasing number of microphone, the complexity would…

Sound · Computer Science 2020-05-20 Lu Ma , Xin Zhao , Pei Zhao , Tengrong Su

Speech super-resolution (SR) reconstructs high-fidelity wideband speech from low-resolution inputs-a task that necessitates reconciling global harmonic coherence with local transient sharpness. While diffusion-based generative models yield…

Sound · Computer Science 2026-01-01 Jiajun Yuan , Xiaochen Wang , Yuhang Xiao , Yulin Wu , Chenhao Hu , Xueyang Lv

The whispering gallery modes (WGMs) of optical resonators have prompted intensive research efforts due to their usefulness in the field of biological sensing, and their employment in nonlinear optics. While much information is available in…

Time series analysis plays a critical role in numerous applications, supporting tasks such as forecasting, classification, anomaly detection, and imputation. In this work, we present the time series pattern machine (TSPM), a model designed…

Machine Learning · Computer Science 2025-05-20 Shiyu Wang , Jiawei Li , Xiaoming Shi , Zhou Ye , Baichuan Mo , Wenze Lin , Shengtong Ju , Zhixuan Chu , Ming Jin

Room impulse responses (RIRs) are essential for many acoustic signal processing tasks, yet measuring them densely across space is often impractical. In this work, we propose RIR-Former, a grid-free, one-step feed-forward model for RIR…

Audio and Speech Processing · Electrical Eng. & Systems 2026-05-12 Shaoheng Xu , Chunyi Sun , Jihui Zhang , Prasanga N. Samarasinghe , Thushara D. Abhayapala

We present in this paper an informed single-channel dereverberation method based on conditional generation with diffusion models. With knowledge of the room impulse response, the anechoic utterance is generated via reverse diffusion using a…

Audio and Speech Processing · Electrical Eng. & Systems 2023-06-22 Jean-Marie Lemercier , Simon Welker , Timo Gerkmann

In this paper we present a unified time-frequency method for speaker extraction in clean and noisy conditions. Given a mixed signal, along with a reference signal, the common approaches for extracting the desired speaker are either applied…

Sound · Computer Science 2022-03-08 Aviad Eisenberg , Sharon Gannot , Shlomo E. Chazan

This paper presents a novel Diffusion-Wavelet (DiWa) approach for Single-Image Super-Resolution (SISR). It leverages the strengths of Denoising Diffusion Probabilistic Models (DDPMs) and Discrete Wavelet Transformation (DWT). By enabling…

Computer Vision and Pattern Recognition · Computer Science 2024-09-24 Brian Moser , Stanislav Frolov , Federico Raue , Sebastian Palacio , Andreas Dengel

Millimeter-wave (mmWave) communication is one of the key enablers of the fifth-generation cellular networks (5G). However, one of the fundamental challenges of mmWave communication is the susceptibility to blockage effects. One way to…

Information Theory · Computer Science 2019-12-30 Yaoshen Cui , Haifan Yin

Conventional microwave imaging schemes, enabled by the ubiquity of coherent sources and detectors, have traditionally relied on frequency bandwidth to retrieve range information, while using mechanical or electronic beamsteering to obtain…

Millimeter-wave (mmWave) and sub-Terahertz (THz) frequencies are expected to play a vital role in 6G wireless systems and beyond due to the vast available bandwidth of many tens of GHz. This paper presents an indoor 3-D spatial statistical…

Information Theory · Computer Science 2021-04-01 Shihao Ju , Yunchou Xing , Ojas Kanhere , Theodore S. Rappaport

In this paper, we introduce a multi-talker distant automatic speech recognition (DASR) system we designed for the DASR task 1 of the CHiME-8 challenge. Our system performs speaker counting, diarization, and ASR. It handles various recording…

In this paper, we present a novel multi-channel speech extraction system to simultaneously extract multiple clean individual sources from a mixture in noisy and reverberant environments. The proposed method is built on an improved…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-17 Jisi Zhang , Catalin Zorila , Rama Doddipatla , Jon Barker

Target speech separation refers to extracting the target speaker's speech from mixed signals. Despite the recent advances in deep learning based close-talk speech separation, the applications to real-world are still an open issue. Two main…

Sound · Computer Science 2020-01-03 Rongzhi Gu , Yuexian Zou

Millimeter-wave communications rely on beamforming gain from both transmitters and receivers to compensate for severe propagation loss. To achieve adequate gain, beam training is required to identify propagation directions. The main…

Signal Processing · Electrical Eng. & Systems 2020-04-06 Han Yan , Veljko Boljanovic , Danijela Cabric

Room equalisation aims to increase the quality of loudspeaker reproduction in reverberant environments, compensating for colouration caused by imperfect room reflections and frequency dependant loudspeaker directivity. A common technique in…

Audio and Speech Processing · Electrical Eng. & Systems 2024-09-17 James Brooks-Park , Martin Bo Møller , Jan Østergaard , Søren Bech , Steven van de Par

The time reversal symmetry of the wave equation allows wave refocusing back at the source. However, this symmetry does not hold in lossy media. We present a new strategy to compensate wave amplitude losses due to attenuation. The strategy…

Classical Physics · Physics 2022-12-15 Crystal T. Wu , Nuno M. Nobre , Emmanuel Fort , Graham D. Riley , Fumie Costen

The source separation-based speech enhancement problem with multiple beamforming in reverberant indoor environments is addressed in this paper. We propose that more generic solutions should cope with time-varying dynamic scenarios with…

Audio and Speech Processing · Electrical Eng. & Systems 2020-11-05 Alejandro Díaz , Diego Pincheira , Rodrigo Mahu , Nestor Becerra Yoma

We propose a diarization system, that estimates "who spoke when" based on spatial information, to be used as a front-end of a meeting transcription system running on the signals gathered from an acoustic sensor network (ASN). Although the…

Audio and Speech Processing · Electrical Eng. & Systems 2023-11-28 Tobias Gburrek , Joerg Schmalenstroeer , Reinhold Haeb-Umbach