English
Related papers

Related papers: A Time-domain Real-valued Generalized Wiener Filte…

200 papers

Solving inverse problems is central to a variety of important applications, such as biomedical image reconstruction and non-destructive testing. These problems are characterized by the sensitivity of direct solution methods with respect to…

Numerical Analysis · Mathematics 2023-05-17 Simon Göppel , Jürgen Frikel , Markus Haltmeier

A discrete-time end-to-end fiber-optical channel model is derived based on the first-order perturbation approach. The model relates the discrete-time input symbol sequences of co-propagating wavelength channels to the received symbol…

Signal Processing · Electrical Eng. & Systems 2020-03-18 Felix Frey , Johannes K. Fischer , Robert F. H. Fischer

In this contribution to the 3rd CHiME Speech Separation and Recognition Challenge (CHiME-3) we extend the acoustic front-end of the CHiME-3 baseline speech recognition system by a coherence-based Wiener filter which is applied to the output…

Sound · Computer Science 2015-09-24 Hendrik Barfuss , Christian Huemmer , Andreas Schwarz , Walter Kellermann

Accurate and reliable identification of the relative transfer functions (RTFs) between microphones with respect to a desired source is an essential component in the design of microphone array beamformers, specifically when applying the…

Audio and Speech Processing · Electrical Eng. & Systems 2024-12-19 Daniel Levi , Amit Sofer , Sharon Gannot

While existing end-to-end beamformers achieve impressive performance in various front-end speech processing tasks, they usually encapsulate the whole process into a black box and thus lack adequate interpretability. As an attempt to fill…

Sound · Computer Science 2022-03-17 Andong Li , Guochen Yu , Chengshi Zheng , Xiaodong Li

While recent progresses in neural network approaches to single-channel speech separation, or more generally the cocktail party problem, achieved significant improvement, their performance for complex mixtures is still not satisfactory. In…

Sound · Computer Science 2018-03-30 Zhuo Chen , Jinyu Li , Xiong Xiao , Takuya Yoshioka , Huaming Wang , Zhenghao Wang , Yifan Gong

Time-frequency (TF) domain dual-path models achieve high-fidelity speech separation. While some previous state-of-the-art (SoTA) models rely on RNNs, this reliance means they lack the parallelizability, scalability, and versatility of…

Audio and Speech Processing · Electrical Eng. & Systems 2024-08-08 Kohei Saijo , Gordon Wichern , François G. Germain , Zexu Pan , Jonathan Le Roux

The rapid advancement of generative artificial intelligence has enabled the creation of highly realistic fake facial images, posing serious threats to personal privacy and the integrity of online information. Existing deepfake detection…

Computer Vision and Pattern Recognition · Computer Science 2025-12-29 Huanhuan Yuan , Yang Ping , Zhengqin Xu , Junyi Cao , Shuai Jia , Chao Ma

In this paper, the problem of sequential beam construction and adaptive channel estimation based on reduced rank (RR) Kalman filtering for frequency-selective massive multiple-input multiple-output (MIMO) systems employing single-carrier…

Information Theory · Computer Science 2017-03-10 Gokhan M. Guvensen , Ender Ayanoglu

Neural channel decoder, as a data-driven channel decoding strategy, has shown very promising improvement on error-correcting capability over the classical methods. However, the success of those deep learning-based decoder comes at the cost…

Information Theory · Computer Science 2026-05-20 Chengwei Zhang , Yifan Du , Siyu Liao

Deep learning model (primarily convolutional networks and LSTM) for time series classification has been studied broadly by the community with the wide applications in different domains like healthcare, finance, industrial engineering and…

Machine Learning · Computer Science 2021-03-29 Minghao Liu , Shengqi Ren , Siyuan Ma , Jiahui Jiao , Yizhou Chen , Zhiguang Wang , Wei Song

Recent proposals of deep beamformers using deep neural networks have attracted significant attention as computational efficient alternatives to adaptive and compressive beamformers. Moreover, deep beamformers are versatile in that image…

Image and Video Processing · Electrical Eng. & Systems 2020-09-07 Shujaat Khan , Jaeyoung Huh , Jong Chul Ye

Generic Boundary Detection (GBD) aims at locating the general boundaries that divide videos into semantically coherent and taxonomy-free units, and could serve as an important pre-processing step for long-form video understanding. Previous…

Computer Vision and Pattern Recognition · Computer Science 2022-10-06 Jing Tan , Yuhong Wang , Gangshan Wu , Limin Wang

High-Frequency (HF) signals are ubiquitous in the industrial world and are of great use for monitoring of industrial assets. Most deep learning tools are designed for inputs of fixed and/or very limited size and many successful applications…

Machine Learning · Computer Science 2022-03-03 Gabriel Michau , Gaetan Frusque , Olga Fink

Speech separation remains an important topic for multi-speaker technology researchers. Convolution augmented transformers (conformers) have performed well for many speech processing tasks but have been under-researched for speech…

Sound · Computer Science 2023-10-11 William Ravenscroft , Stefan Goetze , Thomas Hain

True-time delayers (TTDs) are popular analog devices for facilitating near-field wideband beamforming subject to the spatial-wideband effect. In this paper, an adaptive TTD configuration is proposed for short-range TTDs. Compared to the…

Information Theory · Computer Science 2024-03-28 Hsienchih Ting , Zhaolin Wang , Yuanwei Liu

In this paper, we propose a phase shift deep neural network (PhaseDNN), which provides a uniform wideband convergence in approximating high frequency functions and solutions of wave equations. The PhaseDNN makes use of the fact that common…

Machine Learning · Computer Science 2019-12-17 Wei Cai , Xiaoguang Li , Lizuo Liu

We propose novel two-channel filter banks for signals on graphs. Our designs can be applied to arbitrary graphs, given a positive semi definite variation operator, while using arbitrary vertex partitions for downsampling. The proposed…

Signal Processing · Electrical Eng. & Systems 2023-04-26 Eduardo Pavez , Benjamin Girault , Antonio Ortega , Philip A. Chou

Diffusion Weighted Imaging (DWI) is an advanced imaging technique commonly used in neuroscience and neurological clinical research through a Diffusion Tensor Imaging (DTI) model. Volumetric scalar metrics including fractional anisotropy,…

Image and Video Processing · Electrical Eng. & Systems 2022-11-01 Zihao Tang , Xinyi Wang , Lihaowen Zhu , Mariano Cabezas , Dongnan Liu , Michael Barnett , Weidong Cai , Chengyu Wang

Multivariate time series classification (MTSC) plays a crucial role in various domains, including biomedical signal analysis and motion monitoring. However, existing approaches, particularly deep learning models, often require high…

Machine Learning · Computer Science 2026-04-20 Fernando Moro , Vinicius M. A. Souza