English
Related papers

Related papers: Robust parameter design for Wiener-based binaural …

200 papers

In this work, we introduce a novel framework which combines physics and machine learning methods to analyse acoustic signals. Three methods are developed for this task: a Bayesian inference approach for inferring the spectral acoustics…

Sound · Computer Science 2023-05-30 Yongchao Huang , Yuhang He , Hong Ge

Sound field reconstruction refers to the problem of estimating the acoustic pressure field over an arbitrary region of space, using only a limited set of measurements. Physics-informed neural networks have been adopted to solve the problem…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-05 Stefano Damiano , Toon van Waterschoot

In this paper, a novel architecture for speaker recognition is proposed by cascading speech enhancement and speaker processing. Its aim is to improve speaker recognition performance when speech signals are corrupted by noise. Instead of…

Computation and Language · Computer Science 2020-05-25 Yanpei Shi , Qiang Huang , Thomas Hain

This paper proposes a speech enhancement method which exploits the high potential of residual connections in a Wide Residual Network architecture. This is supported on single dimensional convolutions computed alongside the time domain,…

Audio and Speech Processing · Electrical Eng. & Systems 2019-04-11 Jorge Llombart , Dayana Ribas , Antonio Miguel , Luis Vicente , Alfonso Ortega , Eduardo Lleida

We simulate the effects of different types of noise in state preparation circuits of variational quantum algorithms. We first use a variational quantum eigensolver to find the ground state of a Hamiltonian in presence of noise, and adopt…

Quantum Physics · Physics 2021-08-11 Enrico Fontana , Nathan Fitzpatrick , David Muñoz Ramo , Ross Duncan , Ivan Rungger

In several studies, hybrid neural networks have proven to be more robust against noisy input data compared to plain data driven neural networks. We consider the task of estimating parameters of a mechanical vehicle model based on…

Machine Learning · Computer Science 2020-04-17 Jan Sokolowski , Volker Schulz , Udo Schröder , Hans-Peter Beise

One well established method of interactive image segmentation is the random walker algorithm. Considerable research on this family of segmentation methods has been continuously conducted in recent years with numerous applications. These…

Computer Vision and Pattern Recognition · Computer Science 2022-06-03 Dominik Drees , Florian Eilers , Ang Bian , Xiaoyi Jiang

This study proposes a framework for incorporating wavenumber-domain acoustic reflection coefficients into sound field analysis to characterize direction-dependent material reflection and scattering phenomena. The reflection coefficient is…

Audio and Speech Processing · Electrical Eng. & Systems 2026-01-13 Satoshi Hoshika , Takahiro Iwami , Akira Omoto

Robust audio-visual speech recognition (AVSR) in noisy environments remains challenging, as existing systems struggle to estimate audio reliability and dynamically adjust modality reliance. We propose router-gated cross-modal feature…

Computer Vision and Pattern Recognition · Computer Science 2025-08-27 DongHoon Lim , YoungChae Kim , Dong-Hyun Kim , Da-Hee Yang , Joon-Hyuk Chang

We study identification of stochastic Wiener dynamic systems using so-called indirect inference. The main idea is to first fit an auxiliary model to the observed data and then in a second step, often by simulation, fit a more structured…

Optimization and Control · Mathematics 2015-07-21 Bo Wahlberg , James Welsh , Lennart Ljung

Variational quantum machine learning algorithms have become the focus of recent research on how to utilize near-term quantum devices for machine learning tasks. They are considered suitable for this as the circuits that are run can be…

Quantum Physics · Physics 2022-12-20 Andrea Skolik , Stefano Mangini , Thomas Bäck , Chiara Macchiavello , Vedran Dunjko

In this paper, the capacity of the additive white Gaussian noise (AWGN) channel, affected by time-varying Wiener phase noise is investigated. Tight upper and lower bounds on the capacity of this channel are developed. The upper bound is…

Information Theory · Computer Science 2015-09-22 M. Reza Khanzadi , Rajet Krishnan , Johan Söder , Thomas Eriksson

Dynamic parameterization of acoustic environments has drawn widespread attention in the field of audio processing. Precise representation of local room acoustic characteristics is crucial when designing audio filters for various audio…

Audio and Speech Processing · Electrical Eng. & Systems 2024-04-26 Chunxi Wang , Maoshen Jia , Meiran Li , Changchun Bao , Wenyu Jin

Over the recent years, various deep learning-based methods were proposed for extracting a fixed-dimensional embedding vector from speech signals. Although the deep learning-based embedding extraction methods have shown good performance in…

Audio and Speech Processing · Electrical Eng. & Systems 2021-12-08 Woo Hyun Kang , Jahangir Alam , Abderrahim Fathan

Eurich et al. (2024) recently introduced the computationally efficient monaural and binaural audio quality model (eMoBi-Q). This model integrates both monaural and binaural auditory features and has been validated across six audio datasets…

Audio and Speech Processing · Electrical Eng. & Systems 2025-12-05 Thomas Biberger , Stephan D. Ewert

Hierarchical quantum classifiers, such as quantum convolutional neural networks (QCNNs), represent recent progress toward designing effective and feasible architectures for quantum classification. However, their performance on near-term…

Quantum Physics · Physics 2026-02-26 Taehyun Kim , Israel F. Araujo , Daniel K. Park

The expanding feature set of modern headphones puts a challenge on the design of their control interface. Users may want to separately control each feature or quickly switch between modes that activate different features. Traditional…

Audio and Speech Processing · Electrical Eng. & Systems 2025-03-04 Qiaoyu Yang

Standard system identification methods often provide inconsistent estimates with closed-loop data. With the prediction error method (PEM), this issue is solved by using a noise model that is flexible enough to capture the noise spectrum.…

Systems and Control · Computer Science 2018-09-07 Miguel Galrinho , Cristian R. Rojas , Hakan Hjalmarsson

We propose a diffractive neural network with strong robustness based on Weight Noise Injection training, which achieves accurate and fast optical-based classification while diffraction layers have a certain amount of surface shape error. To…

Image and Video Processing · Electrical Eng. & Systems 2020-06-23 Jiashuo Shi

Most previously proposed dual-channel coherent-to-diffuse-ratio (CDR) estimators are based on a free-field model. When used for binaural signals, e.g., for dereverberation in binaural hearing aids, their performance may degrade due to the…

Sound · Computer Science 2015-06-12 Chengshi Zheng , Andreas Schwarz , Walter Kellermann , Xiaodong Li