English
Related papers

Related papers: Complex ISNMF: a Phase-Aware Model for Monaural Au…

200 papers

A waveform channel is considered where the transmitted signal is corrupted by Wiener phase noise and additive white Gaussian noise. A discrete-time channel model that takes into account the effect of filtering on the phase noise is…

Information Theory · Computer Science 2017-08-15 Hassan Ghozlan , Gerhard Kramer

In this paper, we develop a generalized Bayesian inference framework for a collection of signal-plus-noise matrix models arising in high-dimensional statistics and many applications. The framework is built upon an asymptotically unbiased…

Statistics Theory · Mathematics 2022-04-01 Fangzheng Xie , Dingbo Wu

Automatic music transcription (AMT) aims to infer a latent symbolic representation of a piece of music (piano-roll), given a corresponding observed audio recording. Transcribing polyphonic music (when multiple notes are played…

Machine Learning · Statistics 2018-11-19 Pablo A. Alvarado , Dan Stowell

The so-called independent low-rank matrix analysis (ILRMA) has demonstrated a great potential for dealing with the problem of determined blind source separation (BSS) for audio and speech signals. This method assumes that the spectra from…

Sound · Computer Science 2024-01-04 Jianyu Wang , Shanzheng Guan , Jingdong Chen , Jacob Benesty

Source apportionment analysis, which aims to quantify the attribution of observed concentrations of multiple air pollutants to specific sources, can be formulated as a non-negative matrix factorization (NMF) problem. However, NMF is…

Statistics Theory · Mathematics 2025-10-08 Bora Jin , Abhirup Datta

In this paper we address the problems of modeling the acoustic space generated by a full-spectrum sound source and of using the learned model for the localization and separation of multiple sources that simultaneously emit sparse-spectrum…

Sound · Computer Science 2015-02-06 Antoine Deleforge , Florence Forbes , Radu Horaud

Informed speaker extraction aims to extract a target speech signal from a mixture of sources given prior knowledge about the desired speaker. Recent deep learning-based methods leverage a speaker discriminative model that maps a reference…

Audio and Speech Processing · Electrical Eng. & Systems 2022-02-17 Mohamed Elminshawi , Wolfgang Mack , Emanuël A. P. Habets

We introduce a noise-aware extension to the parametric maximum-likelihood framework for component separation by modeling correlated $1/f^\alpha$ noise as a harmonic-space power law. This approach addresses a key limitation of existing…

Cosmology and Nongalactic Astrophysics · Physics 2025-11-07 Goureesankar Sathyanathan , Josquin Errard , Soumen Basak

Beta process is the standard nonparametric Bayesian prior for latent factor model. In this paper, we derive a structured mean-field variational inference algorithm for a beta process non-negative matrix factorization (NMF) model with…

Machine Learning · Statistics 2014-12-03 Dawen Liang , Matthew D. Hoffman

We present a unifying framework that bridges Bayesian asymptotics and information theory to analyze the asymptotic Shannon capacity of general large-scale MIMO channels including ones with nonlinearities or imperfect hardware. We derive…

Information Theory · Computer Science 2026-02-24 Sheng Yang , Richard Combes

This paper proposes a novel Stage-wise and Prior-aware Neural Speech Phase Prediction (SP-NSPP) model, which predicts the phase spectrum from input amplitude spectrum by two-stage neural networks. In the initial prior-construction stage, we…

Sound · Computer Science 2024-10-08 Fei Liu , Yang Ai , Hui-Peng Du , Ye-Xin Lu , Rui-Chen Zheng , Zhen-Hua Ling

In many dynamical probes of a quantum system, quite often multiple eigenmodes are excited. Therefore, the experimental data can be quite messy due to the mixing of different modes, as well as the background noise, despite that each mode…

Quantum Physics · Physics 2019-04-11 Yadong Wu , Hui Zhai

As an alternative but unified and more fundamental description for quantum physics, Feynman path integrals generalize the classical action principle to a probabilistic perspective, under which the physical observables' estimation translates…

High Energy Physics - Lattice · Physics 2023-03-03 Shile Chen , Oleh Savchuk , Shiqi Zheng , Baoyi Chen , Horst Stoecker , Lingxiao Wang , Kai Zhou

Deriving a good model for multitalker babble noise can facilitate different speech processing algorithms, e.g. noise reduction, to reduce the so-called cocktail party difficulty. In the available systems, the fact that the babble waveform…

Sound · Computer Science 2017-09-19 Nasser Mohammadiha , Arne Leijon

We propose an independence-based joint dereverberation and separation method with a neural source model. We introduce a neural network in the framework of time-decorrelation iterative source steering, which is an extension of independent…

Audio and Speech Processing · Electrical Eng. & Systems 2022-04-04 Kohei Saijo , Robin Scheibler

Despite phenomenal progress in recent years, state-of-the-art music separation systems produce source estimates with significant perceptual shortcomings, such as adding extraneous noise or removing harmonics. We propose a post-processing…

Sound · Computer Science 2022-08-29 Noah Schaffer , Boaz Cogan , Ethan Manilow , Max Morrison , Prem Seetharaman , Bryan Pardo

In this work, with combined belief propagation (BP), mean field (MF) and expectation propagation (EP), an iterative receiver is designed for joint phase noise (PN) estimation, equalization and decoding in a coded communication system. The…

Information Theory · Computer Science 2016-09-07 Wei Wang , Zhongyong Wang , Chuanzong Zhang , Qinghua Guo , Peng Sun , Xingye Wang

Separating vocal elements from musical tracks is a longstanding challenge in audio signal processing. This study tackles the distinct separation of vocal components from musical spectrograms. We employ the Short Time Fourier Transform…

Sound · Computer Science 2024-05-31 Adam Sorrenti

We propose a unified model for three inter-related tasks: 1) to \textit{separate} individual sound sources from a mixed music audio, 2) to \textit{transcribe} each sound source to MIDI notes, and 3) to\textit{ synthesize} new pieces based…

Sound · Computer Science 2021-08-10 Liwei Lin , Qiuqiang Kong , Junyan Jiang , Gus Xia

Nonnegative matrix factorization (NMF) is a linear dimensionality reduction technique for analyzing nonnegative data. A key aspect of NMF is the choice of the objective function that depends on the noise model (or statistics of the noise)…

Machine Learning · Computer Science 2021-02-10 Nicolas Gillis , Le Thi Khanh Hien , Valentin Leplat , Vincent Y. F. Tan