English
Related papers

Related papers: Switching Independent Vector Analysis and Its Exte…

200 papers

Although deep learning based multi-channel speech enhancement has achieved significant advancements, its practical deployment is often limited by constrained computational resources, particularly in low signal-to-noise ratio (SNR)…

Audio and Speech Processing · Electrical Eng. & Systems 2025-05-27 Zheng Wang , Xiaobin Rong , Yu Sun , Tianchi Sun , Zhibin Lin , Jing Lu

A novel extension of Independent Component and Independent Vector Analysis for blind extraction/separation of one or several sources from time-varying mixtures is proposed. The mixtures are assumed to be separable source-by-source in series…

Signal Processing · Electrical Eng. & Systems 2021-05-12 Zbyněk Koldovský , Václav Kautský , Petr Tichavský

Blind source separation (BSS) refers to the process of recovering multiple source signals from observations recorded by an array of sensors. Common approaches to BSS, including independent vector analysis (IVA), and independent low-rank…

Sound · Computer Science 2025-11-11 Jianyu Wang , Shanzheng Guan , Nicolas Dobigeon , Jingdong Chen

This paper presents a computationally efficient approach to blind source separation (BSS) of audio signals, applicable even when there are more sources than microphones (i.e., the underdetermined case). When there are as many sources as…

Sound · Computer Science 2021-01-22 Nobutaka Ito , Rintaro Ikeshita , Hiroshi Sawada , Tomohiro Nakatani

In this paper, we propose a new online independent vector analysis (IVA) algorithm for real-time blind source separation (BSS). In many BSS algorithms, the iterative projection (IP) has been used for updating the demixing matrix, a…

Audio and Speech Processing · Electrical Eng. & Systems 2022-09-05 Taishi Nakashima , Nobutaka Ono

Multichannel blind source separation (MBSS), which focuses on separating signals of interest from mixed observations, has been extensively studied in acoustic and speech processing. Existing MBSS algorithms, such as independent low-rank…

Sound · Computer Science 2025-04-08 Jianyu Wang , Shanzheng Guan , Zhengqiao Zhao , Nicolas Dobigeon , Jingdong Chen

This manuscript proposes a novel robust procedure for the extraction of a speaker of interest (SOI) from a mixture of audio sources. The estimation of the SOI is performed via independent vector extraction (IVE). Since the blind IVE cannot…

Audio and Speech Processing · Electrical Eng. & Systems 2022-07-29 Jiri Malek , Jakub Jansky , Zbynek Koldovsky , Tomas Kounovsky , Jaroslav Cmejla , Jindrich Zdansky

This letter proposes a new blind source separation (BSS) framework termed minimum variance independent component analysis (MVICA), which can potentially achieve the maximum output signal-to-interference ratio (SIR) while also allowing more…

Sound · Computer Science 2022-03-09 Jianju Gu , Longbiao Cheng , Dingding Yao , Junfeng Li , Yonghong Yan

The complete decomposition performed by blind source separation is computationally demanding and superfluous when only the speech of one specific target speaker is desired. In this paper, we propose a computationally efficient blind speech…

Sound · Computer Science 2020-08-04 Lele Liao , Zhaoyi Gu , Jing Lu

Multichannel convolutive blind speech source separation refers to the problem of separating different speech sources from the observed multichannel mixtures without much a priori information about the mixing system. Multichannel nonnegative…

Sound · Computer Science 2024-01-04 Jianyu Wang , Shanzheng Guan

Independent component analysis (ICA) is now a widely used solution for the analysis of multi-subject functional magnetic resonance imaging (fMRI) data. Independent vector analysis (IVA) generalizes ICA to multiple datasets, i.e., to…

Signal Processing · Electrical Eng. & Systems 2023-11-10 Trung Vu , Francisco Laport , Hanlu Yang , Vince D. Calhoun , Tulay Adali

Non-Gaussianity-based Independent Vector Extraction leads to the famous one-unit FastICA/FastIVA algorithm when the likelihood function is optimized using an approximate Newton-Raphson algorithm under the orthogonality constraint. In this…

Signal Processing · Electrical Eng. & Systems 2024-07-15 Zbyněk Koldovský , Jiří Málek , Jaroslav Čmejla , Stephen O'Regan

This paper introduces a new method for multi-channel time domain speech separation in reverberant environments. A fully-convolutional neural network structure has been used to directly separate speech from multiple microphone recordings,…

Audio and Speech Processing · Electrical Eng. & Systems 2020-11-12 Jisi Zhang , Catalin Zorila , Rama Doddipatla , Jon Barker

Recently, Constant Separating Vector (CSV) mixing model has been proposed for the Blind Source Extraction (BSE) of moving sources. In this paper, we experimentally verify the applicability of CSV in the blind extraction of a moving speaker…

Audio and Speech Processing · Electrical Eng. & Systems 2021-02-08 Jakub Janský , Zbyněk Koldovský , Jiří Málek , Tomáš Kounovský , Jaroslav Čmejla

In this paper, we present a statistical beamforming algorithm as a pre-processing step for robust automatic speech recognition (ASR). By modeling the target speech as a non-stationary Laplacian distribution, a mask-based statistical…

Audio and Speech Processing · Electrical Eng. & Systems 2024-01-08 Ui-Hyeop Shin , Hyung-Min Park

Recently, audio-visual speech enhancement has been tackled in the unsupervised settings based on variational auto-encoders (VAEs), where during training only clean data is used to train a generative model for speech, which at test time is…

Audio and Speech Processing · Electrical Eng. & Systems 2021-02-09 Mostafa Sadeghi , Xavier Alameda-Pineda

Self-supervised image denoising implies restoring the signal from a noisy image without access to the ground truth. State-of-the-art solutions for this task rely on predicting masked pixels with a fully-convolutional neural network. This…

Computer Vision and Pattern Recognition · Computer Science 2024-12-31 Mikhail Papkov , Pavel Chizhov , Leopold Parts

In this paper, we propose an invertible deep learning framework called INVVC for voice conversion. It is designed against the possible threats that inherently come along with voice conversion systems. Specifically, we develop an invertible…

Audio and Speech Processing · Electrical Eng. & Systems 2022-01-27 Zexin Cai , Ming Li

Single-channel audio separation aims to separate individual sources from a single-channel mixture. Most existing methods rely on supervised learning with synthetically generated paired data. However, obtaining high-quality paired data in…

Audio and Speech Processing · Electrical Eng. & Systems 2025-12-24 Runwu Shi , Chang Li , Jiang Wang , Rui Zhang , Nabeela Khan , Benjamin Yen , Takeshi Ashizawa , Kazuhiro Nakadai

We propose to learn surrogate functions of universal speech priors for determined blind speech separation. Deep speech priors are highly desirable due to their high modelling power, but are not compatible with state-of-the-art independent…

Audio and Speech Processing · Electrical Eng. & Systems 2020-11-12 Robin Scheibler , Masahito Togami