English
Related papers

Related papers: Gabor frames and deep scattering networks in audio…

200 papers

Acoustic scattering is strongly influenced by boundary geometry of objects over which sound scatters. The present work proposes a method to infer object geometry from scattering features by training convolutional neural networks. The…

Sound · Computer Science 2021-02-12 Ziqi Fan , Vibhav Vineet , Chenshen Lu , T. W. Wu , Kyla McMullen

In this paper we propose a method for automatic local time adap- tation of the spectrogram of an audio signal, based on its decomposition within a Gabor multi-frame. The sparsity of the analyses within each individual frame is evaluated…

Sound · Computer Science 2011-09-29 M. Liuni , A. Röbel , M. Romito , X. Rodet

Generative networks have made it possible to generate meaningful signals such as images and texts from simple noise. Recently, generative methods based on GAN and VAE were developed for graphs and graph signals. However, the mathematical…

Machine Learning · Computer Science 2019-10-18 Dongmian Zou , Gilad Lerman

Speaker-aware source separation methods are promising workarounds for major difficulties such as arbitrary source permutation and unknown number of sources. However, it remains challenging to achieve satisfying performance provided a very…

Sound · Computer Science 2018-07-25 Jun Wang , Jie Chen , Dan Su , Lianwu Chen , Meng Yu , Yanmin Qian , Dong Yu

We consider Hamiltonian deformations of Gabor systems, where the window evolves according to the action of a Schr\"odinger propagator and the phase-space nodes evolve according to the corresponding Hamiltonian flow. We prove the stability…

Mathematical Physics · Physics 2016-11-29 Maurice A. de Gosson , Karlheinz Gröchenig , José Luis Romero

Optoacoustic image formation is conventionally based upon ultrasound time-of-flight readings from multiple detection positions. Herein, we exploit acoustic scattering to physically encode the position of optical absorbers in the acquired…

Biological Physics · Physics 2019-10-30 Xose Luis Dean-Ben , Ali Ozbek , Hernan Lopez-Schier , Daniel Razansky

This paper proposes a novel bidirectional neural vocoder, named BiVocoder, capable both of feature extraction and reverse waveform generation within the short-time Fourier transform (STFT) domain. For feature extraction, the BiVocoder takes…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-05 Hui-Peng Du , Ye-Xin Lu , Yang Ai , Zhen-Hua Ling

We propose the product-of-filters (PoF) model, a generative model that decomposes audio spectra as sparse linear combinations of "filters" in the log-spectral domain. PoF makes similar assumptions to those used in the classic homomorphic…

Machine Learning · Statistics 2014-11-27 Dawen Liang , Matthew D. Hoffman , Gautham J. Mysore

The objective of deep learning methods based on encoder-decoder architectures for music source separation is to approximate either ideal time-frequency masks or spectral representations of the target music source(s). The spectral…

In this report we describe an ongoing line of research for solving single-channel source separation problems. Many monaural signal decomposition techniques proposed in the literature operate on a feature space consisting of a time-frequency…

Sound · Computer Science 2015-04-29 Pablo Sprechmann , Joan Bruna , Yann LeCun

We consider the problem of reconstructing a signal $f$ from its spectrogram, i.e., the magnitudes $|V_\varphi f|$ of its Gabor transform $$V_\varphi f (x,y):=\int_{\mathbb{R}}f(t)e^{-\pi (t-x)^2}e^{-2\pi \i y t}dt, \quad x,y\in…

Functional Analysis · Mathematics 2017-06-15 Philipp Grohs , Martin Rathmair

This paper introduces a Deep Scattering network that utilizes Dual-Tree complex wavelets to extract translation invariant representations from an input signal. The computationally efficient Dual-Tree wavelets decompose the input signal into…

Computer Vision and Pattern Recognition · Computer Science 2017-02-14 Amarjot Singh , Nick Kingsbury

This paper presents a novel method for extracting the vocal track from a musical mixture. The musical mixture consists of a singing voice and a backing track which may comprise of various instruments. We use a convolutional network with…

Sound · Computer Science 2020-02-13 Pritish Chandna , Merlijn Blaauw , Jordi Bonada , Emilia Gomez

To satisfy the requirements of the end-to-end fault diagnosis of gears, an integrated intelligent method of fault diagnosis for gears using acceleration signals was proposed, which was based on Gabor-based Adaptive Short-Time Fourier…

Machine Learning · Computer Science 2025-04-01 Bowei Qiao , Hongwei Wang

Deep speaker embeddings have been demonstrated to outperform their generative counterparts, i-vectors, in recent speaker verification evaluations. To combine the benefits of high performance and generative interpretation, we investigate the…

Audio and Speech Processing · Electrical Eng. & Systems 2020-04-21 Ville Vestman , Kong Aik Lee , Tomi H. Kinnunen

We study spanning properties of a family of functions translated along simple model sets. We characterize tight frame and dual frame generators for such irregular translates and we apply the results to Gabor systems. We use the connection…

Functional Analysis · Mathematics 2019-02-21 Ewa Matusiak

Audio signal processing frequently requires time-frequency representations and in many applications, a non-linear spacing of frequency-bands is preferable. This paper introduces a framework for efficient implementation of invertible signal…

Functional Analysis · Mathematics 2013-05-17 Nicki Holighaus , Monika Dörfler , Gino Angelo Velasco , Thomas Grill

A class of robust estimators of scatter applied to information-plus-impulsive noise samples is studied, where the sample information matrix is assumed of low rank; this generalizes the study of (Couillet et al., 2013b) to spiked random…

Probability · Mathematics 2014-05-01 Romain Couillet

Lifelong audio feature extraction involves learning new sound classes incrementally, which is essential for adapting to new data distributions over time. However, optimizing the model only on new data can lead to catastrophic forgetting of…

Audio and Speech Processing · Electrical Eng. & Systems 2024-02-08 Xilin Jiang , Yinghao Aaron Li , Nima Mesgarani

The purpose of this note is to present a proof of the existence of Gabor frames in general linear position in all finite dimensions. The tools developed in this note are also helpful towards an explicit construction of such a frame, which…

Rings and Algebras · Mathematics 2020-05-04 Romanos-Diogenes Malikiosis
‹ Prev 1 3 4 5 6 7 10 Next ›