English
Related papers

Related papers: Gabor frames and deep scattering networks in audio…

200 papers

We propose harmonic-aligned frame mask for speech signals using non-stationary Gabor transform (NSGT). A frame mask operates on the transfer coefficients of a signal and consequently converts the signal into a counterpart signal. It depicts…

Sound · Computer Science 2019-04-24 Feng Huang , Peter Balazs

Audio zooming, a signal processing technique, enables selective focusing and enhancement of sound signals from a specified region, attenuating others. While traditional beamforming and neural beamforming techniques, centered on creating a…

Audio and Speech Processing · Electrical Eng. & Systems 2023-11-23 Meng Yu , Dong Yu

We perform a Gabor analysis for a large class of evolution equations with constant coefficients. We show that the corresponding propagators have a very sparse Gabor matrix, displaying off-diagonal exponential decay. The results apply to…

Functional Analysis · Mathematics 2014-07-04 Elena Cordero , Fabio Nicola , Luigi Rodino

An acoustic reverberator consisting of a network of delay lines connected via scattering junctions is proposed. All parameters of the reverberator are derived from physical properties of the enclosure it simulates. It allows for simulation…

Sound · Computer Science 2015-07-10 Enzo De Sena , Huseyin Hacihabiboglu , Zoran Cvetkovic , Julius O. Smith

The article describes a system for image recognition using deep convolutional neural networks. Modified network architecture is proposed that focuses on improving convergence and reducing training complexity. The filters in the first layer…

Computer Vision and Pattern Recognition · Computer Science 2019-05-01 Andrey Alekseev , Anatoly Bobe

In an attempt to address the need for skilled clinicians in heart sound interpretation, recent research efforts on automating cardiac auscultation have explored deep learning approaches. The majority of these approaches have been based on…

Sound · Computer Science 2025-10-08 Rami Zewail

Formants are the spectral maxima that result from acoustic resonances of the human vocal tract, and their accurate estimation is among the most fundamental speech processing problems. Recent work has been shown that those frequencies can…

Sound · Computer Science 2022-06-24 Yosi Shrem , Felix Kreuk , Joseph Keshet

It is increasingly considered that human speech perception and production both rely on articulatory representations. In this paper, we investigate whether this type of representation could improve the performances of a deep generative model…

Sound · Computer Science 2021-04-08 Marc-Antoine Georges , Laurent Girin , Jean-Luc Schwartz , Thomas Hueber

The problem of joint estimation of power spectrum and modulation from realizations of frequency modulated stationary wideband signals is considered. The study is motivated by some specific signal classes from which departures to…

Statistics Theory · Mathematics 2013-05-15 Harold Omer , Bruno Torrésani

Deep convolutional neural networks have led to breakthrough results in numerous practical machine learning tasks such as classification of images in the ImageNet data set, control-policy-learning to play Atari games or the board game Go,…

Information Theory · Computer Science 2017-10-25 Thomas Wiatowski , Helmut Bölcskei

The quantum mechanical harmonic oscillator Hamiltonian generates a one-parameter unitary group W(\theta) in L^2(R) which rotates the time-frequency plane. In particular, W(\pi/2) is the Fourier transform. When W(\theta) is applied to any…

Mathematical Physics · Physics 2009-11-07 Gerald Kaiser

Disentangling and recovering physical attributes, such as shape and material, from a few waveform examples is a challenging inverse problem in audio signal processing, with numerous applications in musical acoustics as well as structural…

Sound · Computer Science 2020-07-21 Han Han , Vincent Lostanlen

We experimentally and numerically study the potential of photoacoustic-guiding for light focusing through scattering samples via wavefront-shaping and iterative optimization. We experimentally demonstrate that the focusing efficiency on an…

In this paper, we propose an effective and robust method of spatial feature extraction for acoustic scene analysis utilizing partially synchronized and/or closely located distributed microphones. In the proposed method, a new cepstrum…

Audio and Speech Processing · Electrical Eng. & Systems 2020-04-22 Keisuke Imoto

Transformers have achieved promising results on a variety of tasks. However, the quadratic complexity in self-attention computation has limited the applications, especially in low-resource settings and mobile or edge devices. Existing works…

Sound · Computer Science 2024-01-09 Wentao Zhu

The audio denoising technique has captured widespread attention in the deep neural network field. Recently, the audio denoising problem has been converted into an image generation task, and deep learning-based approaches have been applied…

Sound · Computer Science 2024-06-14 Junhui Li , Pu Wang , Jialu Li , Youshan Zhang

We present a transformer-based speech-declipping model that effectively recovers clipped signals across a wide range of input signal-to-distortion ratios (SDRs). While recent time-domain deep neural network (DNN)-based declippers have…

Audio and Speech Processing · Electrical Eng. & Systems 2024-09-20 Younghoo Kwon , Jung-Woo Choi

We consider the inverse problem of determining the geometry of penetrable objects from scattering data generated by one incident wave at a fixed frequency. We first study an orthogonality sampling type method which is fast, simple to…

Numerical Analysis · Mathematics 2022-07-21 Thu Le , Dinh-Liem Nguyen , Vu Nguyen , Trung Truong

We present a machine learning model for the analysis of randomly generated discrete signals, modeled as the points of an inhomogeneous, compound Poisson point process. Like the wavelet scattering transform introduced by Mallat, our…

Statistics Theory · Mathematics 2021-10-12 Michael Perlmutter , Jieqian He , Matthew Hirn

Gaussian-based representations have enabled efficient physically-based volume rendering at a fraction of the memory cost of regular, discrete, voxel-based distributions. However, several remaining issues hamper their widespread use. One of…

Graphics · Computer Science 2026-02-06 Jorge Condor , Nicolai Hermann , Mehmet Ata Yurtsever , Piotr Didyk