English
Related papers

Related papers: SELEBI: Percussion-aware Time Stretching via Selec…

200 papers

We propose a new algorithm for time stretching music signals based on the theory of nonstationary Gabor frames (NSGFs). The algorithm extends the techniques of the classical phase vocoder (PV) by incorporating adaptive time-frequency (TF)…

Sound · Computer Science 2017-09-14 Emil Solsbæk Ottosen , Monika Dörfler

In this paper we propose a method for automatic local time adap- tation of the spectrogram of an audio signal, based on its decomposition within a Gabor multi-frame. The sparsity of the analyses within each individual frame is evaluated…

Sound · Computer Science 2011-09-29 M. Liuni , A. Röbel , M. Romito , X. Rodet

We propose a novel approach for time-scale modification of audio signals. Unlike traditional methods that rely on the framing technique or the short-time Fourier transform to preserve the frequency during temporal stretching, our neural…

Sound · Computer Science 2023-10-09 Ernie Chu , Ju-Ting Chen , Chia-Ping Chen

A deep neural network solution for time-scale modification (TSM) focused on large stretching factors is proposed, targeting environmental sounds. Traditional TSM artifacts such as transient smearing, loss of presence, and phasiness are…

Audio and Speech Processing · Electrical Eng. & Systems 2022-12-01 Leonardo Fierro , Alec Wright , Vesa Välimäki , Matti Hämäläinen

Most speech enhancement algorithms make use of the short-time Fourier transform (STFT), which is a simple and flexible time-frequency decomposition that estimates the short-time spectrum of a signal. However, the duration of short STFT…

Sound · Computer Science 2015-09-03 Scott Wisdom , Thomas Powers , Les Atlas , James Pitton

This work introduces sequential neural beamforming, which alternates between neural network based spectral separation and beamforming based spatial separation. Our neural networks for separation use an advanced convolutional architecture…

We propose a method for automatic local time-adaptation of the spectrogram of audio signals: it is based on the decomposition of a signal within a Gabor multi-frame through the STFT operator. The sparsity of the analysis in every individual…

Sound · Computer Science 2011-09-29 M. Liuni , A. Röbel , M. Romito , X. Rodet

Time-frequency analysis, such as the Gabor transform, plays an important role in many signal processing applications. The redundancy of such representations is often directly related to the computational load of any algorithm operating in…

Classical Analysis and ODEs · Mathematics 2015-05-13 Ewa Matusiak , Tomer Michaeli , Yonina C. Eldar

The temporal Talbot effect supports the generation of RF signals with high purity in optical domain. This allows for example a remote RF generation without the need for costly high-end electronic circuits and resource sharing. However, one…

Signal Processing · Electrical Eng. & Systems 2020-12-04 Niels Neumann , Zaid Al-Husseini , Dirk Plettemeier

This letter proposes a new time domain absorption approach designed to reduce masking components of speech signals under noisy-reverberant conditions. In this method, the non-stationarity of corrupted signal segments is used to detect…

Audio and Speech Processing · Electrical Eng. & Systems 2020-02-19 G. Zucatelli , R. Coelho

We consider sparseness properties of adaptive time-frequency representations obtained using nonstationary Gabor frames (NSGFs). NSGFs generalize classical Gabor frames by allowing for adaptivity in either time or frequency. It is known that…

Functional Analysis · Mathematics 2018-01-03 Emil Solsbæk Ottosen , Morten Nielsen

In this paper, a speech enhancement method based on noise compensation performed on short time magnitude as well phase spectra is presented. Unlike the conventional geometric approach (GA) to spectral subtraction (SS), here the noise…

Audio and Speech Processing · Electrical Eng. & Systems 2018-03-09 Md Tauhidul Islam , Udoy Saha , K. T. Shahid , Ahmed Bin Hussain , Celia Shahnaz

Most of the current speech data augmentation methods operate on either the raw waveform or the amplitude spectrum of speech. In this paper, we propose a novel speech data augmentation method called PhasePerturbation that operates…

Sound · Computer Science 2023-12-15 Chengxi Lei , Satwinder Singh , Feng Hou , Xiaoyun Jia , Ruili Wang

We present a time-frequency framework adapted to dispersive phase functions via a subdyadic geometry in phase space. On top of this geometry we construct stable Gabor frames with quantitative control of overlap, almost orthogonality, and…

Functional Analysis · Mathematics 2025-11-25 Vicente Vergara

The SpeakerBeam-FE (SBF) method is proposed for speaker extraction. It attempts to overcome the problem of unknown number of speakers in an audio recording during source separation. The mask approximation loss of SBF is sub-optimal, which…

Audio and Speech Processing · Electrical Eng. & Systems 2019-03-26 Chenglin Xu , Wei Rao , Eng Siong Chng , Haizhou Li

This letter introduces an innovative method to enhance the quality of audio time stretching by precisely decomposing a sound into sines, transients, and noise and by improving the processing of the latter component. While there are…

Audio and Speech Processing · Electrical Eng. & Systems 2023-12-25 Eloi Moliner , Leonardo Fierro , Alec Wright , Matti Hämäläinen , Vesa Välimäki

The Viterbi algorithm is a key operator for structured sequence inference in modern data systems, with applications in trajectory analysis, online recommendation, and speech recognition. As these workloads increasingly migrate to…

Distributed, Parallel, and Cluster Computing · Computer Science 2025-10-24 Ziheng Deng , Xue Liu , Jiantong Jiang , Yankai Li , Qingxu Deng , Xiaochun Yang

Surgical instrument segmentation is instrumental to minimally invasive surgeries and related applications. Most previous methods formulate this task as single-frame-based instance segmentation while ignoring the natural temporal and stereo…

Computer Vision and Pattern Recognition · Computer Science 2024-10-10 Qiyuan Wang , Shang Zhao , Zikang Xu , S Kevin Zhou

Seismic attributes calculated by conventional methods are susceptible to noise. Conventional filtering reduces the noise in the cost of losing the spectral bandwidth. The challenge of having a high-resolution and robust signal processing…

Geophysics · Physics 2020-12-02 M. Kazemnia Kakhki , W. J. Mansur , K. Aghazadeh

Cross-domain speech enhancement (SE) is often faced with severe challenges due to the scarcity of noise and background information in an unseen target domain, leading to a mismatch between training and test conditions. This study puts…

Sound · Computer Science 2024-09-04 Chien-Chun Wang , Li-Wei Chen , Hung-Shin Lee , Berlin Chen , Hsin-Min Wang
‹ Prev 1 2 3 10 Next ›