English
Related papers

Related papers: Differentiable Time-Frequency Scattering on GPU

200 papers

Most audio processing pipelines involve transformations that act on fixed-dimensional input representations of audio. For example, when using the Short Time Fourier Transform (STFT) the DFT size specifies a fixed dimension for the input…

Audio and Speech Processing · Electrical Eng. & Systems 2022-03-28 Krishna Subramani , Paris Smaragdis

Time-of-flight (ToF) devices have greatly propelled the advancement of various multi-modal perception applications. However, achieving accurate rendering of time-resolved information remains a challenge, particularly in scenes involving…

Graphics · Computer Science 2024-06-17 Qianyue He , Dongyu Du , Haitian Jiang , Xin Jin

In many multi-microphone algorithms, an estimate of the relative transfer functions (RTFs) of the desired speaker is required. Recently, a computationally efficient RTF vector estimation method was proposed for acoustic sensor networks,…

Audio and Speech Processing · Electrical Eng. & Systems 2023-04-07 Wiebke Middelberg , Simon Doclo

Spatio-Temporal Multivariate time series Forecast (STMF) uses the time series of $n$ spatially distributed variables in a period of recent past to forecast their values in a period of near future. It has important applications in…

Machine Learning · Computer Science 2025-10-29 Zibo Liu , Zhe Jiang , Zelin Xu , Tingsong Xiao , Yupu Zhang , Zhengkun Xiao , Haibo Wang , Shigang Chen

Multifractal Detrended Fluctuation Analysis (MFDFA) is a powerful and widely used technique for characterizing the scaling properties and long-range correlations of complex time series. However, its application often involves significant…

Transformer-based models are at the forefront in long time-series forecasting (LTSF). While in many cases, these models are able to achieve state of the art results, they suffer from a bias toward low-frequencies in the data and high…

Machine Learning · Computer Science 2026-05-13 Elisha Dayag , Nhat Thanh Van Tran , Jack Xin

Linear-time algorithms that are traditionally used to shuffle data on CPUs, such as the method of Fisher-Yates, are not well suited to implementation on GPUs due to inherent sequential dependencies, and existing parallel shuffling…

Distributed, Parallel, and Cluster Computing · Computer Science 2022-02-04 Rory Mitchell , Daniel Stokes , Eibe Frank , Geoffrey Holmes

We present diffSPH, a novel open-source differentiable Smoothed Particle Hydrodynamics (SPH) framework developed entirely in PyTorch with GPU acceleration. diffSPH is designed centrally around differentiation to facilitate optimization and…

Fluid Dynamics · Physics 2025-07-30 Rene Winchenbach , Nils Thuerey

Effective wireless communications are increasingly important in maintaining the successful closed-loop operation of mission-critical industrial Internet-of-Things (IIoT) applications. To meet the ever-increasing demands on better wireless…

Signal Processing · Electrical Eng. & Systems 2021-11-09 Kai Wu , J. Andrew Zhang , Xiaojing Huang , Y. Jay Guo

Estimating Head-Related Transfer Functions (HRTFs) of arbitrary source points is essential in immersive binaural audio rendering. Computing each individual's HRTFs is challenging, as traditional approaches require expensive time and…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-04 Jin Woo Lee , Sungho Lee , Kyogu Lee

Head-related transfer functions (HRTFs) are a set of functions describing the spatial filtering effect of the outer ear (i.e., torso, head, and pinnae) onto sound sources at different azimuth and elevation angles. They are widely used in…

Audio and Speech Processing · Electrical Eng. & Systems 2023-02-24 You Zhang , Yuxiang Wang , Zhiyao Duan

Fourier transform spectroscopy (FTS) has been widely used in a variety of fields in research, industry, and medicine due to its high signal-to-noise ratio, simultaneous acquisition of signals in a broad spectrum, and versatility for…

Time-frequency representation (TFR) allowing for mode reconstruction plays a significant role in interpreting and analyzing the nonstationary signal constituted of various modes. However, it is difficult for most previous methods to handle…

Signal Processing · Electrical Eng. & Systems 2021-09-01 Haijian Zhang , Guang Hua

The scattering transform is a non-linear signal representation method based on cascaded wavelet transform magnitudes. In this paper we introduce phase scattering, a novel approach where we use phase derivatives in a scattering procedure. We…

Sound · Computer Science 2024-07-09 Daniel Haider , Peter Balazs , Nicki Holighaus

Signal extraction from a single-channel mixture with additional undesired signals is most commonly performed using time-frequency (TF) masks. Typically, the mask is estimated with a deep neural network (DNN), and element-wise applied to the…

Sound · Computer Science 2019-12-10 Wolfgang Mack , Emanuël A. P. Habets

Many spatial filtering algorithms used for voice capture in, e.g., teleconferencing applications, can benefit from or even rely on knowledge of Relative Transfer Functions (RTFs). Accordingly, many RTF estimators have been proposed which,…

Audio and Speech Processing · Electrical Eng. & Systems 2021-10-06 Andreas Brendel , Johannes Zeitler , Walter Kellermann

Convolutional neural networks (CNN) are widely used for speech emotion recognition (SER). In such cases, the short time fourier transform (STFT) spectrogram is the most popular choice for representing speech, which is fed as input to the…

Audio and Speech Processing · Electrical Eng. & Systems 2019-08-09 Shruti Gupta , Md. Shah Fahad , Akshay Deepak

Time series forecasting is critical for decision-making across dynamic domains such as energy, finance, transportation, and cloud computing. However, real-world time series often exhibit non-stationarity, including temporal distribution…

Machine Learning · Computer Science 2025-12-01 Junkai Lu , Peng Chen , Chenjuan Guo , Yang Shu , Meng Wang , Bin Yang

We propose to combine cepstrum and nonlinear time-frequency (TF) analysis to study mutiple component oscillatory signals with time-varying frequency and amplitude and with time-varying non-sinusoidal oscillatory pattern. The concept of…

Data Analysis, Statistics and Probability · Physics 2016-11-23 Chen-Yun Lin , Li Su , Hau-tieng Wu

Most studies on speech enhancement generally don't consider the energy distribution of speech in time-frequency (T-F) representation, which is important for accurate prediction of mask or spectra. In this paper, we present a simple yet…

Sound · Computer Science 2022-03-10 Qiquan Zhang , Qi Song , Zhaoheng Ni , Aaron Nicolson , Haizhou Li