Related papers: Audio Compression using Periodic Gabor with Biorth…
A fast, efficient and scalable algorithm is proposed, in this paper, for re-encoding of perceptually quantized wavelet-packet transform (WPT) coefficients of audio and high quality speech and is called "adaptive variable degree-k…
Current numerical abstract interpretation relies on fixed, hand-crafted, instruction-specific transformers tailored to each domain, causing three key limitations: transformers cannot be reused across domains; precise compositional reasoning…
We propose harmonic-aligned frame mask for speech signals using non-stationary Gabor transform (NSGT). A frame mask operates on the transfer coefficients of a signal and consequently converts the signal into a counterpart signal. It depicts…
This study introduces a novel application of a Generative Pre-trained Transformer (GPT) model tailored for photoplethysmography (PPG) signals, serving as a foundation model for various downstream tasks. Adapting the standard GPT…
We report the spectral features of a phase-shifted parity and time ($\mathcal{PT}$)-symmetric fiber Bragg grating (PPTFBG) and demonstrate its functionality as a demultiplexer in the unbroken $\mathcal{PT}$-symmetric regime. The length of…
The performance of Zak-OTFS modulation is critically dependent on the choice of the delay-Doppler (DD) domain pulse shaping filter. The design of pulses for $L^2(\mathbb{R})$ is constrained by the Balian-Low Theorem, which imposes an…
Intense, single-cycle terahertz (THz) pulses offer a promising approach for understanding and controlling the properties of a material on an ultrafast time scale. In particular, resonantly exciting phonons leads to a better understanding of…
Most of the current speech data augmentation methods operate on either the raw waveform or the amplitude spectrum of speech. In this paper, we propose a novel speech data augmentation method called PhasePerturbation that operates…
In this paper, we address the speech denoising problem, where Gaussian, pink and blue additive noises are to be removed from a given speech signal. Our approach is based on a redundant, analysis-sparse representation of the original speech…
Orthogonal moment-based image representations are fundamental in computer vision, but classical methods suffer from high computational complexity and numerical instability at large orders. Zernike and pseudo-Zernike moments, for instance,…
This work presents an analysis of state-of-the-art learning-based image compression techniques. We compare 8 models available in the Tensorflow Compression package in terms of visual quality metrics and processing time, using the KODAK data…
This work aims at presenting a Discontinuous Galerkin (DG) formulation employing a spectral basis for two important models employed in cardiac electrophysiology, namely the monodomain and bidomain models. The use of DG methods is motivated…
A common approach to digital system design involves transforming a continuous-time (s-domain) transfer function into the discrete-time (z-domain) using methods such as Euler or Tustin. These transformations are shown to be specific cases of…
Generative models have demonstrated strong performance in conditional settings and can be viewed as a form of data compression, where the condition serves as a compact representation. However, their limited controllability and…
Head-based signals such as EEG, EMG, EOG, and ECG collected by wearable systems will play a pivotal role in clinical diagnosis, monitoring, and treatment of important brain disorder diseases. However, the real-time transmission of the…
The high frequency performance of strong piezoelectric materials like PZT remains relatively less explored due to the assumption of large dielectric/ferroelectric losses at GHz frequencies. Recently, the advent of magnetoelectric technology…
While existing speech audio codecs designed for compression exploit limited forms of temporal redundancy and allow for multi-scale representations, they tend to represent all features of audio in the same way. In contrast, generative voice…
The fluctuation exchange (FLEX) approximation is applied to study the Holstein-Hubbard model. Due to the retarded nature of the phonon-mediated electron-electron interaction, neither fast Fourier transform (FFT) nor previously developed NRG…
Using a linear combination of atomic orbitals approach, we report a systematic comparison of various Density Functional Theory (DFT) and hybrid exchange-correlation functionals for the prediction of the electronic and structural properties…
Benzoic acid (BA) is a model system for studying proton transfer (PT) reactions. The properties of solid BA subject to high pressure (exceeding 1 kbar = 0.1 GPa) are of particular interest due to the possibility of compression-tuning of the…