Related papers: Audio Inpainting in Time-Frequency Domain with Pha…
Many audio signal processing methods are formulated in the time-frequency (T-F) domain which is obtained by the short-time Fourier transform (STFT). The properties of the STFT are fully characterized by window function, number of frequency…
In this paper, we propose a time-frequency analysis method to obtain instantaneous frequencies and the corresponding decomposition by solving an optimization problem. In this optimization problem, the basis to decompose the signal is not…
The process of reconstructing missing parts of speech audio from context is called speech in-painting. Human perception of speech is inherently multi-modal, involving both audio and visual (AV) cues. In this paper, we introduce and study a…
Video Multimethod Assessment Fusion (VMAF) [1], [2], [3] is a popular tool in the industry for measuring coded video quality. In this study, we propose an auditory-inspired frontend in existing VMAF for creating videos of reference and…
The standard approach for photoacoustic imaging with variable speed of sound is time reversal, which consists in solving a well-posed final-boundary value problem for the wave equation backwards in time. This paper investigates the…
Most existing sound field reconstruction methods target point-to-region reconstruction, interpolating the Acoustic Transfer Functions (ATFs) between a fixed-position sound source and a receiver region. The applicability of these methods is…
Training-free anomalous sound detection (ASD) based on pre-trained audio embedding models has recently garnered significant attention, as it enables the detection of anomalous sounds using only normal reference data while offering improved…
Accurate modeling of spatial acoustics is critical for immersive and intelligible audio in confined, resonant environments such as car cabins. Current tuning methods are manual, hardware-intensive, and static, failing to account for…
In the field of deepfake detection, previous studies focus on using reconstruction or mask and prediction methods to train pre-trained models, which are then transferred to fake audio detection training where the encoder is used to extract…
In this article, a time-domain calibration procedure is proposed for pulsed Terahertz Integrated Circuits (TIC) used in on-chip applications, where the conventional calibration methods are not applicable. The proposed post-detection method…
In this communication we will re-examine the widely studied technique of phase space projection. By imposing a time domain constraint (TDC) on the residual noise, we deduce a more general version of the optimal projector, which includes…
The aim of the study is to demonstrate that some methods are more relevant for implementing the Real-Time Nearfield Acoustic Holography than others. First by focusing on the forward propagation problem, different approaches are compared to…
In this article we study the inverse problem of thermoacoustic tomography (TAT) on a medium with attenuation represented by a time- convolution (or memory) term, and whose consideration is motivated by the modeling of ultrasound waves in…
Nonnegative Matrix Factorization (NMF) is a powerful tool for decomposing mixtures of audio signals in the Time-Frequency (TF) domain. In applications such as source separation, the phase recovery for each extracted component is a major…
The principle of stationary phase (PSP) is re-examined in the context of linear time-frequency (TF) decomposition using Gaussian, gammatone and gammachirp filters at uniform, logarithmic and cochlear spacings in frequency. This necessitates…
We propose a novel approach for time-scale modification of audio signals. Unlike traditional methods that rely on the framing technique or the short-time Fourier transform to preserve the frequency during temporal stretching, our neural…
Ring-array ultrasound computed tomography has recently achieved sufficient maturity for clinical applications like breast imaging. Image reconstruction is achieved with state of art iterative algorithms (full waveform inversion in the…
Almost all known image reconstruction algorithms for photoacoustic and thermoacoustic tomography assume that the acoustic waves leave the region of interest after a finite time. This assumption is reasonable if the reflections from the…
Phase recovery of modified spectrograms is a major issue in audio signal processing applications, such as source separation. This paper introduces a novel technique for estimating the phases of components in complex mixtures within onset…
The reconstruction mechanisms built by the human auditory system during sound reconstruction are still a matter of debate. The purpose of this study is to refine the auditory cortex model introduced in [9], and inspired by the geometrical…