English
Related papers

Related papers: Audio Inpainting in Time-Frequency Domain with Pha…

200 papers

We present a method for audio denoising that combines processing done in both the time domain and the time-frequency domain. Given a noisy audio clip, the method trains a deep neural network to fit this signal. Since the fitting is only…

Sound · Computer Science 2020-06-11 Michael Michelashvili , Lior Wolf

Diffusion models have emerged as a powerful foundation model for visual generations. With an appropriate sampling process, it can effectively serve as a generative prior for solving general inverse problems. Current posterior sampling-based…

Computer Vision and Pattern Recognition · Computer Science 2025-07-01 Shijie Zhou , Huaisheng Zhu , Rohan Sharma , Jiayi Chen , Ruiyi Zhang , Kaiyi Ji , Changyou Chen

Audio zooming, a signal processing technique, enables selective focusing and enhancement of sound signals from a specified region, attenuating others. While traditional beamforming and neural beamforming techniques, centered on creating a…

Audio and Speech Processing · Electrical Eng. & Systems 2023-11-23 Meng Yu , Dong Yu

Phase imaging techniques extract the optical path-length information of a scene, whereas wavefront sensors provide the shape of an optical wavefront. Since these two applications have different technical requirements, they have developed…

Optics · Physics 2018-02-28 F. Soldevila , V. Durán , P. Clemente , J. Lancis , E. Tajahuerce

We consider the problem of reconstructing the sound field in a room using prior information of the boundary geometry, represented as a point cloud. In general, when no boundary information is available, an accurate sound field…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-17 David Sundström , Filip Elvander , Andreas Jakobsson

We study test-time domain adaptation for audio deepfake detection (ADD), addressing three challenges: (i) source-target domain gaps, (ii) limited target dataset size, and (iii) high computational costs. We propose an ADD method using prompt…

Sound · Computer Science 2024-10-15 Hideyuki Oiso , Yuto Matsunaga , Kazuya Kakizaki , Taiki Miyagawa

Automatic image colorization is inherently an ill-posed problem with uncertainty, which requires an accurate semantic understanding of scenes to estimate reasonable colors for grayscale images. Although recent interaction-based methods have…

Computer Vision and Pattern Recognition · Computer Science 2024-12-19 Pengcheng Zhao , Yanxiang Chen , Yang Zhao , Zhao Zhang

In acoustic signal processing, the target signals usually carry semantic information, which is encoded in a hierarchal structure of short and long-term contexts. However, the background noise distorts these structures in a nonuniform way.…

Audio and Speech Processing · Electrical Eng. & Systems 2022-01-26 Tassadaq Hussain , Wei-Chien Wang , Mandar Gogate , Kia Dashtipour , Yu Tsao , Xugang Lu , Adeel Ahsan , Amir Hussain

Fourier reconstruction algorithms significantly outperform conventional back-projection algorithms in terms of computation time. In photoacoustic imaging, these methods require interpolation in the Fourier space domain, which creates…

Numerical Analysis · Mathematics 2016-11-17 M. Haltmeier , O. Scherzer , G. Zangerl

Time-resolved optical filtering (TROF) measures the spectrogram or sonogram by a fast photodiode followed a tunable narrowband optical filter. For periodic signal and to match the sonogram, numerical TROF algorithm is used to find the…

Optics · Physics 2013-01-15 KeangPo Ho , Hsi-Cheng Wang , Hau-Kai Chen , Cheng-Chen Wu

Style Transfer with Inference-Time Optimisation (ST-ITO) is a recent approach for transferring the applied effects of a reference audio to an audio track. It optimises the effect parameters to minimise the distance between the style…

Spectral inference provides fast algorithms and provable optimality for latent topic analysis. But for real data these algorithms require additional ad-hoc heuristics, and even then often produce unusable results. We explain this poor…

Machine Learning · Computer Science 2016-11-02 Moontae Lee , David Bindel , David Mimno

In this paper we consider the inverse problem of identifying the initial data in a fractionally damped wave equation from time trace measurements on a surface, as relevant in photoacoustic or thermoacoustic tomography. We derive and analyze…

Numerical Analysis · Mathematics 2021-11-24 Barbara Kaltenbacher , Anna Schlintl

Gabor analysis is one of the most common instances of time-frequency signal analysis. Choosing a suitable window for the Gabor transform of a signal is often a challenge for practical applications, in particular in audio signal processing.…

Numerical Analysis · Computer Science 2013-11-13 Benjamin Ricaud , Guillaume Stempfel , Bruno Torrésani , Christoph Wiesmeyr , Hélène Lachambre , Darian Onchis

We introduce a new audio processing technique that increases the sampling rate of signals such as speech or music using deep convolutional neural networks. Our model is trained on pairs of low and high-quality audio examples; at test-time,…

Sound · Computer Science 2017-08-03 Volodymyr Kuleshov , S. Zayd Enam , Stefano Ermon

In this paper, we consider signals with intra-wave frequency modulation. To handle this kind of signals effectively, we generalize our data-driven time-frequency analysis by using a shape function to describe the intra-wave frequency…

Information Theory · Computer Science 2016-04-27 Thomas Y. Hou , Zuoqiang Shi

Wavelet domain inpainting refers to the process of recovering the missing coefficients during the image compression or transmission stage. Recently, an efficient algorithm framework which is called Bregmanized operator splitting (BOS) was…

Computer Vision and Pattern Recognition · Computer Science 2013-05-15 Dai-Qiang Chen , Li-Zhi Cheng

Homogeneous diffusion inpainting can reconstruct missing image areas with high quality from a sparse subset of known pixels, provided that their location as well as their gray or color values are well optimized. This property is exploited…

Image and Video Processing · Electrical Eng. & Systems 2024-08-13 Niklas Kämper , Vassillen Chizhov , Joachim Weickert

This paper concerns the problem of estimating multidimensional (MD) frequencies using prior knowledge of the signal spectral sparsity from partial time samples. In many applications, such as radar, wireless communications, and…

Information Theory · Computer Science 2019-04-26 Yinchuan Li , Xu Zhang , Zegang Ding , Xiaodong Wang

The objective of deep learning methods based on encoder-decoder architectures for music source separation is to approximate either ideal time-frequency masks or spectral representations of the target music source(s). The spectral…