English
Related papers

Related papers: Audio Inpainting in Time-Frequency Domain with Pha…

200 papers

In recent years inpainting-based compression methods have been shown to be a viable alternative to classical codecs such as JPEG and JPEG2000. Unlike transform-based codecs, which store coefficients in the transform domain, inpainting-based…

Image and Video Processing · Electrical Eng. & Systems 2024-01-15 Niklas Kämper , Vassillen Chizhov , Joachim Weickert

This paper presents a paraxial modeling approach for vibro-acoustography, a high-frequency ultrasound imaging technique that makes use of the excited low-frequency field to achieve a higher resolution while avoiding speckles. We start from…

Numerical Analysis · Mathematics 2023-10-06 Teresa Rauscher

Image inpainting is the task of filling masked or unknown regions of an image with visually realistic contents, which has been remarkably improved by Deep Neural Networks (DNNs) recently. Essentially, as an inverse problem, the inpainting…

Computer Vision and Pattern Recognition · Computer Science 2022-06-15 Chenjie Cao , Chengrong Wang , Yuntao Zhang , Yanwei Fu

We study inertial versions of primal-dual proximal splitting, also known as the Chambolle--Pock method. Our starting point is the preconditioned proximal point formulation of this method. By adding correctors corresponding to the…

Optimization and Control · Mathematics 2020-05-15 Tuomo Valkonen

We present a new data-driven video inpainting method for recovering missing regions of video frames. A novel deep learning architecture is proposed which contains two sub-networks: a temporal structure inference network and a spatial detail…

Computer Vision and Pattern Recognition · Computer Science 2018-12-04 Chuan Wang , Haibin Huang , Xiaoguang Han , Jue Wang

Audio-visual video segmentation (AVVS) aims to generate pixel-level maps of sound-producing objects that accurately align with the corresponding audio. However, existing methods often face temporal misalignment, where audio cues and…

Computer Vision and Pattern Recognition · Computer Science 2024-12-12 Kexin Li , Zongxin Yang , Yi Yang , Jun Xiao

In neural-based audio feature extraction, ensuring that representations capture disentangled information is crucial for model interpretability. However, existing disentanglement methods often rely on assumptions that are highly dependent on…

Sound · Computer Science 2025-10-07 Benoit Ginies , Xiaoyu Bie , Olivier Fercoq , Gaël Richard

This paper presents a simple but effective method that uses multi-resolution feature maps with convolutional neural networks (CNNs) for anti-spoofing in automatic speaker verification (ASV). The central idea is to alleviate the problem that…

Audio and Speech Processing · Electrical Eng. & Systems 2020-08-21 Qiongqiong Wang , Kong Aik Lee , Takafumi Koshinaka

Advances in AI technology have made voice cloning increasingly accessible, leading to a rise in fraud involving AI-generated audio forgeries. This highlights the need to covertly embed information and verify the authenticity and integrity…

Cryptography and Security · Computer Science 2024-08-28 Guang Yang

A novel sparsity-based algorithm for audio inpainting is proposed. It is an adaptation of the SPADE algorithm by Kiti\'c et al., originally developed for audio declipping, to the task of audio inpainting. The new SPAIN (SParse Audio…

Sound · Computer Science 2020-01-17 Ondřej Mokrý , Pavel Záviška , Pavel Rajmic , Vítězslav Veselý

Recently, a novel measurement setup has been introduced to photoacoustic tomography, that collects data in the form of projections of the full 3D acoustic pressure distribution at a certain time instant. Existing imaging algorithms for this…

Numerical Analysis · Mathematics 2018-08-03 Gerhard Zangerl , Markus Haltmeier , Linh V. Nguyen , Robert Nuster

In this paper we are particularly interested in the image inpainting problem using directional complex tight wavelet frames. Under the assumption that frame coefficients of images are sparse, several iterative thresholding algorithms for…

Information Theory · Computer Science 2014-07-14 Yi Shen , Bin Han , Elena Braverman

We investigate the inverse source problem for the wave equation, arising in photo- and thermoacoustic tomography. There exist quite a few theoretically exact inversion formulas explicitly expressing solution of this problem in terms of the…

Analysis of PDEs · Mathematics 2018-08-01 Ngoc Do , Leonid Kunyansky

Sparse time-frequency (T-F) representations have been an important research topic for more than several decades. Among them, optimization-based methods (in particular, extensions of basis pursuit) allow us to design the representations…

Signal Processing · Electrical Eng. & Systems 2023-08-04 Keidai Arai , Koki Yamada , Kohei Yatabe

Generally, wave field reconstructions obtained by phase-retrieval algorithms are noisy, blurred and corrupted by various artifacts such as irregular waves, spots, etc. These disturbances, arising due to many factors such as non-idealities…

Optics · Physics 2012-07-24 Artem Migukin , Mostafa Agour , Vladimir Katkovnik

In this paper, we propose a solution for improving the quality of temporal sound localization. We employ a multimodal fusion approach to combine visual and audio features. High-quality visual features are extracted using a state-of-the-art…

Sound · Computer Science 2024-07-03 Yurui Huang , Yang Yang , Shou Chen , Xiangyu Wu , Qingguo Chen , Jianfeng Lu

We propose a new algorithm for time stretching music signals based on the theory of nonstationary Gabor frames (NSGFs). The algorithm extends the techniques of the classical phase vocoder (PV) by incorporating adaptive time-frequency (TF)…

Sound · Computer Science 2017-09-14 Emil Solsbæk Ottosen , Monika Dörfler

Recent developments in acoustic signal processing have seen the integration of deep learning methodologies, alongside the continued prominence of classical wave expansion-based approaches, particularly in sound field reconstruction.…

Audio and Speech Processing · Electrical Eng. & Systems 2024-04-24 Marco Olivieri , Xenofon Karakonstantis , Mirco Pezzoli , Fabio Antonacci , Augusto Sarti , Efren Fernandez-Grande

Deep convolutional neural networks are known to specialize in distilling compact and robust prior from a large amount of data. We are interested in applying deep networks in the absence of training dataset. In this paper, we introduce deep…

Sound · Computer Science 2019-12-24 Yapeng Tian , Chenliang Xu , Dingzeyu Li

We present an alternative temporal approach for convolution, providing a new algorithm, called the taches-algorithm. Based on interferences between the successive delayed and amplified output signals associated respectively with the…

Signal Processing · Electrical Eng. & Systems 2024-01-04 Laurent Millot , Gérard Pelé