中文
相关论文

相关论文: Introducing SPAIN (SParse Audio INpainter)

200 篇论文

The state of the art in audio declipping has currently been achieved by SPADE (SParse Audio DEclipper) algorithm by Kiti\'c et al. Until now, the synthesis/sparse variant, S-SPADE, has been considered significantly slower than its…

音频与语音处理 · 电气工程与系统科学 2018-07-19 Pavel Záviška , Pavel Rajmic , Zdeněk Průša , Vítězslav Veselý

We deal with the problem of sparsity-based audio inpainting, i.e. filling in the missing segments of audio. A consequence of the approaches based on mathematical optimization is the insufficient amplitude of the signal in the filled gaps.…

音频与语音处理 · 电气工程与系统科学 2020-11-03 Ondřej Mokrý , Pavel Rajmic

Methods based on sparse representation have found great use in the recovery of audio signals degraded by clipping. The state of the art in declipping has been achieved by the SPADE algorithm by Kiti\'c et. al. (LVA/ICA2015). Our recent…

音频与语音处理 · 电气工程与系统科学 2020-01-17 Pavel Záviška , Pavel Rajmic , Ondřej Mokrý , Zdeněk Průša

Audio inpainting refers to signal processing techniques that aim at restoring missing or corrupted consecutive samples in audio signals. Prior works have shown that $\ell_1$- minimization with appropriate weighting is capable of solving…

声音 · 计算机科学 2022-02-16 Shristi Rajbamshi , Georg Tauböck , Peter Balazs , Nicki Holighaus

Spectroscopic photoacoustic (sPA) imaging uses multiple wavelengths to differentiate chromophores based on their unique optical absorption spectra. This technique has been widely applied in areas such as vascular mapping, tumor detection,…

计算机视觉与模式识别 · 计算机科学 2024-12-24 Fangzhou Lin , Shang Gao , Yichuan Tang , Xihan Ma , Ryo Murakami , Ziming Zhang , John D. Obayemi , Winston W. Soboyejo , Haichong K. Zhang

This technical report shows and discusses in detail how Sparse Audio Declipper (SPADE) algorithms are derived from the signal model using the ADMM approach. The analysis version (A-SPADE) of Kiti\'c et. al. (LVA/ICA 2015) is derived and…

最优化与控制 · 数学 2020-01-17 Pavel Záviška , Ondřej Mokrý , Pavel Rajmic

Recent advances in audio declipping have substantially improved the state of the art.% in certain saturation regimes. Yet, practitioners need guidelines to choose a method, and while existing benchmarks have been instrumental in advancing…

声音 · 计算机科学 2020-12-01 Clément Gaultier , Srđan Kitić , Rémi Gribonval , Nancy Bertin

Self-supervised speech representation models have succeeded in various tasks, but improving them for content-related problems using unlabeled data is challenging. We propose speaker-invariant clustering (Spin), a novel self-supervised…

计算与语言 · 计算机科学 2023-05-19 Heng-Jui Chang , Alexander H. Liu , James Glass

A novel variant of the Janssen method for audio inpainting is presented and compared to other popular audio inpainting methods based on autoregressive (AR) modeling. Both conceptual differences and practical implications are discussed. The…

音频与语音处理 · 电气工程与系统科学 2025-12-09 Ondřej Mokrý , Pavel Rajmic

Audio inpainting aims to reconstruct missing segments in corrupted recordings. Most of existing methods produce plausible reconstructions when the gap lengths are short, but struggle to reconstruct gaps larger than about 100 ms. This paper…

音频与语音处理 · 电气工程与系统科学 2025-01-13 Eloi Moliner , Vesa Välimäki

We introduce a class of Sparse, Physics-based, and partially Interpretable Neural Networks (SPINN) for solving ordinary and partial differential equations (PDEs). By reinterpreting a traditional meshless representation of solutions of PDEs…

机器学习 · 计算机科学 2021-08-13 Amuthan A. Ramabathiran , Prabhu Ramachandran

This paper introduces Robust Spin (R-Spin), a data-efficient domain-specific self-supervision method for speaker and noise-invariant speech representations by learning discrete acoustic units with speaker-invariant clustering (Spin). R-Spin…

计算与语言 · 计算机科学 2024-04-02 Heng-Jui Chang , James Glass

We develop the analysis (cosparse) variant of the popular audio declipping algorithm of Siedenburg et al. (2014). Furthermore, we extend both the old and the new variants by the possibility of weighting the time-frequency coefficients. We…

音频与语音处理 · 电气工程与系统科学 2023-03-08 Pavel Záviška , Pavel Rajmic

Sasaki et al. (2018) presented an efficient audio declipping algorithm, based on the properties of Hankel-structure matrices constructed from time-domain signal blocks. We adapt their approach to solving the audio inpainting problem, where…

音频与语音处理 · 电气工程与系统科学 2024-01-08 Pavel Záviška , Pavel Rajmic , Ondřej Mokrý

In this paper, we present a deep-learning-based framework for audio-visual speech inpainting, i.e., the task of restoring the missing parts of an acoustic speech signal from reliable audio context and uncorrupted visual information. Recent…

音频与语音处理 · 电气工程与系统科学 2021-02-04 Giovanni Morrone , Daniel Michelsanti , Zheng-Hua Tan , Jesper Jensen

Audio inpainting seeks to restore missing segments in degraded recordings. Previous diffusion-based methods exhibit impaired performance when the missing region is large. We introduce the first approach that applies discrete diffusion over…

声音 · 计算机科学 2026-02-18 Tali Dror , Iftach Shoham , Moshe Buchris , Oren Gal , Haim Permuter , Gilad Katz , Eliya Nachmani

Arbitrary text appearance poses a great challenge in scene text recognition tasks. Existing works mostly handle with the problem in consideration of the shape distortion, including perspective distortions, line curvature or other style…

计算机视觉与模式识别 · 计算机科学 2021-10-26 Chengwei Zhang , Yunlu Xu , Zhanzhan Cheng , Shiliang Pu , Yi Niu , Fei Wu , Futai Zou

Recently, great attention was intended toward overcomplete dictionaries and the sparse representations they can provide. In a wide variety of signal processing problems, sparsity serves a crucial property leading to high performance.…

Self-supervised representation learning approaches have grown in popularity due to the ability to train models on large amounts of unlabeled data and have demonstrated success in diverse fields such as natural language processing, computer…

机器学习 · 计算机科学 2023-02-06 John Harvill , Jarred Barber , Arun Nair , Ramin Pishehvar

Music performances, characterized by dense and continuous audio as well as seamless audio-visual integration, present unique challenges for multimodal scene understanding and reasoning. Recent Music Performance Audio-Visual Question…

声音 · 计算机科学 2025-06-03 Xingjian Diao , Tianzhen Yang , Chunhui Zhang , Weiyi Wu , Ming Cheng , Jiang Gui
‹ 上一页 1 2 3 10 下一页 ›