中文
相关论文

相关论文: Introducing SPAIN (SParse Audio INpainter)

200 篇论文

We propose SpeechPainter, a model for filling in gaps of up to one second in speech samples by leveraging an auxiliary textual input. We demonstrate that the model performs speech inpainting with the appropriate content, while maintaining…

声音 · 计算机科学 2022-03-31 Zalán Borsos , Matt Sharifi , Marco Tagliasacchi

Noise reduction techniques based on deep learning have demonstrated impressive performance in enhancing the overall quality of recorded speech. While these approaches are highly performant, their application in audio engineering can be…

声音 · 计算机科学 2023-10-18 Christian J. Steinmetz , Thomas Walther , Joshua D. Reiss

Audio splicing is one of the most common manipulation techniques in the area of audio forensics. In this paper, the magnitudes of acoustic channel impulse response and ambient noise are proposed as the environmental signature. Specifically,…

密码学与安全 · 计算机科学 2014-11-27 Hong Zhao , Yifan Chen , Rui Wang , Hafiz Malik

Here, we propose a new reconstruction method of smooth time-series signals. A key concept of this study is not considering the model in signal space, but in delay-embedded space. In other words, we indirectly represent a time-series signal…

音频与语音处理 · 电气工程与系统科学 2022-03-21 Tatsuya Yokota

We propose an algorithm for the blind separation of single-channel audio signals. It is based on a parametric model that describes the spectral properties of the sounds of musical instruments independently of pitch. We develop a novel…

音频与语音处理 · 电气工程与系统科学 2021-02-03 Sören Schulze , Emily J. King

We present an algorithm for resampling a function from its values on a non-Cartesian grid onto a Cartesian grid. This problem arises in many applications such as MRI, CT, radio astronomy and geophysics. Our algorithm, termed SParse Uniform…

信息论 · 计算机科学 2016-03-17 Amir Kiperwas , Daniel Rosenfeld , Yonina C. Eldar

This paper presents a novel approach called the boundary integrated neural networks (BINNs) for analyzing acoustic radiation and scattering. The method introduces fundamental solutions of the time-harmonic wave equation to encode the…

数值分析 · 数学 2023-07-21 Wenzhen Qu , Yan Gu , Shengdong Zhao , Fajie wang

Physics-Informed Neural Networks (PINNs) are a novel computational approach for solving partial differential equations (PDEs) with noisy and sparse initial and boundary data. Although, efficient quantification of epistemic and aleatoric…

Atomic resolution STEM images often suffer from noise due to low electron doses and instrument imperfections, hence it is challenging to obtain critical structural details required for material analysis. To address the problem, we propose a…

材料科学 · 物理学 2024-12-18 Z. Awan , J. Shabeer , U. Saleem , S. Mehmood , T. Qadeer

We consider audio decoding as an inverse problem and solve it through diffusion posterior sampling. Explicit conditioning functions are developed for input signal measurements provided by an example of a transform domain perceptual audio…

音频与语音处理 · 电气工程与系统科学 2024-09-13 Pedro J. Villasana T. , Lars Villemoes , Janusz Klejsa , Per Hedelin

The rapid advancement of spoofing algorithms necessitates the development of robust detection methods capable of accurately identifying emerging fake audio. Traditional approaches, such as finetuning on new datasets containing these novel…

声音 · 计算机科学 2023-06-16 Xiaohui Zhang , Jiangyan Yi , Jianhua Tao , Chenlong Wang , Le Xu , Ruibo Fu

In this study, we introduce a method for estimating sound fields in reverberant environments using a conditional invertible neural network (CINN). Sound field reconstruction can be hindered by experimental errors, limited spatial data,…

音频与语音处理 · 电气工程与系统科学 2024-04-11 Xenofon Karakonstantis , Efren Fernandez-Grande , Peter Gerstoft

Though achieving excellent performance in some cases, current unsupervised learning methods for single image denoising usually have constraints in applications. In this paper, we propose a new approach which is more general and applicable…

计算机视觉与模式识别 · 计算机科学 2023-04-18 Yutong Xie , Mingze Yuan , Bin Dong , Quanzheng Li

Solving Singularly Perturbed Differential Equations (SPDEs) presents challenges due to the rapid change of their solutions at the boundary layer. In this manuscript, We propose Asymptotic Physics-Informed Neural Networks (ASPINN), a…

机器学习 · 计算机科学 2024-09-23 Sen Wang , Peizhi Zhao , Tao Song

Recent advancements in adapting vision-language pre-training models like CLIP for person re-identification (ReID) tasks often rely on complex adapter design or modality-specific tuning while neglecting cross-modal interaction, leading to…

计算机视觉与模式识别 · 计算机科学 2025-07-02 Yunfei Xie , Yuxuan Cheng , Juncheng Wu , Haoyu Zhang , Yuyin Zhou , Shoudong Han

Confidence alone is often misleading in hyperspectral image classification, as models tend to mistake high predictive scores for correctness while lacking awareness of uncertainty. This leads to confirmation bias, especially under sparse…

计算机视觉与模式识别 · 计算机科学 2025-11-14 Muzhou Yang , Wuzhou Quan , Mingqiang Wei

Distributed representation of words has improved the performance for many natural language tasks. In many methods, however, only one meaning is considered for one label of a word, and multiple meanings of polysemous words depending on the…

计算与语言 · 计算机科学 2020-06-01 Yusuke Takimoto , Yosuke Fukuchi , Shoya Matsumori , Michita Imai

We propose DoPAMINE, a new neural network based multiplicative noise despeckling algorithm. Our algorithm is inspired by Neural AIDE (N-AIDE), which is a recently proposed neural adaptive image denoiser. While the original N-AIDE was…

图像与视频处理 · 电气工程与系统科学 2019-02-08 Sunghwan Joo , Sungmin Cha , Taesup Moon

Many signal processing algorithms break the target signal into overlapping segments (also called windows, or patches), process them separately, and then stitch them back into place to produce a unified output. At the overlaps, the final…

信号处理 · 电气工程与系统科学 2021-03-15 Ignacio Francisco Ramírez Paulino

An effective way to increase the noise robustness of automatic speech recognition is to label noisy speech features as either reliable or unreliable (missing) prior to decoding, and to replace the missing ones by clean speech estimates. We…

声音 · 计算机科学 2009-01-19 J. F. Gemmeke , B. Cranen