中文
相关论文

相关论文: Multiple Hankel matrix rank minimization for audio…

200 篇论文

A novel variant of the Janssen method for audio inpainting is presented and compared to other popular audio inpainting methods based on autoregressive (AR) modeling. Both conceptual differences and practical implications are discussed. The…

音频与语音处理 · 电气工程与系统科学 2025-12-09 Ondřej Mokrý , Pavel Rajmic

The aim of this study is to implement a method to remove ambient noise in biomedical sounds captured in auscultation. We propose an incremental approach based on multichannel non-negative matrix partial co-factorization (NMPCF) for ambient…

In the ocean environment, reverberation is a common and strong interference which significantly degrades the performance of target bearing estimation. Meanwhile, sensor failure is inevitable in actual sonar deployment as the underwater…

信号处理 · 电气工程与系统科学 2020-07-02 Li-ya Xu , Bin Liao , Hao Zhang , Peng Xiao , Jian-jun Huang

Enhancing quality and removing noise during preprocessing is one of the most critical steps in image processing. X-ray images are created by photons colliding with atoms and the variation in scattered noise absorption. This noise leads to a…

计算机视觉与模式识别 · 计算机科学 2024-09-02 Masoud Shahraki Mohammadi , Seyed Javad Seyed Mahdavi Chabok

A sampling method by using scattering amplitude is proposed for shape and location reconstruction in inverse acoustic scattering problems. Only matrix multiplication is involved in the computation, thus the novel sampling method is very…

数值分析 · 数学 2017-09-13 Xiaodong Liu

We address the problem of detecting the number of complex exponentials and estimating their parameters from a noisy signal using the Matrix Pencil (MP) method. We introduce the MP modes and present their informative spectral structure. We…

信号处理 · 电气工程与系统科学 2025-09-30 Yehonatan-Itay Segman , Alon Amar , Ronen Talmon

In this paper, we study a spiked Wigner problem with an inhomogeneous noise profile. Our aim in this problem is to recover the signal passed through an inhomogeneous low-rank matrix channel. While the information-theoretic performances are…

机器学习 · 统计学 2023-02-15 Aleksandr Pak , Justin Ko , Florent Krzakala

In this work, we aim to analyze and optimize the EnCLAP framework, a state-of-the-art model in automated audio captioning. We investigate the impact of modifying the acoustic encoder components, explore pretraining with different dataset…

音频与语音处理 · 电气工程与系统科学 2024-09-04 Jaeyeon Kim , Minjeon Jeon , Jaeyoon Jung , Sang Hoon Woo , Jinjoo Lee

Image foreground extraction is a classical problem in image processing and vision, with a large range of applications. In this dissertation, we focus on the extraction of text and graphics in mixed-content images, and design novel…

计算机视觉与模式识别 · 计算机科学 2018-04-10 Shervin Minaee

Supervised multi-channel audio source separation requires extracting useful spectral, temporal, and spatial features from the mixed signals. The success of many existing systems is therefore largely dependent on the choice of features used…

声音 · 计算机科学 2018-03-05 Emad M. Grais , Dominic Ward , Mark D. Plumbley

In this paper, a Hankel matrix-based fully distributed algorithm is proposed to address a minimal-time deadbeat consensus prediction problem for discrete-time high-order multi-agent systems (MASs). Therein, each agent can predict the…

系统与控制 · 电气工程与系统科学 2023-04-14 Fu-Long Hu , Hai-Tao Zhang , Bowen Xu , Zhe Hu , Wei Ren

Self-supervised representation learning approaches have grown in popularity due to the ability to train models on large amounts of unlabeled data and have demonstrated success in diverse fields such as natural language processing, computer…

机器学习 · 计算机科学 2023-02-06 John Harvill , Jarred Barber , Arun Nair , Ramin Pishehvar

This paper tackles the scarcity of benchmarking data in disentangled auditory representation learning. We introduce SynTone, a synthetic dataset with explicit ground truth explanatory factors for evaluating disentanglement techniques.…

声音 · 计算机科学 2024-02-19 Yusuf Brima , Ulf Krumnack , Simone Pika , Gunther Heidemann

Deep learning, with its robust aotomatic feature extraction capabilities, has demonstrated significant success in audio signal processing. Typically, these methods rely on static, pre-collected large-scale datasets for training, performing…

声音 · 计算机科学 2024-12-19 Qisheng Xu , Yulin Sun , Yi Su , Qian Zhu , Xiaoyi Tan , Hongyu Wen , Zijian Gao , Kele Xu , Yong Dou , Dawei Feng

Recently, the application of low rank minimization to image denoising has shown remarkable denoising results which are equivalent or better than those of the existing state-of-the-art algorithms. However, due to iterative nature of low rank…

计算机视觉与模式识别 · 计算机科学 2015-04-27 Zahid Hussain Shamsi , Hyun Sook Oh , Dai-Gyoung Kim

The careful construction of audio representations has become a dominant feature in the design of approaches to many speech tasks. Increasingly, such approaches have emphasized "disentanglement", where a representation contains only parts of…

Speech in-painting is the task of regenerating missing audio contents using reliable context information. Despite various recent studies in multi-modal perception of audio in-painting, there is still a need for an effective infusion of…

声音 · 计算机科学 2024-06-04 Mahsa Kadkhodaei Elyaderani , Shahram Shirani

With the recent success of representation learning methods, which includes deep learning as a special case, there has been considerable interest in developing representation learning techniques that can incorporate known physical…

机器学习 · 计算机科学 2021-09-10 Harsha Vardhan Tetali , Joel B. Harley , Benjamin D. Haeffele

In this paper, we study a nonlinear spiked random matrix model where a nonlinear function is applied element-wise to a noise matrix perturbed by a rank-one signal. We establish a signal-plus-noise decomposition for this model and identify…

统计理论 · 数学 2024-05-29 Behrad Moniri , Hamed Hassani

The high-intensity, repetitive noise associated with functional magnetic resonance imaging hinders on-line monitoring of subjects' speech and/or recording speech signals suitable for off-line analysis. The proposed algorithm enhances the…

声音 · 计算机科学 2012-07-26 Satrajit S. Ghosh