中文
相关论文

相关论文: FastFCA: A Joint Diagonalization Based Fast Algori…

200 篇论文

In this paper, we present an efficient neural network for end-to-end general purpose audio source separation. Specifically, the backbone structure of this convolutional network is the SUccessive DOwnsampling and Resampling of…

音频与语音处理 · 电气工程与系统科学 2021-05-14 Efthymios Tzinis , Zhepei Wang , Paris Smaragdis

We consider the acoustic source imaging problems using multiple frequency data. Using the data of one observation direction/point, we prove that some information (size and location) of the source support can be recovered. A non-iterative…

偏微分方程分析 · 数学 2017-12-08 Ala Alzaalig , Guanghui Hu , Xiaodong Liu , Jiguang Sun

This paper presents a real-time, energy-efficient embedded system implementing an array of Cascade of Asymmetric Resonators with Fast-Acting Compression (CARFAC) cochlea models for underwater sound analysis. Built on the AMD Kria KV260…

音频与语音处理 · 电气工程与系统科学 2025-08-12 Bram Bremer , Matthew Bigelow , Stuart Anstee , Gregory Cohen , Andre van Schaik , Ying Xu

In this paper, we propose an effective and robust method of spatial feature extraction for acoustic scene analysis utilizing partially synchronized and/or closely located distributed microphones. In the proposed method, a new cepstrum…

音频与语音处理 · 电气工程与系统科学 2020-04-22 Keisuke Imoto

Functional principal component analysis (FPCA) is a fundamental tool and has attracted increasing attention in recent decades, while existing methods are restricted to data with a single or finite number of random functions (much smaller…

统计方法学 · 统计学 2021-01-22 Xiaoyu Hu , Fang Yao

This paper introduces a multi-microphone method for extracting a desired speaker from a mixture involving multiple speakers and directional noise in a reverberant environment. In this work, we propose leveraging the instantaneous relative…

声音 · 计算机科学 2025-02-11 Aviad Eisenberg , Sharon Gannot , Shlomo E. Chazan

Functional data analysis (FDA) methods have computational and theoretical appeals for some high dimensional data, but lack the scalability to modern large sample datasets. To tackle the challenge, we develop randomized algorithms for two…

统计计算 · 统计学 2022-04-11 Shiyuan He , Xiaomeng Yan

This paper presents an unsupervised method that trains neural source separation by using only multichannel mixture signals. Conventional neural separation methods require a lot of supervised data to achieve excellent performance. Although…

声音 · 计算机科学 2019-08-30 Yoshiaki Bando , Yoko Sasaki , Kazuyoshi Yoshii

This paper proposes a new sparse array source enumeration algorithm for underdetermined scenarios with more sources than sensors. The proposed algorithm decomposes the wideband signals into multiple uncorrelated frequency bands, computes…

信号处理 · 电气工程与系统科学 2019-12-30 Yang Liu , John R. Buck

Classical methods for acoustic scene mapping require the estimation of time difference of arrival (TDOA) between microphones. Unfortunately, TDOA estimation is very sensitive to reverberation and additive noise. We introduce an unsupervised…

音频与语音处理 · 电气工程与系统科学 2024-03-14 Idan Cohen , Ofir Lindenbaum , Sharon Gannot

Background: Photoacoustic Microscopy (PAM) integrates optical and acoustic imaging, offering enhanced penetration depth for detecting optical-absorbing components in tissues. Nonetheless, challenges arise in scanning large areas with high…

医学物理 · 物理学 2023-12-15 Irem Loc , Mehmet Burcin Unlu

Sparse principal component analysis (PCA) is a well-established dimensionality reduction technique that is often used for unsupervised feature selection (UFS). However, determining the regularization parameters is rather challenging, and…

机器学习 · 计算机科学 2025-04-07 Long Chen , Xianchao Xiu

Soundscape studies typically attempt to capture the perception and understanding of sonic environments by surveying users. However, for long-term monitoring or assessing interventions, sound-signal-based approaches are required. To this…

音频与语音处理 · 电气工程与系统科学 2023-11-16 Yuanbo Hou , Qiaoqiao Ren , Huizhong Zhang , Andrew Mitchell , Francesco Aletta , Jian Kang , Dick Botteldooren

Query-based audio source extraction seeks to recover a target source from a mixture conditioned on a query. Existing approaches are largely confined to single-channel audio, leaving the spatial information in multi-channel recordings…

音频与语音处理 · 电气工程与系统科学 2025-10-16 Chenxin Yu , Hao Ma , Xu Li , Xiao-Lei Zhang , Mingjie Shao , Chi Zhang , Xuelong Li

Existing methods utilizing spatial information for sound source separation require prior knowledge of the direction of arrival (DOA) of the source or utilize estimated but imprecise localization results, which impairs the separation…

音频与语音处理 · 电气工程与系统科学 2025-04-08 Donghang Wu , Xihong Wu , Tianshu Qu

Context: in large-scale spatial surveys, the Point Spread Function (PSF) varies across the instrument field of view (FOV). Local measurements of the PSFs are given by the isolated stars images. Yet, these estimates may not be directly…

计算机视觉与模式识别 · 计算机科学 2016-11-04 F. M. Ngolè Mboula , J. -L. Starck , K. Okumura , J. Amiaux , P. Hudelot

The front-end factor analysis (FEFA), an extension of principal component analysis (PPCA) tailored to be used with Gaussian mixture models (GMMs), is currently the prevalent approach to extract compact utterance-level features (i-vectors)…

音频与语音处理 · 电气工程与系统科学 2018-05-04 Ville Vestman , Tomi Kinnunen

Fair principal component analysis (FPCA), a ubiquitous dimensionality reduction technique in signal processing and machine learning, aims to find a low-dimensional representation for a high-dimensional dataset in view of fairness. The FPCA…

最优化与控制 · 数学 2023-12-27 Meng Xu , Bo Jiang , Wenqiang Pu , Ya-Feng Liu , Anthony Man-Cho So

Direction of arrival (DoA) estimation is a common sensing problem in radar, sonar, audio, and wireless communication systems. It has gained renewed importance with the advent of the integrated sensing and communication paradigm. To fully…

This study explores the potential of using acoustic features of segmental speech sounds to detect deepfake audio. These features are highly interpretable because of their close relationship with human articulatory processes and are expected…

声音 · 计算机科学 2025-12-12 Tianle Yang , Chengzhe Sun , Siwei Lyu , Phil Rose