中文
相关论文

相关论文: Audio signal interpolation using optimal transport…

200 篇论文

Acoustic-to-articulatory inversion (AAI) is to convert audio into articulator movements, such as ultrasound tongue imaging (UTI) data. An issue of existing AAI methods is only using the personalized acoustic information to derive the…

声音 · 计算机科学 2024-03-13 Yudong Yang , Rongfeng Su , Xiaokang Liu , Nan Yan , Lan Wang

Spatial sound field interpolation relies on suitable models to both conform to available measurements and predict the sound field in the domain of interest. A suitable model can be difficult to determine when the spatial domain of interest…

音频与语音处理 · 电气工程与系统科学 2022-11-30 Manuel Hahmann , Efren Fernandez-Grande

Separating sources is a common challenge in applications such as speech enhancement and telecommunications, where distinguishing between overlapping sounds helps reduce interference and improve signal quality. Additionally, in multichannel…

音频与语音处理 · 电气工程与系统科学 2025-03-25 Linda Fabiani , Sebastian J. Schlecht , Isabel Haasler , Filip Elvander

This study describes a binaural machine hearing system that is capable of performing auditory stream segregation in scenarios where multiple sound sources are present. The process of stream segregation refers to the capability of human…

声音 · 计算机科学 2016-06-27 Christopher Schymura , Thomas Walther , Dorothea Kolossa

Optimal transport as a loss for machine learning optimization problems has recently gained a lot of attention. Building upon recent advances in computational optimal transport, we develop an optimal transport non-negative matrix…

声音 · 计算机科学 2018-09-26 Antoine Rolet , Vivien Seguy , Mathieu Blondel , Hiroshi Sawada

We consider the problem of separating speech sources captured by multiple spatially separated devices, each of which has multiple microphones and samples its signals at a slightly different rate. Most asynchronous array processing methods…

音频与语音处理 · 电气工程与系统科学 2019-12-12 Ryan M. Corey , Andrew C. Singer

In this paper, we consider a scenario of covert communication aided by multiple friendly interference nodes. The objective is to conceal the legitimate communication link under the surveillance of a warden. The main content is as follows:…

信号处理 · 电气工程与系统科学 2024-05-10 Xuyang Zhao. Wei Guo , Yongchao Wang

Optimal transport is a notoriously difficult problem to solve numerically, with current approaches often remaining intractable for very large scale applications such as those encountered in machine learning. Wasserstein barycenters -- the…

机器学习 · 计算机科学 2021-02-25 Julien Lacombe , Julie Digne , Nicolas Courty , Nicolas Bonneel

Sound in indoor spaces forms a complex wavefield due to multiple scattering encountered by the sound. Indoor acoustic communication involving multiple sources and receivers thus inevitably suffers from cross-talks. Here, we demonstrate the…

声音 · 计算机科学 2024-02-13 Hongkuan Zhang , Qiyuan Wang , Mathias Fink , Guancong Ma

In lossless acoustic systems, mode transitions are always time-reversible, consistent with Lorentz reciprocity, giving rise to symmetric sound manipulation in space-time. To overcome this fundamental limitation and break space-time…

Audio source separation is usually achieved by estimating the short-time Fourier transform (STFT) magnitude of each source, and then applying a spectrogram inversion algorithm to retrieve time-domain signals. In particular, the multiple…

声音 · 计算机科学 2020-04-22 Paul Magron , Tuomas Virtanen

Locating where transient signals travel between a source and receiver requires a final step that is needed after using a theory of diffraction such as the integral theorem of Helmholtz and Kirchhoff. Introduced here, the final step accounts…

经典物理 · 物理学 2014-11-18 John L. Spiesberger

The main goal of this paper is to estimate the regional acoustic and geoacoustic shallow-water environment from data collected by a vertical hydrophone array and transmitted by distant time-harmonic point sources. We aim at estimating the…

偏微分方程分析 · 数学 2019-09-04 Laure Dumaz , Josselin Garnier , Guilhem Lepoultier

We extend frequency-domain blind source separation based on independent vector analysis to the case where there are more microphones than sources. The signal is modelled as non-Gaussian sources in a Gaussian background. The proposed…

声音 · 计算机科学 2019-08-08 Robin Scheibler , Nobutaka Ono

The image source method (ISM) is often used to simulate room acoustics due to its ease of use and computational efficiency. The standard ISM is limited to simulations of room impulse responses between point sources and omnidirectional…

音频与语音处理 · 电气工程与系统科学 2023-09-08 Zeyu Xu , Adrian Herzog , Alexander Lodermeyer , Emanuël A. P. Habets , Albert G. Prinn

In this paper we address the problem of simultaneously tracking several moving audio sources, namely the problem of estimating source trajectories from a sequence of observed features. We propose to use the von Mises distribution to model…

声音 · 计算机科学 2019-04-11 Yutong Ban , Xavier Alameda-PIneda , Christine Evers , Radu Horaud

In this paper, we present a novel multi-channel speech extraction system to simultaneously extract multiple clean individual sources from a mixture in noisy and reverberant environments. The proposed method is built on an improved…

音频与语音处理 · 电气工程与系统科学 2021-06-17 Jisi Zhang , Catalin Zorila , Rama Doddipatla , Jon Barker

State-of-the-art under-determined audio source separation systems rely on supervised end-end training of carefully tailored neural network architectures operating either in the time or the spectral domain. However, these methods are…

音频与语音处理 · 电气工程与系统科学 2020-05-29 Vivek Narayanaswamy , Jayaraman J. Thiagarajan , Rushil Anirudh , Andreas Spanias

Audio source separation aims to separate a mixture into target sources. Previous audio source separation systems usually conduct one-step inference, which does not fully explore the separation ability of models. In this work, we reveal that…

声音 · 计算机科学 2025-05-27 Yongyi Zang , Jingyi Li , Qiuqiang Kong

Modern neural network-based speech processing systems usually need to have reverberation resistance, so the training of such systems requires a large amount of reverberation data. In the process of system training, it is now more inclined…

音频与语音处理 · 电气工程与系统科学 2025-08-19 Dong Yang