中文
相关论文

相关论文: Audio signal interpolation using optimal transport…

200 篇论文

This paper proposes a new method to interpolate between two audio signals. As an interpolation parameter is changed, the pitches in one signal slide to the pitches in the other, producing a portamento, or musical glide. The assignment of…

音频与语音处理 · 电气工程与系统科学 2019-06-18 Trevor Henderson , Justin Solomon

Audio source separation is often achieved by estimating the magnitude spectrogram of each source, and then applying a phase recovery (or spectrogram inversion) algorithm to retrieve time-domain signals. Typically, spectrogram inversion is…

声音 · 计算机科学 2023-07-03 Paul Magron , Tuomas Virtanen

This paper introduces a new nonlinear dictionary learning method for histograms in the probability simplex. The method leverages optimal transport theory, in the sense that our aim is to reconstruct histograms using so-called displacement…

This study introduces a novel approach for estimating plane-wave coefficients in sound field reconstruction, specifically addressing challenges posed by error-in-variable phase perturbations. Such systematic errors typically arise from…

信号处理 · 电气工程与系统科学 2026-02-06 Yuyang Liu , Johan Karlsson , Filip Elvander

Time-frequency representations, such as the short-time Fourier transform (STFT), are fundamental tools for analyzing non-stationary signals. However, their ability to achieve sharp localization in both time and frequency is inherently…

信号处理 · 电气工程与系统科学 2026-04-17 David Valdivia , Elsa Cazelles , Cédric Févotte

There has been fascinating work on creating artistic transformations of images by Gatys. This was revolutionary in how we can in some sense alter the 'style' of an image while generally preserving its 'content'. In our work, we present a…

声音 · 计算机科学 2024-12-24 Prateek Verma , Julius O. Smith

We present a novel method for efficiently computing optimal transport maps and Wasserstein barycenters in high-dimensional spaces. Our approach uses conditional normalizing flows to approximate the input distributions as invertible…

机器学习 · 统计学 2025-05-29 Gabriele Visentin , Patrick Cheridito

Separating audio mixtures into individual instrument tracks has been a long standing challenging task. We introduce a novel weakly supervised audio source separation approach based on deep adversarial learning. Specifically, our loss…

声音 · 计算机科学 2018-05-18 Ning Zhang , Junchi Yan , Yuchen Zhou

A wireless acoustic sensor network records audio signals with sampling time and sampling rate offsets between the audio streams, if the analog-digital converters (ADCs) of the network devices are not synchronized. Here, we introduce a new…

音频与语音处理 · 电气工程与系统科学 2021-10-26 Tobias Gburrek , Joerg Schmalenstroeer , Reinhold Haeb-Umbach

The localization of sound sources by the human brain is computationally simulated from a neurobiological perspective. The simulation includes the neural representation of temporal differences in acoustic signals between the ipsilateral and…

神经元与认知 · 定量生物学 2008-10-31 Nikesh S. Dattani

Synthetic magnetism has been recently realized using spatiotemporal modulation patterns, producing non-reciprocal steering of charge-neutral particles such as photons and phonons. Here, we design and experimentally demonstrate a…

应用物理 · 物理学 2022-10-03 Zhaoxian Chen , Zhengwei Li , Jingkai Weng , Bin Liang , Yanqing Lu , Jianchun Cheng , Andrea Alu

A high-precision numerical sound field is the basis of underwater target detection, positioning and communication. A line source in a plane is a common type of sound source in computational ocean acoustics. The exciting waveguide in a…

计算物理 · 物理学 2022-08-17 Houwang Tu , Yongxian Wang , Chunmei Yang , Xiaodong Wang , Shuqing Ma , Wenbin Xiao , Wei Liu

We propose a novel audio watermarking system that is robust to the distortion due to the indoor acoustic propagation channel between the loudspeaker and the receiving microphone. The system utilizes a set of new algorithms that effectively…

多媒体 · 计算机科学 2019-03-21 Yuan-Yen Tai , Mohamed F. Mansour

Passive acoustic sensing is a cost-effective solution for monitoring moving targets such as vessels and aircraft, but its performance is hindered by complex propagation effects like multi-path reflections and motion-induced artefacts.…

声音 · 计算机科学 2026-01-23 Lucas C. F. Domingos , Russell S. A. Brinkworth , Paulo E. Santos , Karl Sammut

As a first step towards a complete computational model of speech learning involving perception-production loops, we investigate the forward mapping between pseudo-motor commands and articulatory trajectories. Two phonological feature sets,…

音频与语音处理 · 电气工程与系统科学 2024-08-09 Angelo Ortiz Tandazo , Thomas Schatz , Thomas Hueber , Emmanuel Dupoux

It is a well-established principle that cross-correlating seismic observations at different receiver locations can yield estimates of band-limited inter-receiver Green's functions. This principle, known as seismic interferometry, is a…

地球物理 · 物理学 2021-08-11 Daniella Ayala-Garcia , Andrew Curtis , Michal Branicki

Research on audio generation has progressively developed along both waveform-based and spectrogram-based directions, giving rise to diverse strategies for representing and generating audio. At the same time, advances in image synthesis have…

声音 · 计算机科学 2026-04-17 Eleonora Ristori , Luca Bindini , Paolo Frasconi

We propose Gaussian optimal transport for Image style transfer in an Encoder/Decoder framework. Optimal transport for Gaussian measures has closed forms Monge mappings from source to target distributions. Moreover interpolates between a…

机器学习 · 计算机科学 2019-05-31 Youssef Mroueh

We study the ability of Wasserstein Generative Adversarial Network (WGAN) to generate missing audio content which is, in context, (statistically similar) to the sound and the neighboring borders. We deal with the challenge of audio…

音频与语音处理 · 电气工程与系统科学 2020-03-18 P. P. Ebner , A. Eltelt

Approximate message passing (AMP) algorithms have shown great promise in sparse signal reconstruction due to their low computational requirements and fast convergence to an exact solution. Moreover, they provide a probabilistic framework…

声音 · 计算机科学 2018-02-02 Turab Iqbal , Wenwu Wang
‹ 上一页 1 2 3 10 下一页 ›