中文
相关论文

相关论文: A Many to One Discrete Auditory Transform

200 篇论文

A crucial aspect for the successful deployment of audio-based models "in-the-wild" is the robustness to the transformations introduced by heterogeneous acquisition conditions. In this work, we propose a method to perform one-shot microphone…

声音 · 计算机科学 2020-10-20 Zalán Borsos , Yunpeng Li , Beat Gfeller , Marco Tagliasacchi

The slowly varying complex envelope of sinusoidal signals can be estimated in real-time using digital downconversion. In this paper, we discuss the requirements on digital downconversion for control applications. Two low-latency…

最优化与控制 · 数学 2021-02-12 Olof Troeng , Lawrence Doolittle

While efficient architectures and a plethora of augmentations for end-to-end image classification tasks have been suggested and heavily investigated, state-of-the-art techniques for audio classifications still rely on numerous…

声音 · 计算机科学 2022-07-06 Avi Gazneli , Gadi Zimerman , Tal Ridnik , Gilad Sharir , Asaf Noy

The intrinsic limitation of the material nonlinearity inevitably results in the poor mode purity, conversion efficiency and real-time reconfigurability of the generated harmonic waves, both in optics and acoustics. Rotational Doppler effect…

应用物理 · 物理学 2022-05-02 Chengbo Hu , Wei Wang , Jincheng Ni , Yujiang Ding , Jingkai Weng , Bin Liang , Cheng-Wei Qiu , Jianchun Cheng

Automatic music transcription (AMT) is the problem of analyzing an audio recording of a musical piece and detecting notes that are being played. AMT is a challenging problem, particularly when it comes to polyphonic music. The goal of AMT…

声音 · 计算机科学 2025-05-08 Yohannis Telila , Tommaso Cucinotta , Davide Bacciu

The state-of-the-art speech enhancement has limited performance in speech estimation accuracy. Recently, in deep learning, the Transformer shows the potential to exploit the long-range dependency in speech by self-attention. Therefore, it…

声音 · 计算机科学 2023-05-10 Yi Li , Yang Sun , Syed Mohsen Naqvi

Transformation-invariant analysis of signals often requires the computation of the distance from a test pattern to a transformation manifold. In particular, the estimation of the distances between a transformed query signal and several…

计算机视觉与模式识别 · 计算机科学 2011-12-26 Elif Vural , Pascal Frossard

Previous speech enhancement methods focus on estimating the short-time spectrum of speech signals due to its short-term stability. However, these methods often only estimate the clean magnitude spectrum and reuse the noisy phase when…

声音 · 计算机科学 2019-10-23 Chuang Geng , Lei Wang

We present a Cycle-GAN based many-to-many voice conversion method that can convert between speakers that are not in the training set. This property is enabled through speaker embeddings generated by a neural network that is jointly trained…

音频与语音处理 · 电气工程与系统科学 2019-05-08 Gokce Keskin , Tyler Lee , Cory Stephenson , Oguz H. Elibol

Many people enjoy classical symphonic music. Its diverse instrumentation makes for a rich listening experience. This diversity adds to the conductor's expressive freedom to shape the sound according to their imagination. As a result, the…

Computational and human perception are often considered separate approaches for studying sound changes over time; few works have touched on the intersection of both. To fill this research gap, we provide a pioneering review contrasting…

计算与语言 · 计算机科学 2024-07-09 Siqi He , Wei Zhao

Multitrack music transcription aims to transcribe a music audio input into the musical notes of multiple instruments simultaneously. It is a very challenging task that typically requires a more complex model to achieve satisfactory result.…

声音 · 计算机科学 2023-06-21 Wei-Tsung Lu , Ju-Chiang Wang , Yun-Ning Hung

In this Letter, we present a theoretical analysis of the acoustic transmission through two-dimensional arrays of straight rigid cylinders placed parallelly in the air. Both periodic and completely random arrangements of the cylinders are…

软凝聚态物质 · 物理学 2009-11-07 You-Yu Chen , Zhen Ye

Most deep learning-based multi-channel speech enhancement methods focus on designing a set of beamforming coefficients to directly filter the low signal-to-noise ratio signals received by microphones, which hinders the performance of these…

声音 · 计算机科学 2022-02-08 Wenzhe Liu , Andong Li , Chengshi Zheng , Xiaodong Li

Amplification underlies the operation of many biological and engineering systems. Simple electrical, optical, and mechanical amplifiers are reciprocal: the backward coupling of the output to the input equals the forward coupling of the…

仪器与探测器 · 物理学 2011-04-15 Tobias Reichenbach , A. J. Hudspeth

Attackers may manipulate audio with the intent of presenting falsified reports, changing an opinion of a public figure, and winning influence and power. The prevalence of inauthentic multimedia continues to rise, so it is imperative to…

声音 · 计算机科学 2022-05-05 Emily R. Bartusiak , Edward J. Delp

Inline holographic imaging presents an ill-posed inverse problem of reconstructing objects' complex amplitude from recorded diffraction patterns. Although recent deep learning approaches have shown promise over classical phase retrieval…

光学 · 物理学 2025-07-02 Chanseok Lee , Fakhriyya Mammadova , Jiseong Barg , Mooseok Jang

In speech processing pipelines, improving the quality and intelligibility of real-world recordings is crucial. While supervised regression is the primary method for speech enhancement, audio tokenization is emerging as a promising…

声音 · 计算机科学 2025-07-18 Luca Della Libera , Cem Subakan , Mirco Ravanelli

Approximate message passing (AMP) is an algorithmic framework for solving linear inverse problems from noisy measurements, with exciting applications such as reconstructing images, audio, hyper spectral images, and various other signals,…

信息论 · 计算机科学 2017-02-13 Junan Zhu , Ryan Pilgrim , Dror Baron

Audio signal processing frequently requires time-frequency representations and in many applications, a non-linear spacing of frequency-bands is preferable. This paper introduces a framework for efficient implementation of invertible signal…

泛函分析 · 数学 2013-05-17 Nicki Holighaus , Monika Dörfler , Gino Angelo Velasco , Thomas Grill