中文
相关论文

相关论文: A Many to One Discrete Auditory Transform

200 篇论文

Invariant-based inverse engineering is an elegant approach to quantum control with corresponding experimental implementations that perform tasks with applications in quantum information processing such as shuttling trapped ions. We build on…

量子物理 · 物理学 2021-12-30 Selwyn Simsek , Florian Mintert

Automatic Music Transcription (AMT) -- the task of converting music audio into note representations -- has seen rapid progress, driven largely by deep learning systems. Due to the limited availability of richly annotated music datasets,…

声音 · 计算机科学 2026-01-27 Lukáš Samuel Marták , Patricia Hu , Gerhard Widmer

This paper describes a method of calculating the transforms, currently obtained via Fourier and reverse Fourier transforms. The method allows calculating efficiently the transforms of a signal having an arbitrary dimension of the digital…

数值分析 · 数学 2025-10-20 Vladimir I Clue

Directly sending audio signals from a transmitter to a receiver across a noisy channel may absorb consistent bandwidth and be prone to errors when trying to recover the transmitted bits. On the contrary, the recent semantic communication…

声音 · 计算机科学 2023-09-15 Eleonora Grassucci , Christian Marinoni , Andrea Rodriguez , Danilo Comminiello

We prove an important property of the binomial transform: it converts multiplication by the discrete variable into a certain difference operator. We also consider the case of dividing by the discrete variable. The properties presented here…

数论 · 数学 2017-02-03 Khristo N. Boyadzhiev

Speech production involves the movement of various articulators, including tongue, jaw, and lips. Estimating the movement of the articulators from the acoustics of speech is known as acoustic-to-articulatory inversion (AAI). Recently, it…

音频与语音处理 · 电气工程与系统科学 2020-06-23 Aravind Illa , Prasanta Kumar Ghosh

Device-guided music transfer adapts playback across unseen devices for users who lack them. Existing methods mainly focus on modifying the timbre, rhythm, harmony, or instrumentation to mimic genres or artists, overlooking the diverse…

声音 · 计算机科学 2025-11-24 Manh Pham Hung , Changshuo Hu , Ting Dang , Dong Ma

Certain quantum devices, such as half-wave plates and quarter-wave plates in quantum optics, are bidirectional, meaning that the roles of their input and output ports can be exchanged. Bidirectional devices can be used in a forward mode and…

量子物理 · 物理学 2023-04-20 Zixuan Liu , Ming Yang , Giulio Chiribella

We introduce the joint time-frequency scattering transform, a time shift invariant descriptor of time-frequency structure for audio classification. It is obtained by applying a two-dimensional wavelet transform in time and log-frequency to…

声音 · 计算机科学 2018-08-06 Joakim Andén , Vincent Lostanlen , Stéphane Mallat

This study proposes a multi-microphone complex spectral mapping approach for speech dereverberation on a fixed array geometry. In the proposed approach, a deep neural network (DNN) is trained to predict the real and imaginary (RI)…

音频与语音处理 · 电气工程与系统科学 2020-03-05 Zhong-Qiu Wang , DeLiang Wang

Signal scaling is a fundamental operation of practical importance in which a signal is enlarged or shrunk in the coordinate direction(s). Scaling or magnification is not trivial for signals of a discrete variable since the signal values may…

信号处理 · 电气工程与系统科学 2021-01-19 Aykut Koç , Burak Bartan , Haldun M. Ozaktas

This paper presents CQT-Diff, a data-driven generative audio model that can, once trained, be used for solving various different audio inverse problems in a problem-agnostic setting. CQT-Diff is a neural diffusion model with an architecture…

音频与语音处理 · 电气工程与系统科学 2023-03-21 Eloi Moliner , Jaakko Lehtinen , Vesa Välimäki

Speech-to-speech translation is a typical sequence-to-sequence learning task that naturally has two directions. How to effectively leverage bidirectional supervision signals to produce high-fidelity audio for both directions? Existing…

计算与语言 · 计算机科学 2023-05-23 Xianchao Wu

We propose an end-to-end music mixing style transfer system that converts the mixing style of an input multitrack to that of a reference song. This is achieved with an encoder pre-trained with a contrastive objective to extract only audio…

音频与语音处理 · 电气工程与系统科学 2023-04-12 Junghyun Koo , Marco A. Martínez-Ramírez , Wei-Hsiang Liao , Stefan Uhlich , Kyogu Lee , Yuki Mitsufuji

Identity, accent, style, and emotions are essential components of human speech. Voice conversion (VC) techniques process the speech signals of two input speakers and other modalities of auxiliary information such as prompts and emotion…

音频与语音处理 · 电气工程与系统科学 2025-12-09 Xining Song , Zhihua Wei , Rui Wang , Haixiao Hu , Yanxiang Chen , Meng Han

A mechanism called switched feedback is introduced; under switched feedback, each channel output goes forward to the receiver(s) or back to the transmitter(s) but never both. By studying the capacity of the Multiple-Access Channel (MAC)…

信息论 · 计算机科学 2025-05-01 Oliver Kosut , Michael Langberg , Michelle Effros

This study focuses on the perception of music performances when contextual factors, such as room acoustics and instrument, change. We propose to distinguish the concept of "performance" from the one of "interpretation", which expresses the…

声音 · 计算机科学 2022-03-08 Federico Simonetta , Federico Avanzini , Stavros Ntalampiras

Visual-to-auditory sensory substitution devices can assist the blind in sensing the visual environment by translating the visual information into a sound pattern. To improve the translation quality, the task performances of the blind are…

计算机视觉与模式识别 · 计算机科学 2019-04-22 Di Hu , Dong Wang , Xuelong Li , Feiping Nie , Qi Wang

We consider the inverse problem of quantitative reconstruction of properties (e.g., bulk modulus, density) of visco-acoustic materials based on measurements of responding waves after stimulation of the medium. Numerical reconstruction is…

偏微分方程分析 · 数学 2022-09-20 Florian Faucher , Otmar Scherzer

Sequence-to-Sequence Text-to-Speech architectures that directly generate low level acoustic features from phonetic sequences are known to produce natural and expressive speech when provided with adequate amounts of training data. Such…

音频与语音处理 · 电气工程与系统科学 2022-07-26 Raul Fernandez , David Haws , Guy Lorberbom , Slava Shechtman , Alexander Sorin