中文
相关论文

相关论文: Audio dequantization using instantaneous frequency

200 篇论文

We address the exploitation of an optical parametric oscillator (OPO) in the task of mitigating, at least partially, phase noise produced by phase diffusion. In particular, we analyze two scenarios where phase diffusion is typically…

We report on the realization of an optical phase noise cancellation technique by passively embedding the optical phase information into a radio frequency (RF) signal and shifting the optical frequency with the amount of phase noise…

仪器与探测器 · 物理学 2020-07-30 Liang Hu , Xueyang Tian , Guiling Wu , Jianping Chen

Environmental sound recordings often contain intelligible speech, raising privacy concerns that limit analysis, sharing and reuse of data. In this paper, we introduce a method that renders speech unintelligible while preserving both the…

声音 · 计算机科学 2025-07-14 Modan Tailleur , Mathieu Lagrange , Pierre Aumond , Vincent Tourre

The residual vector quantization (RVQ) technique plays a central role in recent advances in neural audio codecs. These models effectively synthesize high-fidelity audio from a limited number of codes due to the hierarchical structure among…

音频与语音处理 · 电气工程与系统科学 2025-09-24 Hyeongju Kim , Junhyeok Lee , Jacob Morton , Juheon Lee , Jinhyeok Yang

The phase vocoder (PV) is a widely spread technique for processing audio signals. It employs a short-time Fourier transform (STFT) analysis-modify-synthesis loop and is typically used for time-scaling of signals by means of using different…

声音 · 计算机科学 2022-02-16 Zdenek Prusa , Nicki Holighaus

Quantum annealing (QA) is one of the efficient methods to calculate the ground-state energy of a problem Hamiltonian. In the absence of noise, QA can accurately estimate the ground-state energy if the adiabatic condition is satisfied.…

量子物理 · 物理学 2022-10-18 Yuta Shingu , Tetsuro Nikuni , Shiro Kawabata , Yuichiro Matsuzaki

We introduce a new method for reducing phase noise in oscillators, thereby improving their frequency precision. The noise reduction device consists of a pair of coupled nonlinear resonating elements that are driven parametrically by the…

介观与纳米尺度物理 · 物理学 2012-07-18 Eyal Kenig , M. C. Cross , Ron Lifshitz , R. B. Karabalin , L. G. Villanueva , M. H. Matheny , M. L. Roukes

ITU-R BS.1387 states a method for objective assessment of perceived audio quality. This Recommendation, known also as PEAQ (Perceptual Evaluation of Audio Quality) is based on a psychoacoustic model of the human ear and was standardized by…

音频与语音处理 · 电气工程与系统科学 2019-10-31 Luis F. Abanto-Leon , Guillermo Kemper Vasquez , Joel Telles

Recent neural audio codecs have achieved impressive reconstruction quality, typically relying on quantization methods such as Residual Vector Quantization (RVQ), Vector Quantization (VQ) and Finite Scalar Quantization (FSQ). However, these…

声音 · 计算机科学 2026-05-19 Tal Shuster , Eliya Nachmani

Diffusion models have recently dominated image synthesis tasks. However, the iterative denoising process is expensive in computations at inference time, making diffusion models less practical for low-latency and scalable real-world…

计算机视觉与模式识别 · 计算机科学 2023-11-02 Yefei He , Luping Liu , Jing Liu , Weijia Wu , Hong Zhou , Bohan Zhuang

We introduce a simple and linear SNR (strictly speaking, periodic to random power ratio) estimator (0dB to 80dB without additional calibration/linearization) for providing reliable descriptions of aperiodicity in speech corpus. The main…

音频与语音处理 · 电气工程与系统科学 2018-07-06 Hideki Kawahara , Ken-Ichi Sakakibara , Masanori Morise , Hideki Banno , Tomoki Toda

In neural-based audio feature extraction, ensuring that representations capture disentangled information is crucial for model interpretability. However, existing disentanglement methods often rely on assumptions that are highly dependent on…

声音 · 计算机科学 2025-10-07 Benoit Ginies , Xiaoyu Bie , Olivier Fercoq , Gaël Richard

In the field of deepfake detection, previous studies focus on using reconstruction or mask and prediction methods to train pre-trained models, which are then transferred to fake audio detection training where the encoder is used to extract…

Neural audio compression has emerged as a promising technology for efficiently representing speech, music, and general audio. However, existing methods suffer from significant performance degradation at limited bitrates, where the available…

声音 · 计算机科学 2026-05-08 Jin Wang , Wenbin Jiang , Xiangbo Wang , Yubo You , Sheng Fang

Phase noise correction is crucial to exploit full advantage of orthogonal frequency division multiplexing (OFDM) in modern high-data-rate communications. OFDM channel estimation with simultaneous phase noise compensation has therefore drawn…

信息论 · 计算机科学 2017-04-25 Zhongju Wang , Prabhu Babu , Daniel P. Palomar

Existing fraud detection methods predominantly rely on transcribed text, suffering from ASR errors and missing crucial acoustic cues like vocal tone and environmental context. This limits their effectiveness against complex deceptive…

Modern neural speech enhancement models usually include various forms of phase information in their training loss terms, either explicitly or implicitly. However, these loss terms are typically designed to reduce the distortion of phase…

声音 · 计算机科学 2022-02-25 Doyeon Kim , Hyewon Han , Hyeon-Kyeong Shin , Soo-Whan Chung , Hong-Goo Kang

Neural audio codecs (NACs) typically encode the short-term energy (gain) and normalized structure (shape) of speech/audio signals jointly within the same latent space. As a result, they are poorly robust to a global variation of the input…

声音 · 计算机科学 2026-02-18 Samir Sadok , Laurent Girin , Xavier Alameda-Pineda

Wave is crucial to acquiring information from the world and its interaction with matter is determined by the wavelength, or frequency. Search for the ability to shift frequency often points to material nonlinearity, which is significant…

应用物理 · 物理学 2020-09-24 Yumin Zhang , Keming Wu , Chunqi Wang , Lixi Huang

Phase retrieval (PR) aims to recover a signal from the magnitudes of a set of inner products. This problem arises in many audio signal processing applications which operate on a short-time Fourier transform magnitude or power spectrogram,…

声音 · 计算机科学 2021-02-24 Pierre-Hugo Vial , Paul Magron , Thomas Oberlin , Cédric Févotte