中文
相关论文

相关论文: An Explicit Consistency-Preserving Loss Function f…

200 篇论文

Variational Autoencoders (VAEs) are powerful generative models, however their generated samples are known to suffer from a characteristic blurriness, as compared to the outputs of alternative generating techniques. Extensive research…

图像与视频处理 · 电气工程与系统科学 2024-01-09 Vibhu Dalal

Conventional methods for speech enhancement rely on handcrafted loss functions (e.g., time or frequency domain losses) or deep feature losses (e.g., using WavLM or wav2vec), which often fail to capture subtle signal properties essential for…

声音 · 计算机科学 2025-05-28 Saisamarth Rajesh Phaye , Milos Cernak , Andrew Harper

Sparse modeling is one of the efficient techniques for imaging that allows recovering lost information. In this paper, we present a novel iterative phase-retrieval algorithm using a sparse representation of the object amplitude and phase.…

计算机视觉与模式识别 · 计算机科学 2011-08-17 Artem Migukin , Vladimir Katkovnik , Jaakko Astola

The problem of phase retrieval is a classic one in optics and arises when one is interested in recovering an unknown signal from the magnitude (intensity) of its Fourier transform. While there have existed quite a few approaches to phase…

信息论 · 计算机科学 2015-10-28 Kishore Jaganathan , Yonina C. Eldar , Babak Hassibi

Speech super-resolution (SSR) enhances low-resolution speech by increasing the sampling rate. While most SSR methods focus on magnitude reconstruction, recent research highlights the importance of phase reconstruction for improved…

Simulating the long-term dynamics of multi-scale and multi-physics systems poses a significant challenge in understanding complex phenomena across science and engineering. The complexity arises from the intricate interactions between scales…

机器学习 · 计算机科学 2025-09-22 Da Long , Shandian Zhe , Samuel Williams , Leonid Oliker , Zhe Bai

Recent advances in deep learning have significantly improved multichannel speech enhancement algorithms, yet conventional training loss functions such as the scale-invariant signal-to-distortion ratio (SDR) may fail to preserve fine-grained…

声音 · 计算机科学 2025-06-24 Nasser-Eddine Monir , Paul Magron , Romain Serizel

In many areas of imaging science, it is difficult to measure the phase of linear measurements. As such, one often wishes to reconstruct a signal from intensity measurements, that is, perform phase retrieval. In this paper, we provide a…

信息论 · 计算机科学 2013-09-13 Boris Alexeev , Afonso S. Bandeira , Matthew Fickus , Dustin G. Mixon

This paper proposes an approach to the joint modeling of the short-time Fourier transform magnitude and phase spectrograms with a deep generative model. We assume that the magnitude follows a Gaussian distribution and the phase follows a…

声音 · 计算机科学 2022-07-18 Aditya Arie Nugraha , Kouhei Sekiguchi , Kazuyoshi Yoshii

We propose a novel iterative phase estimation framework, termed multi-source Griffin-Lim algorithm (MSGLA), for speech enhancement (SE) under additive noise conditions. The core idea is to leverage the ad-hoc consistency constraint of…

音频与语音处理 · 电气工程与系统科学 2025-07-04 Chun-Wei Ho , Pin-Jui Ku , Hao Yen , Sabato Marco Siniscalchi , Yu Tsao , Chin-Hui Lee

The quality of inverse problem solutions obtained through deep learning [Barbastathis et al, 2019] is limited by the nature of the priors learned from examples presented during the training phase. In the case of quantitative phase retrieval…

图像与视频处理 · 电气工程与系统科学 2019-07-30 Mo Deng , Shuai Li , Alexandre Goy , Iksung Kang , George Barbastathis

The problem of signal recovery from its Fourier transform magnitude is of paramount importance in various fields of engineering and has been around for over 100 years. Due to the absence of phase information, some form of additional…

信息论 · 计算机科学 2015-07-02 Kishore Jaganathan , Samet Oymak , Babak Hassibi

Deep neural network (DNN) based end-to-end optimization in the complex time-frequency (T-F) domain or time domain has shown considerable potential in monaural speech separation. Many recent studies optimize loss functions defined solely in…

声音 · 计算机科学 2022-01-05 Zhong-Qiu Wang , Gordon Wichern , Jonathan Le Roux

The popular VQ-VAE models reconstruct images through learning a discrete codebook but suffer from a significant issue in the rapid quality degradation of image reconstruction as the compression rate rises. One major reason is that a higher…

计算机视觉与模式识别 · 计算机科学 2023-11-07 Xinmiao Lin , Yikang Li , Jenhao Hsiao , Chiuman Ho , Yu Kong

Vocal dereverberation remains a challenging task in audio processing, particularly for real-time applications where both accuracy and efficiency are crucial. Traditional deep learning approaches often struggle to suppress reverberation…

声音 · 计算机科学 2025-10-02 Daniel G. Williams

The speaker extraction algorithm extracts the target speech from a mixture speech containing interference speech and background noise. The extraction process sometimes over-suppresses the extracted target speech, which not only creates…

音频与语音处理 · 电气工程与系统科学 2022-06-22 Zexu Pan , Meng Ge , Haizhou Li

Recently, deep neural networks (DNNs) have been successfully used for speech enhancement, and DNN-based speech enhancement is becoming an attractive research area. While time-frequency masking based on the short-time Fourier transform…

音频与语音处理 · 电气工程与系统科学 2020-08-21 Yuichiro Koyama , Tyler Vuong , Stefan Uhlich , Bhiksha Raj

The ill-posed problem of phase retrieval in optics, using one or more intensity measurements, has a multitude of applications using electromagnetic or matter waves. Many phase retrieval algorithms are computed on pixel arrays using discrete…

图像与视频处理 · 电气工程与系统科学 2022-09-21 J. A. Pollock , K. S. Morgan , L. C. P. Croton , M. K. Croughan , G. Ruben , N. Yagi , H. Sekiguchi , M. J. Kitchen

The problem of phase retrieval, i.e., the problem of recovering a function from the magnitudes of its Fourier transform, naturally arises in various fields of physics, such as astronomy, radar, speech recognition, quantum mechanics and,…

泛函分析 · 数学 2020-02-17 Philipp Grohs , Sarah Koppensteiner , Martin Rathmair

A core challenge for signal data recovery is to model the distribution of signal matrix (SM) data based on measured low-quality data in biomedical engineering of magnetic particle imaging (MPI). For acquiring the high-resolution…

信号处理 · 电气工程与系统科学 2025-01-09 Liwen Zhang , Zhaoji Miao , Fan Yang , Gen Shi , Jie He , Yu An , Hui Hui , Jie Tian