English
Related papers

Related papers: Phase reconstruction from amplitude spectrograms b…

200 papers

This paper proposes a novel algorithm for image phase retrieval, i.e., for recovering complex-valued images from the amplitudes of noisy linear combinations (often the Fourier transform) of the sought complex images. The algorithm is…

Signal Processing · Electrical Eng. & Systems 2018-10-19 Joshin P. Krishnan , José M. Bioucas-Dias , Vladimir Katkovnik

This paper presents a novel neural speech phase prediction model which predicts wrapped phase spectra directly from amplitude spectra. The proposed model is a cascade of a residual convolutional network and a parallel estimation…

Sound · Computer Science 2024-03-27 Yang Ai , Zhen-Hua Ling

In this paper, we introduce a dictionary learning based approach applied to the problem of real-time reconstruction of MR image sequences that are highly undersampled in k-space. Unlike traditional dictionary learning, our method integrates…

Computer Vision and Pattern Recognition · Computer Science 2015-02-19 Dornoosh Zonoobi , Shahrooz Faghih Roohi , Ashraf A. Kassim

Discrete diffusion models have recently emerged as strong alternatives to autoregressive language models, matching their performance through large-scale training. However, inference-time control remains relatively underexplored. In this…

Machine Learning · Computer Science 2026-04-09 Meihua Dang , Jiaqi Han , Minkai Xu , Kai Xu , Akash Srivastava , Stefano Ermon

Convolutional Neural Networks (CNN) based image reconstruction methods have been intensely used for X-ray computed tomography (CT) reconstruction applications. Despite great success, good performance of this data-based approach critically…

Computer Vision and Pattern Recognition · Computer Science 2019-01-31 Ziling Wu , Abdulaziz Alorf , Ting Yang , Ling Li , Yunhui Zhu

This study presents a novel model for invertible sentence embeddings using a residual recurrent network trained on an unsupervised encoding task. Rather than the probabilistic outputs common to neural machine translation models, our…

Computation and Language · Computer Science 2023-04-07 Jeremy Wilkerson

Single Domain Generalization (SDG) for object detection aims to train a model on a single source domain that can generalize effectively to unseen target domains. While recent methods like CLIP-based semantic augmentation have shown promise,…

Computer Vision and Pattern Recognition · Computer Science 2025-11-14 Mengzhu Wang , Changyuan Deng , Shanshan Wang , Nan Yin , Long Lan , Liang Yang

Recently, digital holographic imaging techniques (including methods with heterodyne detection) have found increased attention in the terahertz (THz) frequency range. However, holographic techniques rely on the use of a reference beam in…

Optics · Physics 2022-12-14 Mingjun Xiang , Hui Yuan , Lingxiao Wang , Kai Zhou , Hartmut G. Roskos

This paper presents a neural vocoder named HiNet which reconstructs speech waveforms from acoustic features by predicting amplitude and phase spectra hierarchically. Different from existing neural vocoders such as WaveNet, SampleRNN and…

Sound · Computer Science 2020-02-06 Yang Ai , Zhen-Hua Ling

Recent developments in acoustic signal processing have seen the integration of deep learning methodologies, alongside the continued prominence of classical wave expansion-based approaches, particularly in sound field reconstruction.…

Audio and Speech Processing · Electrical Eng. & Systems 2024-04-24 Marco Olivieri , Xenofon Karakonstantis , Mirco Pezzoli , Fabio Antonacci , Augusto Sarti , Efren Fernandez-Grande

We study the problem of recovering the phase from magnitude measurements; specifically, we wish to reconstruct a complex-valued signal x of C^n about which we have phaseless samples of the form y_r = |< a_r,x >|^2, r = 1,2,...,m (knowledge…

Information Theory · Computer Science 2016-11-17 Emmanuel Candes , Xiaodong Li , Mahdi Soltanolkotabi

Although deep neural network (DNN)-based speech enhancement (SE) methods outperform the previous non-DNN-based ones, they often degrade the perceptual quality of generated outputs. To tackle this problem, we introduce a DNN-based generative…

Audio and Speech Processing · Electrical Eng. & Systems 2023-08-31 Ryosuke Sawata , Naoki Murata , Yuhta Takida , Toshimitsu Uesaka , Takashi Shibuya , Shusuke Takahashi , Yuki Mitsufuji

Recent studies in deep learning-based speech separation have proven the superiority of time-domain approaches to conventional time-frequency-based methods. Unlike the time-frequency domain approaches, the time-domain separation systems…

Audio and Speech Processing · Electrical Eng. & Systems 2020-03-30 Yi Luo , Zhuo Chen , Takuya Yoshioka

Oversmoothing remains a persistent problem when applying deep learning to off-axis quantitative phase imaging (QPI). End-to-end U-Nets favour low-frequency content and under-represent fine, diagnostic detail. We trace this issue to spectral…

Image and Video Processing · Electrical Eng. & Systems 2025-06-16 Yi Zhang

Synthesizing human's movements such as dancing is a flourishing research field which has several applications in computer graphics. Recent studies have demonstrated the advantages of deep neural networks (DNNs) for achieving remarkable…

Machine Learning · Computer Science 2019-06-24 Nelson Yalta , Shinji Watanabe , Kazuhiro Nakadai , Tetsuya Ogata

Graph Neural Networks (GNNs) are deep-learning architectures designed for graph-type data, where understanding relationships among individual observations is crucial. However, achieving promising GNN performance, especially on unseen data,…

Machine Learning · Computer Science 2024-05-22 Lequan Lin , Dai Shi , Andi Han , Zhiyong Wang , Junbin Gao

We investigate the learning dynamics of fully-connected neural networks through the lens of gradient signal-to-noise ratio (SNR), examining the behavior of first-order optimizers like Adam in non-convex objectives. By interpreting the…

Machine Learning · Computer Science 2024-03-28 Sokratis J. Anagnostopoulos , Juan Diego Toscano , Nikolaos Stergiopulos , George Em Karniadakis

Signal extraction from a single-channel mixture with additional undesired signals is most commonly performed using time-frequency (TF) masks. Typically, the mask is estimated with a deep neural network (DNN), and element-wise applied to the…

Sound · Computer Science 2019-12-10 Wolfgang Mack , Emanuël A. P. Habets

Inspired by the remarkable learning and prediction performance of deep neural networks (DNNs), we apply one special type of DNN framework, known as model-driven deep unfolding neural network, to reconfigurable intelligent surface…

Signal Processing · Electrical Eng. & Systems 2021-12-06 Jiguang He , Henk Wymeersch , Marco Di Renzo , Markku Juntti

This paper addresses the reconstruction of periodic structures using phase or phaseless near-field measurements. We introduce a novel illumination strategy based on the quasi-periodic condition. Employing the Dirichlet-to-Neumann (DtN) map,…

Mathematical Physics · Physics 2024-04-12 Jue Wang , Yujie Wang , Lei Zhang , Enxi Zheng