中文
相关论文

相关论文: Log Complex Color for Visual Pattern Recognition o…

200 篇论文

This paper introduces a novel audio-to-image encoding framework that integrates multiple dimensions of voice characteristics into a single RGB image for speaker recognition. In this method, the green channel encodes raw audio data, the red…

声音 · 计算机科学 2025-03-11 Youness Atif

Conventional optical coherent receivers capture the full electrical field, including amplitude and phase, of a signal waveform by measuring its interference against a stable continuous-wave local oscillator (LO). In optical coherent…

信号处理 · 电气工程与系统科学 2020-06-24 Haoshuo Chen , Nicolas K. Fontaine , Joan M. Gene , Roland Ryf , David T. Neilson , Gregory Raybon

We present a new system for simultaneous estimation of keys, chords, and bass notes from music audio. It makes use of a novel chromagram representation of audio that takes perception of loudness into account. Furthermore, it is fully based…

声音 · 计算机科学 2011-07-26 Yizhao Ni , Matt Mcvicar , Raul Santos-Rodriguez , Tijl De Bie

Most deep learning-based models for speech enhancement have mainly focused on estimating the magnitude of spectrogram while reusing the phase from noisy speech for reconstruction. This is due to the difficulty of estimating the phase of…

声音 · 计算机科学 2019-04-03 Hyeong-Seok Choi , Jang-Hyun Kim , Jaesung Huh , Adrian Kim , Jung-Woo Ha , Kyogu Lee

Music source separation in the time-frequency domain is commonly achieved by applying a soft or binary mask to the magnitude component of (complex) spectrograms. The phase component is usually not estimated, but instead copied from the…

声音 · 计算机科学 2021-03-25 Andreas Jansson , Rachel M. Bittner , Nicola Montecchio , Tillman Weyde

The waveforms of attosecond pulses produced by high-harmonic generation carry information on the electronic structure and dynamics in atomic and molecular systems. Current methods for the temporal characterization of such pulses have…

Computational spectral imaging is drawing increasing attention owing to the snapshot advantage, and amplitude, phase, and wavelength encoding systems are three types of representative implementations. Fairly comparing and understanding the…

图像与视频处理 · 电气工程与系统科学 2023-12-22 Xinyuan Liu , Lizhi Wang , Lingen Li , Chang Chen , Xue Hu , Fenglong Song , Youliang Yan

One of the decisions that arise when designing a neural network for any application is how the data should be represented in order to be presented to, and possibly generated by, a neural network. For audio, the choice is less obvious than…

声音 · 计算机科学 2017-06-30 L. Wyse

This paper considers the question of recovering the phase of an object from intensity-only measurements, a problem which naturally appears in X-ray crystallography and related disciplines. We study a physically realistic setup where one can…

信息论 · 计算机科学 2013-11-08 Emmanuel Candes , Xiaodong Li , Mahdi Soltanolkotabi

We describe a new algorithm to solve a particular phase retrieval problem, that has wide applications in audio processing: the reconstruction of a function from its scalogram, that is from the modulus of its wavelet transform. It is a…

最优化与控制 · 数学 2017-04-11 Irène Waldspurger

Like the ordinary power spectrum, higher-order spectra (HOS) describe signal properties that are invariant under translations in time. Unlike the power spectrum, HOS retain phase information from which details of the signal waveform can be…

信号处理 · 电气工程与系统科学 2019-08-27 Christopher K. Kovach , Matthew A. Howard

This paper introduces an improved image processing method usable in capacitive imaging applications. Standard capacitive imaging tends to prefer amplitude-based images over the use of phase due to better signal-to-noise ratios. The new…

信号处理 · 电气工程与系统科学 2022-01-06 Silvio Amato , David Hutchins , Xiaokang Yin , Marco Ricci , Stefano Laureti

Spectroscopy is an indispensable tool in understanding the structures and dynamics of molecular systems. However computational modelling of spectroscopy is challenging due to the exponential scaling of computational complexity with system…

量子物理 · 物理学 2021-06-22 Chee-Kong Lee , Chang-Yu Hsieh , Shengyu Zhang , Liang Shi

Multiplicative noise models occur in the study of several coherent imaging systems, such as synthetic aperture radar and sonar, and ultrasound and laser imaging. This type of noise is also commonly referred to as speckle. Multiplicative…

最优化与控制 · 数学 2009-03-25 Jose M. Bioucas-Dias , Mario A. T. Figueiredo

Coherent diffractive imaging is a technique that recovers the sample image by numerically inverting its diffraction pattern. We propose a generalization of this method for the inversion of multi-wavelength data. Using this approach, we show…

光学 · 物理学 2020-05-08 Erik Malm , Edwin Fohtung , Anders Mikkelsen

Style transfer is a technique for combining two images based on the activations and feature statistics in a deep learning neural network architecture. This paper studies the analogous task in the audio domain and takes a critical look at…

声音 · 计算机科学 2020-08-10 M. Huzaifah , L. Wyse

We present a reconstruction technique for simultaneous retrieval of absorption and phase shifting properties of an object recorded by in-line holography. The routine is experimentally tested by applying it to optical holograms of a pure…

光学 · 物理学 2010-10-28 Tatiana Latychevskaia , Hans-Werner Fink

Recent advancements in deep learning have significantly impacted the field of speech signal processing, particularly in the analysis and manipulation of complex spectrograms. This survey provides a comprehensive overview of the…

音频与语音处理 · 电气工程与系统科学 2025-10-06 Yuying Xie , Zheng-Hua Tan

Quantitative phase imaging (QPI) is important in many applications such as microscopy and crystallography. To quantitatively reveal phase information, people could either employ interference to map phase distribution into intensity fringes,…

光学 · 物理学 2020-11-11 Xianye Li , Yafei sun , Yikang He , Xun Li , Baoqing Sun

With active research in audio compression techniques yielding substantial breakthroughs, spectral reconstruction of low-quality audio waves remains a less indulged topic. In this paper, we propose a novel approach for reconstructing higher…

声音 · 计算机科学 2021-08-10 Darshan Deshpande , Harshavardhan Abichandani