Related papers: Log Complex Color for Visual Pattern Recognition o…
This paper proposes APSS, a novel neural speech separation model with parallel amplitude and phase spectrum estimation. Unlike most existing speech separation methods, the APSS distinguishes itself by explicitly estimating the phase…
Encoding more information into wave fields is a central goal in imaging, communication, and wave control. Optical holography benefits from polarization multiplexing, but acoustic holography remains largely limited to pressure-only encoding…
We consider the imaging problem of the reconstruction of a three-dimensional object via optical diffraction tomography under the assumptions of the Born approximation. Our focus lies in the situation that a rigid object performs an…
We present a method to remove unknown convolutive noise introduced to speech by reverberations of recording environments, utilizing some amount of training speech data from the reverberant environment, and any available non-reverberant…
We present a novel method for the compensation of long duration data loss in audio signals, in particular music. The concealment of such signal defects is based on a graph that encodes signal structure in terms of time-persistent spectral…
Spectral rendering is essential for the production of physically-plausible synthetic images, but requires to introduce several changes in the content generation pipeline. In particular, the authoring of spectral material properties (e.g.,…
We present a novel approach to inspecting galaxy spectra using sound, via their direct audio representation ('spectral audification'). We discuss the potential of this as a complement to (or stand-in for) visual approaches. We surveyed 58…
Light passing through scattering media will be strongly scattered and diffused into complex speckle pattern, which however contains almost all the spatial information and color information of the objects. Although various technologies have…
Audio codecs power discrete music generative modelling, music streaming and immersive media by shrinking PCM audio to bandwidth-friendly bit-rates. Recent works have gravitated towards processing in the spectral domain; however,…
We consider imaging the reflectivity of scatterers from intensity-only data recorded by a single moving transducer that both emits and receives signals, forming a synthetic aperture. By exploiting frequency illumination diversity, we obtain…
Measuring the spectral phase of a pulse is key for performing wavelength resolved ultrafast measurements in the few femtosecond regime. However, accurate measurements in real experimental conditions can be challenging. We show that the…
In high-frequency photoacoustic imaging with uniform illumination, homogeneous photo-absorbing structures may be invisible because of their large size or limited-view issues. Here we show that, by exploiting dynamic speckle illumination, it…
The rise of deep learning algorithms has led many researchers to withdraw from using classic signal processing methods for sound generation. Deep learning models have achieved expressive voice synthesis, realistic sound textures, and…
Ultrathin flat meta-optics have shown great promise for holography in recent years. However, most of the reported meta-optical holograms rely on only phase modulation and neglect the amplitude information. Modulation of both amplitude and…
Pattern recognition from audio signals is an active research topic encompassing audio tagging, acoustic scene classification, music classification, and other areas. Spectrogram and mel-frequency cepstral coefficients (MFCC) are among the…
We propose a reliable direct imaging method based on the reverse time migration for finding extended obstacles with phaseless total field data. We prove that the imaging resolution of the method is essentially the same as the imaging…
A simple method of phase-and-amplitude extraction is derived that corrects for image blurring induced by partially spatially coherent incident illumination using only a single intensity image as input. The method is based on Fresnel…
Hair appearance is a complex phenomenon due to hair geometry and how the light bounces on different hair fibers. For this reason, reproducing a specific hair color in a rendering environment is a challenging task that requires manual work…
Musicians and audio engineers sculpt and transform their sounds by connecting multiple processors, forming an audio processing graph. However, most deep-learning methods overlook this real-world practice and assume fixed graph settings. To…
Sequential scientific data span many resolutions and domains, and unifying them into a common representation is a key step toward developing foundation models for the sciences. Astronomical spectra exemplify this challenge: massive surveys…