English
Related papers

Related papers: Log Complex Color for Visual Pattern Recognition o…

200 papers

Capturing high-frequency data concerning the condition of complex systems, e.g. by acoustic monitoring, has become increasingly prevalent. Such high-frequency signals typically contain time dependencies ranging over different time scales…

Sound · Computer Science 2022-06-14 Gaetan Frusque , Olga Fink

Learning features from data has shown to be more successful than using hand-crafted features for many machine learning tasks. In music information retrieval (MIR), features learned from windowed spectrograms are highly variant to…

Sound · Computer Science 2019-07-16 Stefan Lattner , Monika Dörfler , Andreas Arzt

Time-frequency representations of audio signals often resemble texture images. This paper derives a simple audio classification algorithm based on treating sound spectrograms as texture images. The algorithm is inspired by an earlier visual…

Computer Vision and Pattern Recognition · Computer Science 2008-09-29 Guoshen Yu , Jean-Jacques Slotine

Several Scientific and engineering applications require merging of sampled images for complex perception development. In most cases, for such requirements, images are merged at intensity level. Even though it gives fairly good perception of…

Computer Vision and Pattern Recognition · Computer Science 2014-08-01 T. R. Gopalakrishnan Nair , Richa Sharma

Phase retrieval refers to algorithmic methods for recovering a signal from its phaseless measurements. Local search algorithms that work directly on the non-convex formulation of the problem have been very popular recently. Due to the…

Information Theory · Computer Science 2020-03-06 Rishabh Dudeja , Milad Bakhshizadeh , Junjie Ma , Arian Maleki

Temporal imaging systems are outstanding tools for single-shot observation of optical signals that have irregular and ultrafast dynamics. They allow long time windows to be recorded with femtosecond resolution, and do not rely on complex…

We learn audio representations by solving a novel self-supervised learning task, which consists of predicting the phase of the short-time Fourier transform from its magnitude. A convolutional encoder is used to map the magnitude spectrum of…

Audio and Speech Processing · Electrical Eng. & Systems 2019-10-29 Félix de Chaumont Quitry , Marco Tagliasacchi , Dominik Roblek

In the field of deepfake detection, previous studies focus on using reconstruction or mask and prediction methods to train pre-trained models, which are then transferred to fake audio detection training where the encoder is used to extract…

Signal analysis and classification is fraught with high levels of noise and perturbation. Computer-vision-based deep learning models applied to spectrograms have proven useful in the field of signal classification and detection; however,…

Machine Learning · Computer Science 2024-09-04 Joel Brogan , Olivera Kotevska , Anibely Torres , Sumit Jha , Mark Adams

In this study, we experimentally investigate the application of a transient signal with complex frequencies to the absorption and transmission of sound waves. Indeed, the emission of a wave with an exponentially varying amplitude in time is…

Applied Physics · Physics 2025-03-18 Anis Maddi , Gaelle Poignand , Vassos Achilleos , Vincent Pagneux , Guillaume Penelet

Speech separation has been very successful with deep learning techniques. Substantial effort has been reported based on approaches over spectrogram, which is well known as the standard time-and-frequency cross-domain representation for…

Sound · Computer Science 2019-04-17 Gene-Ping Yang , Chao-I Tuan , Hung-Yi Lee , Lin-shan Lee

Audio captioning aims to generate text descriptions of audio clips. In the real world, many objects produce similar sounds. How to accurately recognize ambiguous sounds is a major challenge for audio captioning. In this work, inspired by…

Audio and Speech Processing · Electrical Eng. & Systems 2023-05-30 Xubo Liu , Qiushi Huang , Xinhao Mei , Haohe Liu , Qiuqiang Kong , Jianyuan Sun , Shengchen Li , Tom Ko , Yu Zhang , Lilian H. Tang , Mark D. Plumbley , Volkan Kılıç , Wenwu Wang

We present an algorithm for coherent diffractive imaging with phaseless measurements. It treats the forward model as a combination of coherent and incoherent waves. The algorithm reconstructs absorption and phase contrast that quantifies…

Computational Physics · Physics 2022-08-25 Miguel Moscoso , Alexei Novikov , George Papanicolaou , Chrysoula Tsogka

We present a system for measuring the amplitude and phase profiles of the pressure field of a harmonic acoustic wave with the goal of reconstructing the volumetric sound field. Unlike optical holograms that cannot be reconstructed exactly…

Soft Condensed Matter · Physics 2018-08-09 Hillary W. Gao , Kimberly I. Mishra , Annemarie Winters , Sidney Wolin , David G. Grier

We show that an intensity speckle can be directly interpreted as the properties of incident light - amplitude, phase, polarization, and coherency over spatial positions. Revisiting the speckle-correlation scattering matrix (SSM) method [Lee…

Optics · Physics 2019-08-07 KyeoReh Lee , YongKeun Park

Neural vocoders are central to speech synthesis; despite their success, most still suffer from limited prosody modeling and inaccurate phase reconstruction. We propose a vocoder that introduces prosody-guided harmonic attention to enhance…

Sound · Computer Science 2026-01-22 Mohammed Salah Al-Radhi , Riad Larbi , Mátyás Bartalis , Géza Németh

Speckle noise is an inherent disturbance in coherent imaging systems such as digital holography, synthetic aperture radar, optical coherence tomography, or ultrasound systems. These systems usually produce only single observation per view…

Image and Video Processing · Electrical Eng. & Systems 2022-05-19 Tsung-Ming Tai , Yun-Jie Jhang , Wen-Jyi Hwang , Chau-Jern Cheng

Self-supervised pre-training models have been used successfully in several machine learning domains. However, only a tiny amount of work is related to music. In our work, we treat a spectrogram of music as a series of patches and design a…

Sound · Computer Science 2022-10-31 Leyi Zhao , Yi Li

Respiratory sound contains crucial information for the early diagnosis of fatal lung diseases. Since the COVID-19 pandemic, there has been a growing interest in contact-free medical care based on electronic stethoscopes. To this end,…

Audio and Speech Processing · Electrical Eng. & Systems 2024-12-30 Sangmin Bae , June-Woo Kim , Won-Yang Cho , Hyerim Baek , Soyoun Son , Byungjo Lee , Changwan Ha , Kyongpil Tae , Sungnyun Kim , Se-Young Yun

Previous audio generation mainly focuses on specified sound classes such as speech or music, whose form and content are greatly restricted. In this paper, we go beyond specific audio generation by using natural language description as a…

Sound · Computer Science 2023-05-04 Guangwei Li , Xuenan Xu , Lingfeng Dai , Mengyue Wu , Kai Yu
‹ Prev 1 4 5 6 7 8 10 Next ›