中文
相关论文

相关论文: Log Complex Color for Visual Pattern Recognition o…

200 篇论文

Capturing high-frequency data concerning the condition of complex systems, e.g. by acoustic monitoring, has become increasingly prevalent. Such high-frequency signals typically contain time dependencies ranging over different time scales…

声音 · 计算机科学 2022-06-14 Gaetan Frusque , Olga Fink

Learning features from data has shown to be more successful than using hand-crafted features for many machine learning tasks. In music information retrieval (MIR), features learned from windowed spectrograms are highly variant to…

声音 · 计算机科学 2019-07-16 Stefan Lattner , Monika Dörfler , Andreas Arzt

Time-frequency representations of audio signals often resemble texture images. This paper derives a simple audio classification algorithm based on treating sound spectrograms as texture images. The algorithm is inspired by an earlier visual…

计算机视觉与模式识别 · 计算机科学 2008-09-29 Guoshen Yu , Jean-Jacques Slotine

Several Scientific and engineering applications require merging of sampled images for complex perception development. In most cases, for such requirements, images are merged at intensity level. Even though it gives fairly good perception of…

计算机视觉与模式识别 · 计算机科学 2014-08-01 T. R. Gopalakrishnan Nair , Richa Sharma

Phase retrieval refers to algorithmic methods for recovering a signal from its phaseless measurements. Local search algorithms that work directly on the non-convex formulation of the problem have been very popular recently. Due to the…

信息论 · 计算机科学 2020-03-06 Rishabh Dudeja , Milad Bakhshizadeh , Junjie Ma , Arian Maleki

Temporal imaging systems are outstanding tools for single-shot observation of optical signals that have irregular and ultrafast dynamics. They allow long time windows to be recorded with femtosecond resolution, and do not rely on complex…

We learn audio representations by solving a novel self-supervised learning task, which consists of predicting the phase of the short-time Fourier transform from its magnitude. A convolutional encoder is used to map the magnitude spectrum of…

音频与语音处理 · 电气工程与系统科学 2019-10-29 Félix de Chaumont Quitry , Marco Tagliasacchi , Dominik Roblek

In the field of deepfake detection, previous studies focus on using reconstruction or mask and prediction methods to train pre-trained models, which are then transferred to fake audio detection training where the encoder is used to extract…

Signal analysis and classification is fraught with high levels of noise and perturbation. Computer-vision-based deep learning models applied to spectrograms have proven useful in the field of signal classification and detection; however,…

机器学习 · 计算机科学 2024-09-04 Joel Brogan , Olivera Kotevska , Anibely Torres , Sumit Jha , Mark Adams

In this study, we experimentally investigate the application of a transient signal with complex frequencies to the absorption and transmission of sound waves. Indeed, the emission of a wave with an exponentially varying amplitude in time is…

应用物理 · 物理学 2025-03-18 Anis Maddi , Gaelle Poignand , Vassos Achilleos , Vincent Pagneux , Guillaume Penelet

Speech separation has been very successful with deep learning techniques. Substantial effort has been reported based on approaches over spectrogram, which is well known as the standard time-and-frequency cross-domain representation for…

声音 · 计算机科学 2019-04-17 Gene-Ping Yang , Chao-I Tuan , Hung-Yi Lee , Lin-shan Lee

Audio captioning aims to generate text descriptions of audio clips. In the real world, many objects produce similar sounds. How to accurately recognize ambiguous sounds is a major challenge for audio captioning. In this work, inspired by…

音频与语音处理 · 电气工程与系统科学 2023-05-30 Xubo Liu , Qiushi Huang , Xinhao Mei , Haohe Liu , Qiuqiang Kong , Jianyuan Sun , Shengchen Li , Tom Ko , Yu Zhang , Lilian H. Tang , Mark D. Plumbley , Volkan Kılıç , Wenwu Wang

We present an algorithm for coherent diffractive imaging with phaseless measurements. It treats the forward model as a combination of coherent and incoherent waves. The algorithm reconstructs absorption and phase contrast that quantifies…

计算物理 · 物理学 2022-08-25 Miguel Moscoso , Alexei Novikov , George Papanicolaou , Chrysoula Tsogka

We present a system for measuring the amplitude and phase profiles of the pressure field of a harmonic acoustic wave with the goal of reconstructing the volumetric sound field. Unlike optical holograms that cannot be reconstructed exactly…

软凝聚态物质 · 物理学 2018-08-09 Hillary W. Gao , Kimberly I. Mishra , Annemarie Winters , Sidney Wolin , David G. Grier

We show that an intensity speckle can be directly interpreted as the properties of incident light - amplitude, phase, polarization, and coherency over spatial positions. Revisiting the speckle-correlation scattering matrix (SSM) method [Lee…

光学 · 物理学 2019-08-07 KyeoReh Lee , YongKeun Park

Neural vocoders are central to speech synthesis; despite their success, most still suffer from limited prosody modeling and inaccurate phase reconstruction. We propose a vocoder that introduces prosody-guided harmonic attention to enhance…

声音 · 计算机科学 2026-01-22 Mohammed Salah Al-Radhi , Riad Larbi , Mátyás Bartalis , Géza Németh

Speckle noise is an inherent disturbance in coherent imaging systems such as digital holography, synthetic aperture radar, optical coherence tomography, or ultrasound systems. These systems usually produce only single observation per view…

图像与视频处理 · 电气工程与系统科学 2022-05-19 Tsung-Ming Tai , Yun-Jie Jhang , Wen-Jyi Hwang , Chau-Jern Cheng

Self-supervised pre-training models have been used successfully in several machine learning domains. However, only a tiny amount of work is related to music. In our work, we treat a spectrogram of music as a series of patches and design a…

声音 · 计算机科学 2022-10-31 Leyi Zhao , Yi Li

Respiratory sound contains crucial information for the early diagnosis of fatal lung diseases. Since the COVID-19 pandemic, there has been a growing interest in contact-free medical care based on electronic stethoscopes. To this end,…

音频与语音处理 · 电气工程与系统科学 2024-12-30 Sangmin Bae , June-Woo Kim , Won-Yang Cho , Hyerim Baek , Soyoun Son , Byungjo Lee , Changwan Ha , Kyongpil Tae , Sungnyun Kim , Se-Young Yun

Previous audio generation mainly focuses on specified sound classes such as speech or music, whose form and content are greatly restricted. In this paper, we go beyond specific audio generation by using natural language description as a…

声音 · 计算机科学 2023-05-04 Guangwei Li , Xuenan Xu , Lingfeng Dai , Mengyue Wu , Kai Yu