English
Related papers

Related papers: Gabor frames and deep scattering networks in audio…

200 papers

While deep neural networks have facilitated significant advancements in the field of speech enhancement, most existing methods are developed following either empirical or relatively blind criteria, lacking adequate guidelines in pipeline…

Sound · Computer Science 2023-03-29 Andong Li , Guochen Yu , Chengshi Zheng , Wenzhe Liu , Xiaodong Li

Light scattering within scattering media presents a substantial obstacle to optical transmission. A speckle pattern with random amplitude and phase distribution is observed when coherent light travels through strong scattering media.…

Optics · Physics 2022-09-01 Hui Liu , Xiangyu Zhu , Xiaoxue Zhang , Xudong Chen , Zhili Lin

Existing audio-text retrieval (ATR) methods are essentially discriminative models that aim to maximize the conditional likelihood, represented as p(candidates|query). Nevertheless, this methodology fails to consider the intrinsic data…

Sound · Computer Science 2024-10-18 Yifei Xin , Xuxin Cheng , Zhihong Zhu , Xusheng Yang , Yuexian Zou

We present a novel hybrid sound propagation algorithm for interactive applications. Our approach is designed for dynamic scenes and uses a neural network-based learned scattered field representation along with ray tracing to generate…

Sound · Computer Science 2021-09-28 Zhenyu Tang , Hsien-Yu Meng , Dinesh Manocha

We explore frame-level audio feature learning for chord recognition using artificial neural networks. We present the argument that chroma vectors potentially hold enough information to model harmonic content of audio for chord recognition,…

Sound · Computer Science 2016-12-16 Filip Korzeniowski , Gerhard Widmer

In this expository note we present an introduction to the Gabor wave front set. As is often the case, this tool in microlocal analysis has been introduced and reinvented in different forms which turn out to be equivalent or intimately…

Classical Analysis and ODEs · Mathematics 2020-04-06 Luigi Rodino , S. Ivan Trapasso

We consider sparseness properties of adaptive time-frequency representations obtained using nonstationary Gabor frames (NSGFs). NSGFs generalize classical Gabor frames by allowing for adaptivity in either time or frequency. It is known that…

Functional Analysis · Mathematics 2018-01-03 Emil Solsbæk Ottosen , Morten Nielsen

We describe some recent advances in the numerical solution of acoustic scattering problems. A major focus of the paper is the efficient solution of high frequency scattering problems via hybrid numerical-asymptotic boundary element methods.…

Numerical Analysis · Mathematics 2014-10-23 Simon N. Chandler-Wilde , Stephen Langdon

Contemporary speech enhancement predominantly relies on audio transforms that are trained to reconstruct a clean speech waveform. The development of high-performing neural network sound recognition systems has raised the possibility of…

Audio and Speech Processing · Electrical Eng. & Systems 2025-11-18 Mark R. Saddler , Andrew Francl , Jenelle Feather , Kaizhi Qian , Yang Zhang , Josh H. McDermott

We present a new encoder-decoder Vision Transformer architecture, Patcher, for medical image segmentation. Unlike standard Vision Transformers, it employs Patcher blocks that segment an image into large patches, each of which is further…

Image and Video Processing · Electrical Eng. & Systems 2023-05-31 Yanglan Ou , Ye Yuan , Xiaolei Huang , Stephen T. C. Wong , John Volpi , James Z. Wang , Kelvin Wong

The graph Hilbert transform (GHT) is a key tool in constructing analytic signals and extracting envelope and phase information in graph signal processing. However, its utility is limited by confinement to the graph Fourier domain, a fixed…

Signal Processing · Electrical Eng. & Systems 2025-09-23 Daxiang Li , Zhichao Zhang

We obtain Gabor frame characterisations of modulation spaces defined via a class of translation-modulation invariant Banach spaces of distributions that was recently introduced in $[10]$. We show that these spaces admit an atomic…

Functional Analysis · Mathematics 2021-02-08 Andreas Debrouwere , Bojan Prangoski

Current theoretical treatment of mode splitting and scattering loss resulting from sub-wavelength scatterers attached to the surface of high-quality-factor whispering-gallery-mode microresonators is not satisfactory. Different models have…

Optics · Physics 2015-06-05 Qing Li , Ali A. Eftekhar , Zhixuan Xia , Ali Adibi

This paper presents a novel approach for computing substructure characteristic modes. This method leverages electromagnetic scattering matrices and spherical wave expansion to directly decompose electromagnetic fields. Unlike conventional…

Classical Physics · Physics 2024-09-13 Chenbo Shi , Jin Pan , Xin Gu , Shichen Liang , Le Zuo

As wireless networks transition toward 6G, high mobility, clustered scattering, and hardware impairments increasingly challenge classical assumptions on channel sparsity, resolvability, and stationarity. In these regimes, performance…

Signal Processing · Electrical Eng. & Systems 2026-05-05 Hamza Haif , Abdelali Arous , Huseyin Arslan

Most of the research on data-driven speech representation learning has focused on raw audios in an end-to-end manner, paying little attention to their internal phonological or gestural structure. This work, investigating the speech…

Audio and Speech Processing · Electrical Eng. & Systems 2022-06-22 Jiachen Lian , Alan W Black , Louis Goldstein , Gopala Krishna Anumanchipalli

Complex spectrum and magnitude are considered as two major features of speech enhancement and dereverberation. Traditional approaches always treat these two features separately, ignoring their underlying relationship. In this paper, we…

Audio and Speech Processing · Electrical Eng. & Systems 2022-05-06 Yihui Fu , Yun Liu , Jingdong Li , Dawei Luo , Shubo Lv , Yukai Jv , Lei Xie

With the emergence of GAN-based vocoders, the discriminator, as a crucial component, has been developed recently. In our work, we focus on improving the time-frequency based discriminator. Particularly, Short-Time Fourier Transform (STFT)…

Audio and Speech Processing · Electrical Eng. & Systems 2025-12-04 Nan Xu , Zhaolong Huang , Xiao Zeng

We focus on automatic feature extraction for raw audio heartbeat sounds, aimed at anomaly detection applications in healthcare. We learn features with the help of an autoencoder composed by a 1D non-causal convolutional encoder and a…

Sound · Computer Science 2021-02-25 Robert-George Colt , Csongor-Huba Várady , Riccardo Volpi , Luigi Malagò

Background: Windowed Fourier decompositions (WFD) are widely used in measuring stationary and non-stationary spectral phenomena and in describing pairwise relationships among multiple signals. Although a variety of WFDs see frequent…

Quantitative Methods · Quantitative Biology 2019-01-30 Christopher K. Kovach , Phillip E. Gander
‹ Prev 1 8 9 10 Next ›