English
Related papers

Related papers: wav2shape: Hearing the Shape of a Drum Machine

200 papers

This paper concerns the inverse source scattering problems of recovering random sources for acoustic and elastic waves. The underlying sources are assumed to be random functions driven by an additive white noise. The inversion process aims…

Numerical Analysis · Mathematics 2024-12-10 Yan Chang , Yukun Guo , Zhipeng Yang , Yue Zhao

We describe and analyze algorithms for shape-constrained symbolic regression, which allows the inclusion of prior knowledge about the shape of the regression function. This is relevant in many areas of engineering -- in particular whenever…

Neural and Evolutionary Computing · Computer Science 2021-07-21 Christian Haider , Fabricio Olivetti de França , Bogdan Burlacu , Gabriel Kronberger

In this work, we consider the inverse electromagnetic scattering problem for a magneto-dielectric cylinder covering an impedance cylinder of arbitrary shape. We solve it by introducing a divide-and-conquer framework using specially designed…

Numerical Analysis · Mathematics 2025-11-27 Leonidas Mindrinos , Nikolaos Pallikarakis , Nikolaos L Tsitsas

Reconstructing the room transfer functions needed to calculate the complex sound field in a room has several important real-world applications. However, an unpractical number of microphones is often required. Recently, in addition to…

Audio and Speech Processing · Electrical Eng. & Systems 2024-06-12 Francesca Ronchini , Luca Comanducci , Mirco Pezzoli , Fabio Antonacci , Augusto Sarti

Audio processors whose parameters are modified periodically over time are often referred as time-varying or modulation based audio effects. Most existing methods for modeling these type of effect units are often optimized to a very specific…

Audio and Speech Processing · Electrical Eng. & Systems 2019-06-24 Marco A. Martínez Ramírez , Emmanouil Benetos , Joshua D. Reiss

Diffusion model shows remarkable potential on sparse-view computed tomography (SVCT) reconstruction. However, when a network is trained on a limited sample space, its generalization capability may be constrained, which degrades performance…

Image and Video Processing · Electrical Eng. & Systems 2025-11-11 Zekun Zhou , Tan Liu , Bing Yu , Yanru Gong , Liu Shi , Qiegen Liu

In this paper we address the problems of modeling the acoustic space generated by a full-spectrum sound source and of using the learned model for the localization and separation of multiple sources that simultaneously emit sparse-spectrum…

Sound · Computer Science 2015-02-06 Antoine Deleforge , Florence Forbes , Radu Horaud

Sound field reconstruction involves estimating sound fields from a limited number of spatially distributed observations. This work introduces a differentiable physics approach for sound field reconstruction, where the initial conditions of…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-07 Samuel A. Verburg , Efren Fernandez-Grande , Peter Gerstoft

Resonances are common in wave physics and their full and rigorous characterization is crucial to correctly tailor the response of a system in both time and frequency domains. However, they have been conventionally described by the quality…

Optics · Physics 2025-04-09 Isam Ben Soltane , Nicolas Bonod

We investigate the potential of stochastic neural networks for learning effective waveform-based acoustic models. The waveform-based setting, inherent to fully end-to-end speech recognition systems, is motivated by several comparative…

Machine Learning · Statistics 2021-08-17 Dino Oglic , Zoran Cvetkovic , Peter Sollich

We present a single-stage casual waveform-to-waveform multichannel model that can separate moving sound sources based on their broad spatial locations in a dynamic acoustic scene. We divide the scene into two spatial regions containing,…

Sound · Computer Science 2022-07-01 Dejan Markovic , Alexandre Defossez , Alexander Richard

Applications of deep learning to automatic multitrack mixing are largely unexplored. This is partly due to the limited available data, coupled with the fact that such data is relatively unstructured and variable. To address these…

Audio and Speech Processing · Electrical Eng. & Systems 2020-10-21 Christian J. Steinmetz , Jordi Pons , Santiago Pascual , Joan Serrà

Sound event detection systems typically consist of two stages: extracting hand-crafted features from the raw audio waveform, and learning a mapping between these features and the target sound events using a classifier. Recently, the focus…

Sound · Computer Science 2018-05-11 Emre Çakır , Tuomas Virtanen

The inverse acoustic scattering problems using multi-frequency backscattering far field patterns at isolated directions are studied. The underlying object could be point like scatterers, small scatterers, extended inhomogeneities and…

Analysis of PDEs · Mathematics 2019-10-18 Xia Ji , Xiaodong Liu

Conventionally, audio super-resolution models fixed the initial and the target sampling rates, which necessitate the model to be trained for each pair of sampling rates. We introduce NU-Wave 2, a diffusion model for neural audio upsampling…

Audio and Speech Processing · Electrical Eng. & Systems 2022-09-28 Seungu Han , Junhyeok Lee

We are concerned with the inverse scattering problem of extracting the geometric structures of an unknown/inaccessible inhomogeneous medium by using the corresponding acoustic far-field measurement. Using the intrinsic geometric properties…

Analysis of PDEs · Mathematics 2017-06-15 Jingzhi Li , Xiaofei Li , Hongyu Liu

The reconstruction mechanisms built by the human auditory system during sound reconstruction are still a matter of debate. The purpose of this study is to refine the auditory cortex model introduced in [9], and inspired by the geometrical…

Analysis of PDEs · Mathematics 2021-03-09 Rand Asswad , Ugo Boscain , Giuseppina Turco , Dario Prandi , Ludovic Sacchelli

Spurred by the potential of deep learning, computational music generation has gained renewed academic interest. A crucial issue in music generation is that of user control, especially in scenarios where the music generation process is…

Sound · Computer Science 2019-08-05 Stefan Lattner , Maarten Grachten

While deep generative models have become the leading methods for algorithmic composition, it remains a challenging problem to control the generation process because the latent variables of most deep-learning models lack good…

Sound · Computer Science 2020-08-18 Ziyu Wang , Dingsu Wang , Yixiao Zhang , Gus Xia

Most audio processing pipelines involve transformations that act on fixed-dimensional input representations of audio. For example, when using the Short Time Fourier Transform (STFT) the DFT size specifies a fixed dimension for the input…

Audio and Speech Processing · Electrical Eng. & Systems 2022-03-28 Krishna Subramani , Paris Smaragdis