English
Related papers

Related papers: Image2Reverb: Cross-Modal Reverb Impulse Response …

200 papers

We introduce SoundVista, a method to generate the ambient sound of an arbitrary scene at novel viewpoints. Given a pre-acquired recording of the scene from sparsely distributed microphones, SoundVista can synthesize the sound of that scene…

Underground stations are a common communication situation in towns: we talk with friends or colleagues, listen to announcements or shop for titbits while background noise and reverberation are challenging communication. Here, we perform an…

Sound · Computer Science 2025-09-15 Ľuboš Hládek , Stephan D. Ewert , Bernhard U. Seeber

Simulation involves predicting responses of a physical system. In this article, we simulate opto-acoustic signals generated in a three-dimensional volume due to absorption of an optical pulse. A separable computational model is developed…

Computational Physics · Physics 2021-07-22 Jason Zalev , Michael C. Kolios

Realistic sound simulation plays a critical role in many applications. A key element in sound simulation is the room impulse response (RIR), which characterizes how sound propagates from a source to a listener within a given space. Recent…

Sound · Computer Science 2025-09-19 Chen Si , Qianyi Wu , Chaitanya Amballa , Romit Roy Choudhury

We investigate the benefit of combining blind audio recordings with 3D scene information for novel-view acoustic synthesis. Given audio recordings from 2-4 microphones and the 3D geometry and material of a scene containing multiple unknown…

Deep learning approaches have emerged that aim to transform an audio signal so that it sounds as if it was recorded in the same room as a reference recording, with applications both in audio post-production and augmented reality. In this…

Audio and Speech Processing · Electrical Eng. & Systems 2021-07-16 Christian J. Steinmetz , Vamsi Krishna Ithapu , Paul Calamia

Conditional text-to-image generation is an active area of research, with many possible applications. Existing research has primarily focused on generating a single image from available conditioning information in one step. One practical…

Computer Vision and Pattern Recognition · Computer Science 2019-09-24 Alaaeldin El-Nouby , Shikhar Sharma , Hannes Schulz , Devon Hjelm , Layla El Asri , Samira Ebrahimi Kahou , Yoshua Bengio , Graham W. Taylor

Training audio-to-image generative models requires an abundance of diverse audio-visual pairs that are semantically aligned. Such data is almost always curated from in-the-wild videos, given the cross-modal semantic correspondence that is…

Sound · Computer Science 2025-01-10 Darius Petermann , Mahdi M. Kalayeh

Blind acoustic parameter estimation consists in inferring the acoustic properties of an environment from recordings of unknown sound sources. Recent works in this area have utilized deep neural networks trained either partially or…

Sound · Computer Science 2022-07-20 Prerak Srivastava , Antoine Deleforge , Emmanuel Vincent

The inference of the absorption configuration of an existing room solely using acoustic signals can be challenging. This research presents two methods for estimating the room dimensions and frequency-dependent absorption coefficients using…

Sound · Computer Science 2023-04-26 Yuanxin Xia , Cheol-Ho Jeong

Room geometry inference (RGI) aims at estimating room shapes from measured room impulse responses (RIRs) and has received lots of attention for its importance in environment-aware audio rendering and virtual acoustic representation of a…

Audio and Speech Processing · Electrical Eng. & Systems 2024-01-22 Inmo Yeon , Jung-Woo Choi

The ability to accurately estimate room impulse responses (RIRs) is integral to many applications of spatial audio processing. Regrettably, estimating the RIR using ambient signals, such as speech or music, remains a challenging problem due…

Signal Processing · Electrical Eng. & Systems 2024-03-07 David Sundström , Anton Björkman , Andreas Jakobsson , Filip Elvander

In the next generations of cellular communication networks, higher density of base stations and higher frequency bands will be adopted. If being reflected by targets, the communication signal also brings information of the targets, in…

Signal Processing · Electrical Eng. & Systems 2020-11-18 Husheng Li

Can machines recording an audio-visual scene produce realistic, matching audio-visual experiences at novel positions and novel view directions? We answer it by studying a new task -- real-world audio-visual scene synthesis -- and a…

Computer Vision and Pattern Recognition · Computer Science 2023-10-17 Susan Liang , Chao Huang , Yapeng Tian , Anurag Kumar , Chenliang Xu

We introduce a computationally efficient and tunable feedback delay network (FDN) architecture for real-time room impulse response (RIR) rendering that addresses the computational and latency challenges inherent in traditional convolution…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-02 Armin Gerami , Ramani Duraiswami

Single-channel speech dereverberation aims at extracting a dry speech signal from a recording affected by the acoustic reflections in a room. However, most current deep learning-based approaches for speech dereverberation are not…

Sound · Computer Science 2024-07-12 Louis Bahrman , Mathieu Fontaine , Jonathan Le Roux , Gaël Richard

Geometrical acoustics is well suited for simulating room reverberation in interactive real-time applications. While the image source model (ISM) is exceptionally fast, the restriction to specular reflections impacts its perceptual…

Audio and Speech Processing · Electrical Eng. & Systems 2025-01-28 Stephan D. Ewert , Nico Gößling , Oliver Buttler , Steven van de Par , Hongmei Hu

We study imaging with an array of sensors that probes a medium with single frequency electromagnetic waves and records the scattered electric field. The medium is known and homogenous except for some small and penetrable inclusions. The…

Analysis of PDEs · Mathematics 2016-09-21 Liliana Borcea , Josselin Garnier

In this work, we introduce a novel framework which combines physics and machine learning methods to analyse acoustic signals. Three methods are developed for this task: a Bayesian inference approach for inferring the spectral acoustics…

Sound · Computer Science 2023-05-30 Yongchao Huang , Yuhang He , Hong Ge

Data-driven acoustic echo cancellation (AEC) methods, predominantly trained on synthetic or constrained real-world datasets, encounter performance declines in unseen echo scenarios, especially in real environments where echo paths are not…

Sound · Computer Science 2025-06-09 Fei Zhao , Shulin He , Xueliang Zhang