Related papers: Virtual acoustics in inhomogeneous media with sing…
When a signal is emitted from a source, recorded by an array of transducers, time reversed and re-emitted into the medium, it will refocus approximately on the source location. We analyze the refocusing resolution in a high frequency,…
Single-channel audio separation aims to separate individual sources from a single-channel mixture. Most existing methods rely on supervised learning with synthetically generated paired data. However, obtaining high-quality paired data in…
Given an input sound signal and a target virtual sound source, sound spatialisation algorithms manipulate the signal so that a listener perceives it as though it were emitted from the target source. There exist several established…
The objective of this work is to extract target speaker's voice from a mixture of voices using visual cues. Existing works on audio-visual speech separation have demonstrated their performance with promising intelligibility, but maintaining…
We address the challenge of sound propagation simulations in 3D virtual rooms with moving sources, which have applications in virtual/augmented reality, game audio, and spatial computing. Solutions to the wave equation can describe wave…
This paper proposes a real-time system integrating an acoustic material estimation from visual appearance and an on-the-fly mapping in the 3-dimension. The proposed method estimates the acoustic materials of surroundings in indoor scenes…
We analyse the ultrasound waves reflected by multiple bubbles in the linearized time-dependent acoustic model. The generated time-dependent wave field is estimated close to the bubbles. The motivation of this study comes from the therapy…
When waves propagate through a complex or heterogeneous medium the wave field is corrupted by the heterogeneities. Such corruption limits the performance of imaging or communication schemes. One may then ask the question: is there an…
In this paper, we consider inverse time-harmonic acoustic and electromagnetic scattering from locally perturbed rough surfaces in three dimensions. The scattering interface is supposed to be the graph of a Lipschitz continuous function with…
Range estimation of a far field sound source in a reverberant environment is known to be a notoriously difficult problem, hence most localization methods are only capable of estimating the source's Direction-of-Arrival (DoA). In an earlier…
How does audio describe the world around us? In this paper, we propose a method for generating an image of a scene from sound. Our method addresses the challenges of dealing with the large gaps that often exist between sight and sound. We…
An immersive acoustic experience enabled by spatial audio is just as crucial as the visual aspect in creating realistic virtual environments. However, existing methods for room impulse response estimation rely either on data-demanding…
Contrary to geometric acoustics-based simulations where the spatial information is available in a tangible form, it is not straightforward to auralize wave-based simulations. A variety of methods have been proposed that compute the ear…
Light and sound waves have the fascinating property that they can move objects through the transfer of linear or angular momentum. This ability has led to the development of optical and acoustic tweezers, with applications ranging from…
Inviscid hydrodynamics mediates forces through pressure and other, typically irrotational, external forces. Acoustically induced forces must be consistent with arising from such a pressure field. The use of "acoustic stress" is shown to…
Binaural audio gives the listener an immersive experience and can enhance augmented and virtual reality. However, recording binaural audio requires specialized setup with a dummy human head having microphones in left and right ears. Such a…
This investigation is concerned with the 2D acoustic scattering problem of a plane wave propagating in a non-lossy fluid host and soliciting a linear, isotropic, macroscopically-homogeneous, lossy, flat-plane layer in which the mass density…
Attention is a powerful concept in computer vision. End-to-end networks that learn to focus selectively on regions of an image or video often perform strongly. However, other image regions, while not necessarily containing the signal of…
Video and audio content creation serves as the core technique for the movie industry and professional users. Recently, existing diffusion-based methods tackle video and audio generation separately, which hinders the technique transfer from…
Audio-visual sound source localization task aims to spatially localize sound-making objects within visual scenes by integrating visual and audio cues. However, existing methods struggle with accurately localizing sound-making objects in…