English
Related papers

Related papers: Virtual acoustics in inhomogeneous media with sing…

200 papers

We present a single-stage casual waveform-to-waveform multichannel model that can separate moving sound sources based on their broad spatial locations in a dynamic acoustic scene. We divide the scene into two spatial regions containing,…

Sound · Computer Science 2022-07-01 Dejan Markovic , Alexandre Defossez , Alexander Richard

An inverse scattering problem for the 3D acoustic equation in time domain is considered. The unknown spatially distributed speed of sound is the subject of the solution of this problem. A single location of the point source is used. Using a…

Mathematical Physics · Physics 2019-01-01 Michael V. Klibanov , Jingzhi Li , Wenlong Zhang

We tackle the multi-party speech recovery problem through modeling the acoustic of the reverberant chambers. Our approach exploits structured sparsity models to perform room modeling and speech recovery. We propose a scheme for…

Machine Learning · Computer Science 2012-10-26 Afsaneh Asaei , Mohammad Golbabaee , Hervé Bourlard , Volkan Cevher

We aim to monitor and characterize signals in the subsurface by combining these passive signals with recorded reflection data at the surface of the Earth. To achieve this, we propose a method to create virtual receivers from reflection data…

Geophysics · Physics 2023-08-15 Joeri Brackenhoff , Jan Thorbecke , Kees Wapenaar

This paper introduces an area-based source separation method designed for virtual meeting scenarios. The aim is to preserve speech signals from an unspecified number of sources within a defined spatial area in front of a linear microphone…

Audio and Speech Processing · Electrical Eng. & Systems 2024-08-20 Martin Strauss , Okan Köpüklü

In recent years, metamaterials have gained considerable attention as a promising material technology due to their unique properties and customizable design, distinguishing them from traditional materials. This article delves into the value…

Applied Physics · Physics 2024-04-09 Sichao Qu , Min Yang , Yunfei Xu , Songwen Xiao , Nicholas X. Fang

We introduce SoundSpaces 2.0, a platform for on-the-fly geometry-based audio rendering for 3D environments. Given a 3D mesh of a real-world environment, SoundSpaces can generate highly realistic acoustics for arbitrary sounds captured from…

Audio-visual navigation task requires an agent to find a sound source in a realistic, unmapped 3D environment by utilizing egocentric audio-visual observations. Existing audio-visual navigation works assume a clean environment that solely…

Sound · Computer Science 2022-02-23 Yinfeng Yu , Wenbing Huang , Fuchun Sun , Changan Chen , Yikai Wang , Xiaohong Liu

The thud of a bouncing ball, the onset of speech as lips open -- when visual and audio events occur together, it suggests that there might be a common, underlying event that produced both signals. In this paper, we argue that the visual and…

Computer Vision and Pattern Recognition · Computer Science 2018-10-10 Andrew Owens , Alexei A. Efros

We introduce a new approach for audio-visual speech separation. Given a video, the goal is to extract the speech associated with a face in spite of simultaneous background sounds and/or other human speakers. Whereas existing methods focus…

Computer Vision and Pattern Recognition · Computer Science 2021-04-07 Ruohan Gao , Kristen Grauman

The aim of the study is to demonstrate that some methods are more relevant for implementing the Real-Time Nearfield Acoustic Holography than others. First by focusing on the forward propagation problem, different approaches are compared to…

Classical Physics · Physics 2008-10-13 Jean-Hugh Thomas , Vincent Grulier , Sébastien Paillasseur , Jean-Claude Pascal

Time reversal methods are widely used to achieve wave focusing in acoustics and electromagnetics. A typical time reversal experiment requires that a transmitter be initially present at the target focusing point, which limits the application…

Chaotic Dynamics · Physics 2016-05-11 Bo Xiao , Thomas M. Antonsen , Edward Ott , Steven M. Anlage

In Extended Reality (XR), complex acoustic environments often overwhelm users, compromising both scene awareness and social engagement due to entangled sound sources. We introduce MoXaRt, a real-time XR system that uses audio-visual cues to…

This paper studies audio-visual noise suppression for egocentric videos -- where the speaker is not captured in the video. Instead, potential noise sources are visible on screen with the camera emulating the off-screen speaker's view of the…

Sound · Computer Science 2023-05-04 Roshan Sharma , Weipeng He , Ju Lin , Egor Lakomkin , Yang Liu , Kaustubh Kalgaonkar

Voice conversion models have developed for decades, and current mainstream research focuses on non-streaming voice conversion. However, streaming voice conversion is more suitable for practical application scenarios than non-streaming voice…

Sound · Computer Science 2022-06-16 Ziyi Chen , Haoran Miao , Pengyuan Zhang

This paper concerns time-harmonic inverse source problems with a single far-field pattern in two dimensions, where the source term is compactly supported in an a priori given inhomogeneous background medium. For convex-polygonal source…

Analysis of PDEs · Mathematics 2020-03-13 Guanghui Hu , Jingzhi Li

An audiovisual speaker conversion method is presented for simultaneously transforming the facial expressions and voice of a source speaker into those of a target speaker. Transforming the facial and acoustic features together makes it…

Audio and Speech Processing · Electrical Eng. & Systems 2018-12-04 Fuming Fang , Xin Wang , Junichi Yamagishi , Isao Echizen

Recently, with the advancement of AIGC, deep learning-based video-to-audio (V2A) technology has garnered significant attention. However, existing research mostly focuses on mono audio generation that lacks spatial perception, while the…

Sound · Computer Science 2025-08-22 Lei Zhao , Rujin Chen , Chi Zhang , Xiao-Lei Zhang , Xuelong Li

In non-destructive and biomedical imaging, spatial patterns inside a sample are imaged without destroying it. Therefore, propagating waves, including electromagnetic or ultrasonic signals, or even diffuse heat are generated or modified by…

Applied Physics · Physics 2025-07-09 Peter Burgholzer , Lukas Gahleitner , Guenther Mayr

Conventional approaches to sound localization and separation are based on microphone arrays in artificial systems. Inspired by the selective perception of human auditory system, we design a multi-source listening system which can separate…

Sound · Computer Science 2019-11-11 Xuecong Sun , Han Jia , Zhe Zhang , Yuzhen Yang , Zhaoyong Sun , Jun Yang