中文
相关论文

相关论文: Fully Reversing the Shoebox Image Source Method: F…

200 篇论文

For robust visual-inertial SLAM in perceptually-challenging indoor environments,recent studies exploit line features to extract descriptive information about scene structure to deal with the degeneracy of point features. But existing…

计算机视觉与模式识别 · 计算机科学 2024-07-02 Wanting Li , Shuo Wang , Yongcai Wang , Yu Shao , Xuewei Bai , Deying Li

A new impulse response (IR) dataset called "MeshRIR" is introduced. Currently available datasets usually include IRs at an array of microphones from several source positions under various room conditions, which are basically designed for…

音频与语音处理 · 电气工程与系统科学 2021-07-26 Shoichi Koyama , Tomoya Nishida , Keisuke Kimura , Takumi Abe , Natsuki Ueno , Jesper Brunnström

Seismic waves are the most sensitive probe of the Earth's interior we have. With the dense data sets available in exploration, images of subsurface structures can be obtained through processes such as migration. Unfortunately, relating…

地球物理 · 物理学 2009-05-05 R. B. Schlottmann

Single-channel speech dereverberation aims at extracting a dry speech signal from a recording affected by the acoustic reflections in a room. However, most current deep learning-based approaches for speech dereverberation are not…

声音 · 计算机科学 2024-07-12 Louis Bahrman , Mathieu Fontaine , Jonathan Le Roux , Gaël Richard

Acoustic Environment Matching (AEM) is the task of transferring clean audio into a target acoustic environment, enabling engaging applications such as audio dubbing and auditory immersive virtual reality (VR). Recovering similar room…

声音 · 计算机科学 2026-04-01 Chenpei Huang , Lingfeng Yao , Kyu In Lee , Lan Emily Zhang , Xun Chen , Miao Pan

Reconstructing unknown external source functions is an important perception capability for a large range of robotics domains including manipulation, aerial, and underwater robotics. In this work, we propose a Physics-Informed Neural Network…

机器人学 · 计算机科学 2024-11-05 Youngsun Wi , Jayjun Lee , Miquel Oller , Nima Fazeli

Spike camera is a new type of bio-inspired vision sensor that records light intensity in the form of a spike array with high temporal resolution (20,000 Hz). This new paradigm of vision sensor offers significant advantages for many vision…

计算机视觉与模式识别 · 计算机科学 2023-08-08 Lin Zhu , Yunlong Zheng , Mengyue Geng , Lizhi Wang , Hua Huang

In this paper, we propose a new algorithm that efficiently separates a directional source and diffuse background noise based on independent low-rank matrix analysis (ILRMA). ILRMA is one of the state-of-the-art techniques of blind source…

声音 · 计算机科学 2019-06-19 Yuki Kubo , Norihiro Takamune , Daichi Kitamura , Hiroshi Saruwatari

We study the problem of simultaneously reconstructing a polygonal room and a trajectory of a device equipped with a (nearly) collocated omnidirectional source and receiver. The device measures arrival times of echoes of pulses emitted by…

机器人学 · 计算机科学 2016-09-01 Miranda Krekovic , Ivan Dokmanic , Martin Vetterli

An inverse source reconstruction (ISR) based 3-D near-field (NF) passive radar microwave imaging method utilizing modulated signals is presented. The modulated signals from a non-cooperative transmitter are scattered by the targets of…

图像与视频处理 · 电气工程与系统科学 2026-05-06 Quanfeng Wang , Alexander H. Paulus , Thomas F. Eibert

The room impulse response (RIR) encodes, among others, information about the distance of an acoustic source from the sensors. Deep neural networks (DNNs) have been shown to be able to extract that information for acoustic distance…

声音 · 计算机科学 2024-08-27 Tobias Gburrek , Adrian Meise , Joerg Schmalenstroeer , Reinhold Haeb-Umbach

Unsupervised physical parameter estimation from video lacks a common benchmark: existing methods evaluate on non-overlapping synthetic data, the sole real-world dataset is restricted to single-body systems, and no established protocol…

计算机视觉与模式识别 · 计算机科学 2026-03-23 Rasul Khanbayov , Mohamed Rayan Barhdadi , Erchin Serpedin , Hasan Kurban

We develop a novel iterative direct sampling method (IDSM) for solving linear or nonlinear elliptic inverse problems with partial Cauchy data. It integrates three innovations: a data completion scheme to reconstruct missing boundary…

数值分析 · 数学 2025-11-12 Bangti Jin , Fengru Wang , Jun Zou

Accurate estimation of Room Impulse Response (RIR), which captures an environment's acoustic properties, is important for speech processing and AR/VR applications. We propose AV-RIR, a novel multi-modal multi-task learning approach to…

声音 · 计算机科学 2024-04-25 Anton Ratnarajah , Sreyan Ghosh , Sonal Kumar , Purva Chiniya , Dinesh Manocha

Learning-based methods have become ubiquitous in speaker localization. Existing systems rely on simulated training sets for the lack of sufficiently large, diverse and annotated real datasets. Most room acoustics simulators used for this…

声音 · 计算机科学 2023-05-26 Prerak Srivastava , Antoine Deleforge , Archontis Politis , Emmanuel Vincent

The estimation of room impulse responses (RIRs) between static loudspeaker and microphone locations can be done using a number of well-established measurement and inference procedures. While these procedures assume a time-invariant acoustic…

音频与语音处理 · 电气工程与系统科学 2024-11-14 Kathleen MacWilliam , Thomas Dietzen , Randall Ali , Toon van Waterschoot

For audio in augmented reality (AR), knowledge of the users' real acoustic environment is crucial for rendering virtual sounds that seamlessly blend into the environment. As acoustic measurements are usually not feasible in practical AR…

声音 · 计算机科学 2024-09-24 Francesc Lluís , Nils Meyer-Kahlen

Diffusion models are widely used in applications ranging from image generation to inverse problems. However, training diffusion models typically requires clean ground-truth images, which are unavailable in many applications. We introduce…

图像与视频处理 · 电气工程与系统科学 2025-05-20 Chicago Y. Park , Shirin Shoushtari , Hongyu An , Ulugbek S. Kamilov

The generation of room impulse responses (RIRs) using deep neural networks has attracted growing research interest due to its applications in virtual and augmented reality, audio postproduction, and related fields. Most existing approaches…

声音 · 计算机科学 2025-07-17 Silvia Arellano , Chunghsin Yeh , Gautam Bhattacharya , Daniel Arteaga

We propose a novel method called the Relevance Subject Machine (RSM) to solve the person re-identification (re-id) problem. RSM falls under the category of Bayesian sparse recovery algorithms and uses the sparse representation of the input…

计算机视觉与模式识别 · 计算机科学 2017-04-03 Igor Fedorov , Ritwik Giri , Bhaskar D. Rao , Truong Q. Nguyen