中文
相关论文

相关论文: Audio-Visual Calibration with Polynomial Regressio…

200 篇论文

This paper introduces a modification of phase transform on singular value decomposition (SVD-PHAT) to localize multiple sound sources. This work aims to improve localization accuracy and keeps the algorithm complexity low for real-time…

音频与语音处理 · 电气工程与系统科学 2019-07-01 Francois Grondin , James Glass

This paper introduces a new localization method called SVD-PHAT. The SVD-PHAT method relies on Singular Value Decomposition of the SRP-PHAT projection matrix. A k-d tree is also proposed to speed up the search for the most likely direction…

音频与语音处理 · 电气工程与系统科学 2019-02-12 Francois Grondin , James Glass

This paper introduces a variant of the Singular Value Decomposition with Phase Transform (SVD-PHAT), named Difference SVD-PHAT (DSVD-PHAT), to achieve robust Sound Source Localization (SSL) in noisy conditions. Experiments are performed on…

音频与语音处理 · 电气工程与系统科学 2019-07-31 Francois Grondin , James Glass

We propose an efficient method to estimate source power spectral densities (PSDs) in a multi-source reverberant environment using a spherical microphone array. The proposed method utilizes the spatial correlation between the spherical…

声音 · 计算机科学 2018-05-21 Abdullah Fahim , Prasanga N. Samarasinghe , Thushara D. Abhayapala

Joint audio-visual speaker tracking requires that the locations of microphones and cameras are known and that they are given in a common coordinate system. Sensor self-localization algorithms, however, are usually separately developed for…

声音 · 计算机科学 2015-04-14 Florian Jacob , Reinhold Haeb-Umbach

Acoustic cameras have found many applications in practice. Accurate and reliable extrinsic calibration of the microphone array and visual sensors within acoustic cameras is crucial for fusing visual and auditory measurements. Existing…

机器人学 · 计算机科学 2025-02-11 Zhi Li , Jiang Wang , Xiaoyang Li , He Kong

In this paper, robust detection, tracking and geometry estimation methods are developed and combined into a system for estimating time-difference estimates, microphone localization and sound source movement. No assumptions on the 3D…

Neural approaches have shown a significant progress on camera-based reconstruction. But they require either a fairly dense sampling of the viewing sphere, or pre-training on an existing dataset, thereby limiting their generalizability. In…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Mohammed Brahimi , Bjoern Haefner , Zhenzhang Ye , Bastian Goldluecke , Daniel Cremers

Current 3D photoacoustic tomography (PAT) systems offer either high image quality or high frame rates but are not able to deliver high spatial and temporal resolution simultaneously, which limits their ability to image dynamic processes in…

Photoacoustic tomography (PAT) is a non-invasive imaging modality that requires recovering the initial data of the wave equation from certain measurements of the solution outside the object. In the standard PAT measurement setup, the used…

偏微分方程分析 · 数学 2021-12-07 Linh V. Nguyen , Markus Haltmeier , Richard Kowar , Ngoc Do

The reconstruction of a scene via a stereo-camera system is a two-steps process, where at first images from different cameras are matched to identify the set of point-to-point correspondences that then will actually be reconstructed in the…

计算机视觉与模式识别 · 计算机科学 2021-01-15 Riccardo Beschi , Xiao Feng , Stefania Melillo , Leonardo Parisi , Lorena Postiglione

This paper presents a method to reconstruct the 3D structure of generic convex rooms from sound signals. Differently from most of the previous approaches, the method is fully uncalibrated in the sense that no knowledge about the microphones…

声音 · 计算机科学 2016-06-21 Marco Crocco , Andrea Trucco , Alessio Del Bue

This paper introduces SMP-PHAT, which performs direction of arrival (DoA) of sound estimation with a microphone array by merging pairs of microphones that are parallel in space. This approach reduces the number of pairwise cross-correlation…

We tackle the problem of automatic calibration of radially distorted cameras in challenging conditions. Accurately determining distortion parameters typically requires either 1) solving the full Structure from Motion (SfM) problem involving…

计算机视觉与模式识别 · 计算机科学 2025-09-16 Daniil Sinitsyn , Linus Härenstam-Nielsen , Daniel Cremers

Nearly all 3D displays need calibration for correct rendering. More often than not, the optical elements in a 3D display are misaligned from the designed parameter setting. As a result, 3D magic does not perform well as intended. The…

计算机视觉与模式识别 · 计算机科学 2017-04-26 Hyoseok Hwang , Hyun Sung Chang , Dongkyung Nam , In So Kweon

A microphone array can provide a mobile robot with the capability of localizing, tracking and separating distant sound sources in 2D, i.e., estimating their relative elevation and azimuth. To combine acoustic data with visual information in…

音频与语音处理 · 电气工程与系统科学 2020-07-23 Simon Michaud , Samuel Faucher , François Grondin , Jean-Samuel Lauzon , Mathieu Labbé , Dominic Létourneau , François Ferland , François Michaud

Spatial audio is an essential medium to audiences for 3D visual and auditory experience. However, the recording devices and techniques are expensive or inaccessible to the general public. In this work, we propose a self-supervised audio…

声音 · 计算机科学 2019-05-15 Yu-Ding Lu , Hsin-Ying Lee , Hung-Yu Tseng , Ming-Hsuan Yang

We propose a novel sparse representation for heavily underdetermined multichannel sound mixtures, i.e., with much more sources than microphones. The proposed approach operates in the complex Fourier domain, thus preserving spatial…

声音 · 计算机科学 2014-10-10 Antoine Deleforge , Walter Kellermann

Photo-acoustic tomography (PAT) aims to leverage the photo-acoustic coupling between optical absorption of light sources and ultrasound (US) emission to obtain high contrast reconstructions of optical parameters with the high resolution of…

偏微分方程分析 · 数学 2016-09-21 Guillaume Bal , Amir Moradifam

Humans can robustly recognize and localize objects by integrating visual and auditory cues. While machines are able to do the same now with images, less work has been done with sounds. This work develops an approach for dense semantic…

计算机视觉与模式识别 · 计算机科学 2020-03-10 Arun Balajee Vasudevan , Dengxin Dai , Luc Van Gool
‹ 上一页 1 2 3 10 下一页 ›