中文
相关论文

相关论文: Evaluation of an open-source implementation of the…

200 篇论文

Accurately estimating sound source positions is crucial for robot audition. However, existing sound source localization methods typically rely on a microphone array with at least two spatially preconfigured microphones. This requirement…

机器人学 · 计算机科学 2025-06-23 Jiang Wang , Runwu Shi , Benjamin Yen , He Kong , Kazuhiro Nakadai

Gas source localization (GSL) with an autonomous robot is a problem with many prospective applications, from finding pipe leaks to emergency-response scenarios. In this work, we present a new method to perform GSL in realistic indoor…

机器人学 · 计算机科学 2024-07-23 Pepe Ojeda , Javier Monroy , Javier Gonzalez-Jimenez

In this paper, the problem of extending narrowband multichannel sound source localization algorithms to the wideband case is addressed. The DOA estimation of narrowband algorithms is based on the estimate of inter-channel phase differences…

音频与语音处理 · 电气工程与系统科学 2019-06-24 Kainan Chen , Wenyu Jin , Bharadwaj Desikan

This paper presents a factor graph formulation and particle-based sum-product algorithm (SPA) for robust sequential localization in multipath-prone environments. The proposed algorithm jointly performs data association, sequential…

信号处理 · 电气工程与系统科学 2023-06-16 Alexander Venus , Erik Leitinger , Stefan Tertinek , Klaus Witrisal

In this paper we use the MAP criterion to locate a region containing a source. Sensors placed in a field of interest divide the latter into smaller regions and take measurements that are transmitted over noisy wireless channels. We propose…

最优化与控制 · 数学 2009-03-19 S. H. Dandach , F. Bullo

This paper proposes a data-driven algorithm of locating the source of forced oscillations and suggests the physical interpretation of the method. By leveraging the sparsity of the forced oscillation sources along with the low-rank nature of…

信号处理 · 电气工程与系统科学 2019-08-28 Tong Huang , Nikolaos M. Freris , P. R. Kumar , Le Xie

We address the problem of online localization and tracking of multiple moving speakers in reverberant environments. The paper has the following contributions. We use the direct-path relative transfer function (DP-RTF), an inter-channel…

声音 · 计算机科学 2019-04-11 Xiaofei Li , Yutong Ban , Laurent Girin , Xavier Alameda-Pineda , Radu Horaud

This document describes our submission to the 2018 LOCalization And TrAcking (LOCATA) challenge (Tasks 1, 3, 5). We estimate the 3D position of a speaker using the Global Coherence Field (GCF) computed from multiple microphone pairs of a…

声音 · 计算机科学 2019-01-28 Xinyuan Qian , Andrea Cavallaro , Alessio Brutti , Maurizio Omologo

Accurate sound source localization (SSL), such as direction-of-arrival (DoA) estimation, relies on consistent multichannel data. However, batteryless systems often suffer from missing data due to the stochastic nature of energy harvesting,…

机器学习 · 计算机科学 2025-07-21 Subrata Biswas , Mohammad Nur Hossain Khan , Violet Colwell , Jack Adiletta , Bashima Islam

Multi-speaker localization and tracking using microphone array recording is of importance in a wide range of applications. One of the challenges with multi-speaker tracking is to associate direction estimates with the correct speaker. Most…

音频与语音处理 · 电气工程与系统科学 2024-10-16 Hanan Beit-On , Vladimir Tourbabin , Boaz Rafaely

Accurate localization and perception are pivotal for enhancing the safety and reliability of vehicles. However, current localization methods suffer from reduced accuracy when the line-of-sight (LOS) path is obstructed, or a combination of…

信号处理 · 电气工程与系统科学 2024-08-16 Yinuo Du , Hanying Zhao , Yang Liu , Xinlei Yu , Yuan Shen

Dynamic objects in the environment, such as people and other agents, lead to challenges for existing simultaneous localization and mapping (SLAM) approaches. To deal with dynamic environments, computer vision researchers usually apply some…

机器人学 · 计算机科学 2021-08-04 Tianwei Zhang , Huayan Zhang , Xiaofei Li , Junfeng Chen , Tin Lun Lam , Sethu Vijayakumar

This study presents a system for sound source localization in time domain using a deep residual neural network. Data from the linear 8 channel microphone array with 3 cm spacing is used by the network for direction estimation. We propose to…

声音 · 计算机科学 2018-08-21 Dmitry Suvorov , Ge Dong , Roman Zhukov

This paper implements Simultaneous Localization and Mapping (SLAM) technique to construct a map of a given environment. A Real Time Appearance Based Mapping (RTAB-Map) approach was taken for accomplishing this task. Initially, a 2d…

机器人学 · 计算机科学 2018-09-11 Sagarnil Das

This paper addresses source localization problem in a random shallow water channel. We present an extension of the generalized MUSIC method to the case, %in which when the signal correlation matrix is imprecisely known. The algorithm is…

大气与海洋物理 · 物理学 2014-10-29 Alexander Sazontov , Ivan Smirnov , Alexander Matveyev

The steered response power (SRP) is a popular approach to compute a map of the acoustic scene, typically used for acoustic source localization. The SRP map is obtained as the frequency-weighted output power of a beamformer steered towards a…

音频与语音处理 · 电气工程与系统科学 2024-11-25 Thomas Dietzen , Enzo De Sena , Toon van Waterschoot

Sound event localization aims at estimating the positions of sound sources in the environment with respect to an acoustic receiver (e.g. a microphone array). Recent advances in this domain most prominently focused on utilizing deep…

Sound sources localization using multichannel signal processing has been a subject of active research for decades. In recent years, the use of deep learning in audio signal processing has allowed to drastically improve performances for…

音频与语音处理 · 电气工程与系统科学 2021-06-16 Hadrien Pujol , Éric Bavu , Alexandre Garcia

Dereverberation of a moving speech source in the presence of other directional interferers, is a harder problem than that of stationary source and interference cancellation. We explore joint multi channel linear prediction (MCLP) and…

音频与语音处理 · 电气工程与系统科学 2019-10-23 Srikanth Raj Chetupalli , Thippur V. Sreenivas

How to visually localize multiple sound sources in unconstrained videos is a formidable problem, especially when lack of the pairwise sound-object annotations. To solve this problem, we develop a two-stage audiovisual learning framework…

计算机视觉与模式识别 · 计算机科学 2020-07-15 Rui Qian , Di Hu , Heinrich Dinkel , Mengyue Wu , Ning Xu , Weiyao Lin