English
Related papers

Related papers: PSD Estimation of Multiple Sound Sources in a Reve…

200 papers

Neural networks (NNs) have been widely applied in speech processing tasks, and, in particular, those employing microphone arrays. Nevertheless, most existing NN architectures can only deal with fixed and position-specific microphone arrays.…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-14 Yochai Yemini , Ethan Fetaya , Haggai Maron , Sharon Gannot

Extracting direct-path spatial features is critical for sound source localization in adverse acoustic environments. This paper proposes a full-band and narrow-band fusion network for estimating direct-path inter-channel phase difference…

Audio and Speech Processing · Electrical Eng. & Systems 2023-06-01 Yabo Wang , Bing Yang , Xiaofei Li

This paper addresses the problem of localizing audio sources using binaural measurements. We propose a supervised formulation that simultaneously localizes multiple sources at different locations. The approach is intrinsically efficient…

Sound · Computer Science 2016-04-18 Antoine Deleforge , Radu Horaud , Yoav Schechner , Laurent Girin

An original experimental method for determining the empirical probability distribution function (PDF) of the quality factor (Q) of a mode-stirred reverberation chamber is presented. Spectral averaging of S-parameters across a relatively…

Optics · Physics 2015-06-18 Luk R. Arnaut , Mihai I. Andries , Jerome Sol , Philippe Besnier

Performing an adequate evaluation of sound event detection (SED) systems is far from trivial and is still subject to ongoing research. The recently proposed polyphonic sound detection (PSD)-receiver operating characteristic (ROC) and PSD…

Audio and Speech Processing · Electrical Eng. & Systems 2022-02-01 Janek Ebbers , Romain Serizel , Reinhold Haeb-Umbach

In this paper, the task of channel sounding using software defined radios (SDRs) is considered. In contrast to classical channel sounding equipment, SDRs are general purpose devices and require additional steps to be implemented when…

Signal Processing · Electrical Eng. & Systems 2022-05-24 Julian Ahrens , Lia Ahrens , Michael Zentarra , Hans D. Schotten

A dictionary learning based audio source classification algorithm is proposed to classify a sample audio signal as one amongst a finite set of different audio sources. Cosine similarity measure is used to select the atoms during dictionary…

Sound · Computer Science 2015-10-28 K V Vijay Girish , T V Ananthapadmanabha , A G Ramakrishnan

Separating different speaker properties from a multi-speaker environment is challenging. Instead of separating a two-speaker signal in signal space like speech source separation, a speaker embedding de-mixing approach is proposed. The…

Sound · Computer Science 2021-02-08 Yanpei Shi , Thomas Hain

Accurately estimating sound source positions is crucial for robot audition. However, existing sound source localization methods typically rely on a microphone array with at least two spatially preconfigured microphones. This requirement…

Robotics · Computer Science 2025-06-23 Jiang Wang , Runwu Shi , Benjamin Yen , He Kong , Kazuhiro Nakadai

X-ray reverberation has become a powerful tool to probe the disc-corona geometry near black holes. Here, we develop Machine Learning (ML) models to extract the X-ray reverberation features imprinted in the Power Spectral Density (PSD) of…

High Energy Astrophysical Phenomena · Physics 2021-08-18 P. Chainakun , N. Mankatwit , P. Thongkonsing , A. J. Young

We present in this paper a spectrally accurate numerical method for computing the spherical/vector spherical harmonic expansion of a function/vector field with given (elemental) nodal values on a spherical surface. Built upon suitable…

Numerical Analysis · Mathematics 2017-09-18 Bo Wang , Li-Lian Wang , Ziqing Xie

Distributed microphone arrays composed of multiple subarrays enable blind source separation over a wide spatial area. Directly applying fast multichannel nonnegative matrix factorization (FastMNMF) to all subarrays can exploit observations…

Audio and Speech Processing · Electrical Eng. & Systems 2026-05-20 Hirotaka Nishikori , Nobutaka Ito , Kouei Yamaoka , Norihiro Takamune , Hiroshi Saruwatari

The increasing popularity of spatial audio in applications such as teleconferencing, entertainment, and virtual reality has led to the recent developments of binaural reproduction methods. However, only a few of these methods are…

Audio and Speech Processing · Electrical Eng. & Systems 2025-02-17 Ami Berger , Vladimir Tourbabin , Jacob Donley , Zamir Ben-Hur , Boaz Rafaely

This paper addresses the problem of binaural localization of a single speech source in noisy and reverberant environments. For a given binaural microphone setup, the binaural response corresponding to the direct-path propagation of a single…

Sound · Computer Science 2016-09-08 Xiaofei Li , Laurent Girin , Radu Horaud , Sharon Gannot

We present a single channel data driven method for non-intrusive estimation of full-band reverberation time and full-band direct-to-reverberant ratio. The method extracts a number of features from reverberant speech and builds a model using…

Sound · Computer Science 2015-10-16 Pablo Peso Parada , Dushyant Sharma , Toon van Waterschoot , Patrick A. Naylor

Spatial frequency estimation from a mixture of noisy sinusoids finds applications in various fields. While subspace-based methods offer cost-effective super-resolution parameter estimation, they demand precise array calibration, posing…

Signal Processing · Electrical Eng. & Systems 2024-10-23 Tianyi Liu , Sai Pavan Deram , Khaled Ardah , Martin Haardt , Marc E. Pfetsch , Marius Pesavento

Blind Speech Separation (BSS) aims to separate multiple speech sources from audio mixtures recorded by a microphone array. The problem is challenging because it is a blind inverse problem, i.e., the microphone array geometry, the room…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-18 Zhongweiyang Xu , Xulin Fan , Zhong-Qiu Wang , Xilin Jiang , Romit Roy Choudhury

Audio source separation is often achieved by estimating the magnitude spectrogram of each source, and then applying a phase recovery (or spectrogram inversion) algorithm to retrieve time-domain signals. Typically, spectrogram inversion is…

Sound · Computer Science 2023-07-03 Paul Magron , Tuomas Virtanen

Coherent diffraction imaging (CDI) is high-resolution lensless microscopy that has been applied to image a wide range of specimens using synchrotron radiation, X-ray free electron lasers, high harmonic generation, soft X-ray laser and…

Optics · Physics 2014-01-23 Jose A Rodriguez , Rui Xu , Chien-Chun Chen , Yunfei Zou , Jianwei Miao

In this paper, robust detection, tracking and geometry estimation methods are developed and combined into a system for estimating time-difference estimates, microphone localization and sound source movement. No assumptions on the 3D…