English
Related papers

Related papers: Ambisonics Networks -- The Effect Of Radial Functi…

200 papers

Learning from data in the quaternion domain enables us to exploit internal dependencies of 4D signals and treating them as a single entity. One of the models that perfectly suits with quaternion-valued data processing is represented by 3D…

Audio and Speech Processing · Electrical Eng. & Systems 2022-12-16 Danilo Comminiello , Marco Lella , Simone Scardapane , Aurelio Uncini

This document illustrates how to process the signals from the microphones of a rigid-sphere higher-order ambisonic microphone array so that they are encoded with N3D normalization and ACN channel order and thereby can be used with the…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-02 Jens Ahrens

Acoustic signal processing in the spherical harmonics domain (SHD) is an active research area that exploits the signals acquired by higher order microphone arrays. A very important task is that concerning the localization of active sound…

Audio and Speech Processing · Electrical Eng. & Systems 2024-01-29 Maximo Cobos , Mirco Pezzoli , Fabio Antonacci , Augusto Sarti

Recently, the Spherical Wavelet Framework (SWF) was proposed to combine the benefits of Ambisonics and Object-Based Audio (OBA) by utilising highly localised basis functions. SWF can enhance the sweet-spot area and reduce localisation blur…

Audio and Speech Processing · Electrical Eng. & Systems 2025-10-28 Ş. Ekmen , H. Lee

Ambisonics rendering has become an integral part of 3D audio for headphones. It works well with existing recording hardware, the processing cost is mostly independent of the number of sound sources, and it elegantly allows for rotating the…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-27 Or Berebi , Fabian Brinkmann , Stefan Weinzierl , Boaz Rafaely

Overfitting is one of the most critical challenges in deep neural networks, and there are various types of regularization methods to improve generalization performance. Injecting noises to hidden units during training, e.g., dropout, is…

Machine Learning · Computer Science 2017-11-10 Hyeonwoo Noh , Tackgeun You , Jonghwan Mun , Bohyung Han

Advanced remote applications such as Networked Music Performance (NMP) require solutions to guarantee immersive real-world-like interaction among users. Therefore, the adoption of spatial audio formats, such as Ambisonics, is fundamental to…

Audio and Speech Processing · Electrical Eng. & Systems 2025-08-04 Paolo Ostan , Carlo Centofanti , Mirco Pezzoli , Alberto Bernardini , Claudia Rinaldi , Fabio Antonacci

Ambisonics Signal Matching (ASM) is a recently proposed signal-independent approach to encoding Ambisonic signal from wearable microphone arrays, enabling efficient and standardized spatial sound reproduction. However, reproduction accuracy…

Audio and Speech Processing · Electrical Eng. & Systems 2025-07-08 Yhonatan Gayer , Vladimir Tourbabin , Zamir Ben-Hur , David Alon , Boaz Rafaely

Acoustical signal processing of directional representations of sound fields, including source, receiver, and scatterer transfer functions, are often expressed and modeled in the spherical harmonic domain (SHD). Certain such modeling…

Audio and Speech Processing · Electrical Eng. & Systems 2024-07-10 Archontis Politis

We present a deep neural network approach for encoding microphone array signals into Ambisonics that generalizes to arbitrary microphone array configurations with fixed microphone count but varying locations and frequency-dependent…

Audio and Speech Processing · Electrical Eng. & Systems 2026-02-02 Mikko Heikkinen , Archontis Politis , Konstantinos Drossos , Tuomas Virtanen

Parametric sound field synthesis methods, such as the Spatial Decomposition Method (SDM) and Higher-Order Spatial Impulse Response Rendering (HO-SIRR), are widely used for the analysis and auralization of sound fields. This paper studies…

Audio and Speech Processing · Electrical Eng. & Systems 2024-11-04 Alan Pawlak , Hyunkook Lee , Aki Mäkivirta , Thomas Lund

Speech enhancement in ad-hoc microphone arrays is often hindered by the asynchronization of the devices composing the microphone array. Asynchronization comes from sampling time offset and sampling rate offset which inevitably occur when…

Audio and Speech Processing · Electrical Eng. & Systems 2023-08-01 Nicolas Furnon , Romain Serizel , Slim Essid , Irina Illina

Characterizing sound field diffuseness has many practical applications, from room acoustics analysis to speech enhancement and sound field reproduction. In this paper we investigate how spherical microphone arrays (SMAs) can be used to…

Sound · Computer Science 2016-07-04 Nicolas Epain , Craig T. Jin

The vulnerability of neural network classifiers to adversarial attacks is a major obstacle to their deployment in safety-critical applications. Regularization of network parameters during training can be used to improve adversarial…

Machine Learning · Computer Science 2024-05-28 Sheng Yang , Jacob A. Zavatone-Veth , Cengiz Pehlevan

Sound field reconstruction (SFR) augments the information of a sound field captured by a microphone array. Conventional SFR methods using basis function decomposition are straightforward and computationally efficient, but may require more…

Audio and Speech Processing · Electrical Eng. & Systems 2024-02-15 Fei Ma , Sipei Zhao , Ian S. Burnett

Deep neural network (DNN)-based speech enhancement algorithms in microphone arrays have now proven to be efficient solutions to speech understanding and speech recognition in noisy environments. However, in the context of ad-hoc microphone…

Signal Processing · Electrical Eng. & Systems 2020-11-04 Nicolas Furnon , Romain Serizel , Irina Illina , Slim Essid

Sound field reproduction with undistorted sound quality and precise spatial localization is desirable for automotive audio systems. However, the complexity of automotive cabin acoustic environment often necessitates a trade-off between…

Audio and Speech Processing · Electrical Eng. & Systems 2025-09-12 Yufan Qian , Tianshu Qu , Xihong Wu

Training convolutional recurrent neural networks on first-order Ambisonics signals is a well-known approach when estimating the direction of arrival for speech/sound signals. In this work, we investigate whether increasing the order of…

Audio and Speech Processing · Electrical Eng. & Systems 2021-05-07 Nils Poschadel , Robert Hupke , Stephan Preihs , Jürgen Peissig

We propose a spatial diffuseness feature for deep neural network (DNN)-based automatic speech recognition to improve recognition accuracy in reverberant and noisy environments. The feature is computed in real-time from multiple microphone…

Computation and Language · Computer Science 2015-09-02 Andreas Schwarz , Christian Huemmer , Roland Maas , Walter Kellermann

Diffusion imaging is an important method in the field of neuroscience, as it is sensitive to changes within the tissue microstructure of the human brain. However, a major challenge when using MRI to derive quantitative measures is that the…

Computer Vision and Pattern Recognition · Computer Science 2018-08-07 Simon Koppers , Luke Bloy , Jeffrey I. Berman , Chantal M. W. Tax , J. Christopher Edgar , Dorit Merhof