中文
相关论文

相关论文: A first-order DirAC-based parametric Ambisonic cod…

200 篇论文

Scene-based spatial audio formats, such as Ambisonics, are playback system agnostic and may therefore be favoured for delivering immersive audio experiences to a wide range of (potentially unknown) devices. The number of channels required…

音频与语音处理 · 电气工程与系统科学 2024-01-25 Christoph Hold , Leo McCormack , Archontis Politis , Ville Pulkki

This paper presents the 3D soundfield synthesis of the pressure field radiated by directional acoustic sources using both the multimodal method and higher-order ambisonics (HOA). Ambisonics is a technique for encoding and reproducing…

经典物理 · 物理学 2024-09-20 Philippe Thorner , Eric Bavu , Jean-Baptiste Doc , Christophe Langrenne

Spatial audio formats like Ambisonics are playback device layout-agnostic and well-suited for applications such as teleconferencing and virtual reality. Conventional Ambisonic encoding methods often rely on spherical microphone arrays for…

音频与语音处理 · 电气工程与系统科学 2024-09-17 Yue Qiao , Vinay Kothapally , Meng Yu , Dong Yu

The paper presents a method for improving spatial resolution of first-order ambisonic audio. The method is based on time/frequency decomposition of the audio with subsequent extraction of a directed plane wave from each frequency component.…

声音 · 计算机科学 2023-12-14 Denis Likhachov , Nick Petrovsky , Elias Azarov

Spatial audio enhances immersion by reproducing 3D sound fields, with Ambisonics offering a scalable format for this purpose. While first-order Ambisonics (FOA) notably facilitates hardware-efficient acquisition and storage of sound fields…

音频与语音处理 · 电气工程与系统科学 2026-03-31 Amit Milstein , Nir Shlezinger , Boaz Rafaely

Ambisonics is a spatial audio format describing a sound field. First-order Ambisonics (FOA) is a popular format comprising only four channels. This limited channel count comes at the expense of spatial accuracy. Ideally one would be able to…

音频与语音处理 · 电气工程与系统科学 2025-08-04 Ismael Nawfal , Symeon Delikaris Manias , Mehrez Souden , Juha Merimaa , Joshua Atkins , Elisabeth McMullin , Shadi Pirhosseinloo , Daniel Phillips

Neural audio codecs have been widely studied for mono and stereo signals, but spatial audio remains largely unexplored. We present the first discrete neural spatial audio codec for first-order ambisonics (FOA). Building on the WavTokenizer…

声音 · 计算机科学 2025-10-28 Parthasaarathy Sudarsanam , Sebastian Braun , Hannes Gamper

Diffusion probabilistic models have recently achieved remarkable success in generating high quality image and video data. In this work, we build on this class of generative models and introduce a method for lossy compression of high…

计算机视觉与模式识别 · 计算机科学 2023-03-30 Noor Fathima Ghouse , Jens Petersen , Auke Wiggers , Tianlin Xu , Guillaume Sautière

This contribution introduces a dataset of 7th-order Ambisonic Room Impulse Responses (HOA-RIRs), created using the Image Source Method. By employing higher-order Ambisonics, our dataset enables precise spatial audio reproduction, a critical…

声音 · 计算机科学 2025-06-02 Shivam Saini , Jürgen Peissig

The recently standardized 3GPP codec for Immersive Voice and Audio Services (IVAS) includes a parametric mode for efficiently coding multiple audio objects at low bit rates. In this mode, parametric side information is obtained from both…

音频与语音处理 · 电气工程与系统科学 2025-07-09 Andrea Eichenseer , Srikanth Korse , Guillaume Fuchs , Markus Multrus

Using deep neural networks (DNNs) for encoding of microphone array (MA) signals to the Ambisonics spatial audio format can surpass certain limitations of established conventional methods, but existing DNN-based methods need to be trained…

音频与语音处理 · 电气工程与系统科学 2025-01-15 Mikko Heikkinen , Archontis Politis , Konstantinos Drossos , Tuomas Virtanen

Ambisonics encoding of microphone array signals can enable various spatial audio applications, such as virtual reality or telepresence, but it is typically designed for uniformly-spaced spherical microphone arrays. This paper proposes a…

音频与语音处理 · 电气工程与系统科学 2024-01-12 Mikko Heikkinen , Archontis Politis , Tuomas Virtanen

This paper presents virtual upmixing of steering vectors captured by a fewer-channel spherical microphone array. This challenge has conventionally been addressed by recovering the directions and signals of sound sources from first-order…

音频与语音处理 · 电气工程与系统科学 2026-02-23 Emilio Picard , Diego Di Carlo , Aditya Arie Nugraha , Mathieu Fontaine , Kazuyoshi Yoshii

We present a deep neural network approach for encoding microphone array signals into Ambisonics that generalizes to arbitrary microphone array configurations with fixed microphone count but varying locations and frequency-dependent…

音频与语音处理 · 电气工程与系统科学 2026-02-02 Mikko Heikkinen , Archontis Politis , Konstantinos Drossos , Tuomas Virtanen

Direction of arrival (DOA) estimation employing low-resolution analog-to-digital convertors (ADCs) has emerged as a challenging and intriguing problem, particularly with the rise in popularity of large-scale arrays. The substantial…

信号处理 · 电气工程与系统科学 2024-01-05 Junkai Ji , Wei Mao , Feng Xi , Shengyao Chen

We introduce ImmerseDiffusion, an end-to-end generative audio model that produces 3D immersive soundscapes conditioned on the spatial, temporal, and environmental conditions of sound objects. ImmerseDiffusion is trained to generate…

声音 · 计算机科学 2025-02-11 Mojtaba Heydari , Mehrez Souden , Bruno Conejo , Joshua Atkins

Audio denoising is critical in signal processing, enhancing intelligibility and fidelity for applications like restoring musical recordings. This paper presents a proof-of-concept for adapting a state-of-the-art neural audio codec, the…

声音 · 计算机科学 2025-11-04 Daniel Jimon , Mircea Vaida , Adriana Stan

Advanced remote applications such as Networked Music Performance (NMP) require solutions to guarantee immersive real-world-like interaction among users. Therefore, the adoption of spatial audio formats, such as Ambisonics, is fundamental to…

音频与语音处理 · 电气工程与系统科学 2025-08-04 Paolo Ostan , Carlo Centofanti , Mirco Pezzoli , Alberto Bernardini , Claudia Rinaldi , Fabio Antonacci

Dithering is a technique commonly used to improve the perceptual quality of lossy data compression. In this work, we analytically and experimentally justify the use of dithering for ASR input compression. We formalize an understanding of…

音频与语音处理 · 电气工程与系统科学 2025-12-12 Ellison Murray , Morriel Kasher , Predrag Spasojevic

A multichannel extension to the RVQGAN neural coding method is proposed, and realized for data-driven compression of third-order Ambisonics audio. The input- and output layers of the generator and discriminator models are modified to accept…

声音 · 计算机科学 2024-12-13 Toni Hirvonen , Mahmoud Namazi
‹ 上一页 1 2 3 10 下一页 ›