English
Related papers

Related papers: Spheroidal Ambisonics: a Spatial Audio Framework U…

200 papers

Traditional video-to-audio generation techniques primarily focus on perspective video and non-spatial audio, often missing the spatial cues necessary for accurately representing sound sources in 3D environments. To address this limitation,…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-04 Huadai Liu , Tianyi Luo , Kaicheng Luo , Qikai Jiang , Peiwen Sun , Jialei Wang , Rongjie Huang , Qian Chen , Wen Wang , Xiangtai Li , Shiliang Zhang , Zhijie Yan , Zhou Zhao , Wei Xue

The field of sonification, using of non-speech audio for data analysis, is already established in space sciences. Meetings like "The audible Universe" focus on sonification tools applied in astronomy to represent complex data like nebulae…

Instrumentation and Methods for Astrophysics · Physics 2024-02-05 Johanna Casado , Beatriz García

We propose a practical framework to synthesize the broadband sound-field on a small rigid surface based on the physics of sound propagation. The sound-field is generated as a composite map of two components: the room component and the…

Sound · Computer Science 2024-07-16 Mohamed F. Mansour

A simple approach to microphone- and speaker-arrays is described in which the microphone array is regarded as a sampling grid for the acoustic field, and the corresponding speaker-array is treated as a "spatial digital to analog converter"…

Sound · Computer Science 2019-11-19 Julius O. Smith

Sonar-based indoor mapping systems have been widely employed in robotics for several decades. While such systems are still the mainstream in underwater and pipe inspection settings, the vulnerability to noise reduced, over time, their…

Robotics · Computer Science 2024-09-19 Usama Saqib , Letizia Marchegiani , Jesper Rindom Jensen

Humans can robustly recognize and localize objects by using visual and/or auditory cues. While machines are able to do the same with visual data already, less work has been done with sounds. This work develops an approach for scene…

Sound · Computer Science 2022-03-01 Dengxin Dai , Arun Balajee Vasudevan , Jiri Matas , Luc Van Gool

We present an indoor acoustic simulation framework that supports both ultrasonic and audible signaling. The framework opens the opportunity for fast indoor acoustic data generation and positioning development. The improved…

Audio and Speech Processing · Electrical Eng. & Systems 2023-06-22 Daan Delabie , Chesney Buyle , Bert Cox , Liesbet Van der Perre , Lieven De Strycker

The sampling of sound fields involves the measurement of spatially dependent room impulse responses, where the Nyquist-Shannon sampling theorem applies in both the temporal and spatial domain. Therefore, sampling inside a volume of interest…

Sound · Computer Science 2016-09-30 Fabrice Katzberg , Radoslaw Mazur , Marco Maass , Philipp Koch , Alfred Mertins

There is a growing effort in creating chiral transport of sound waves. However, most approaches so far are confined to the macroscopic scale. Here, we propose a new approach suitable to the nanoscale which is based on pseudomagnetic fields.…

Mesoscale and Nanoscale Physics · Physics 2017-09-06 Christian Brendel , Vittorio Peano , Oskar Painter , Florian Marquardt

We introduce nonparaxial spatially accelerating waves whose two-dimensional transverse profiles propagate along semicircular trajectories while approximately preserving their shape. We derive these waves by considering imaginary…

Optics · Physics 2015-03-13 Miguel A. Alonso , Miguel A. Bandres

Recent molecular communication (MC) research has integrated more detailed computational models to capture the dynamics of practical biophysical systems. This research focuses on developing realistic models for MC transceivers inspired by…

Can uncorrelated surrounding sound sources be used to generate extended diffuse sound fields? By definition, targets are a constant sound pressure level, a vanishing average sound intensity, uncorrelated sound waves arriving isotropically…

Audio and Speech Processing · Electrical Eng. & Systems 2024-02-22 Franz Zotter , Stefan Riedel , Lukas Gölles , Matthias Frank

Binaural audio provides human listeners with an immersive spatial sound experience, but most existing videos lack binaural audio recordings. We propose an audio spatialization method that draws on visual information in videos to convert…

Computer Vision and Pattern Recognition · Computer Science 2021-11-23 Rishabh Garg , Ruohan Gao , Kristen Grauman

Animals hear and vocalize across frequency ranges that differ substantially from humans, often extending into the ultrasonic domain. Yet most computational bioacoustics systems rely on audio models pre-trained at 16 kHz, restricting their…

Modeling and forecasting the solar wind-driven global magnetic field perturbations is an open challenge. Current approaches depend on simulations of computationally demanding models like the Magnetohydrodynamics (MHD) model or sampling…

Humans can robustly recognize and localize objects by integrating visual and auditory cues. While machines are able to do the same now with images, less work has been done with sounds. This work develops an approach for dense semantic…

Computer Vision and Pattern Recognition · Computer Science 2020-03-10 Arun Balajee Vasudevan , Dengxin Dai , Luc Van Gool

We study the problem of multimodal physical scene understanding, where an embodied agent needs to find fallen objects by inferring object properties, direction, and distance of an impact sound source. Previous works adopt feed-forward…

Robotics · Computer Science 2024-07-17 Jie Yin , Andrew Luo , Yilun Du , Anoop Cherian , Tim K. Marks , Jonathan Le Roux , Chuang Gan

In this article, we present a space-frequency theory for spherical harmonics based on the spectral decomposition of a particular space-frequency operator. The presented theory is closely linked to the theory of ultraspherical polynomials on…

Numerical Analysis · Mathematics 2013-07-16 Wolfgang Erb , Sonja Mathias

In this work, we address the challenge of generalizable audio deepfake detection (ADD) across diverse speech synthesis paradigms-including conventional text-to-speech (TTS) systems and modern diffusion or flow-matching (FM) based…

Audio and Speech Processing · Electrical Eng. & Systems 2025-11-17 Farhan Sheth , Girish , Mohd Mujtaba Akhtar , Muskaan Singh

The use of spatial information with multiple microphones can improve far-field automatic speech recognition (ASR) accuracy. However, conventional microphone array techniques degrade speech enhancement performance when there is an array…

Audio and Speech Processing · Electrical Eng. & Systems 2021-12-23 Kenichi Kumatani , Minhua Wu , Shiva Sundaram , Nikko Strom , Bjorn Hoffmeister
‹ Prev 1 4 5 6 7 8 10 Next ›