中文
相关论文

相关论文: Velocity Potential Neural Field for Efficient Ambi…

200 篇论文

A recurrent neural network model of phonological pattern learning is proposed. The model is a relatively simple neural network with one recurrent layer, and displays biases in learning that mimic observed biases in human learning.…

计算与语言 · 计算机科学 2024-05-31 Amanda Doucette

Most of the prior studies in the spatial \ac{DoA} domain focus on a single modality. However, humans use auditory and visual senses to detect the presence of sound sources. With this motivation, we propose to use neural networks with audio…

声音 · 计算机科学 2021-05-14 Xinyuan Qian , Maulik Madhavi , Zexu Pan , Jiadong Wang , Haizhou Li

We introduce a novel all neural model for low-latency directional speech extraction. The model uses direction of arrival (DOA) embeddings from a predefined spatial grid, which are transformed and fused into a recurrent neural network based…

声音 · 计算机科学 2024-07-09 Ashutosh Pandey , Sanha Lee , Juan Azcarreta , Daniel Wong , Buye Xu

The learning-from-observation (LfO) framework aims to map human demonstrations to a robot to reduce programming effort. To this end, an LfO system encodes a human demonstration into a series of execution units for a robot, which are…

机器人学 · 计算机科学 2021-03-25 Naoki Wake , Iori Yanokura , Kazuhiro Sasabuchi , Katsushi Ikeuchi

Over the past few decades, extensive research has been devoted to the design of artificial reverberation algorithms aimed at emulating the room acoustics of physical environments. Despite significant advancements, automatic parameter tuning…

音频与语音处理 · 电气工程与系统科学 2024-10-10 Alessandro Ilic Mezza , Riccardo Giampiccolo , Enzo De Sena , Alberto Bernardini

Traditional sound design workflows rely on manual alignment of audio events to visual cues, as in Foley sound design, where everyday actions like footsteps or object interactions are recreated to match the on-screen motion. This process is…

The network studied here is based on a standard model in physics, but it appears in various applications ranging from spintronics to neuroscience. When the network is forced by an external signal common to all its elements, there are shown…

适应与自组织系统 · 物理学 2020-08-18 Frank Hoppensteadt

Large audio-language models have made rapid progress in recognizing what is present in an audio clip, but spatial audio-language understanding still lacks a clear task interface. A model must also decide where sound events occur, which…

声音 · 计算机科学 2026-05-12 Yuhuan You , Lai Wei , Xihong Wu , Tianshu Qu

Recently deep learning and machine learning approaches have been widely employed for various applications in acoustics. Nonetheless, in the area of sound field processing and reconstruction classic methods based on the solutions of wave…

音频与语音处理 · 电气工程与系统科学 2025-01-07 Mirco Pezzoli , Fabio Antonacci , Augusto Sarti

In sound field control applications, it is commonly assumed that one has access to an accurate representation of the sound field in the region of interest. This is a problematic assumption since the reconstruction of a sound field from…

音频与语音处理 · 电气工程与系统科学 2026-05-21 David Sundström , Filip Tronarp , Johan Lindström , Andreas Jakobsson

Recently, the Spherical Wavelet Framework (SWF) was proposed to combine the benefits of Ambisonics and Object-Based Audio (OBA) by utilising highly localised basis functions. SWF can enhance the sweet-spot area and reduce localisation blur…

音频与语音处理 · 电气工程与系统科学 2025-10-28 Ş. Ekmen , H. Lee

Immersive spatial audio has become increasingly critical for applications ranging from AR/VR to home entertainment and automotive sound systems. However, existing generative methods remain constrained to low-dimensional formats such as…

音频与语音处理 · 电气工程与系统科学 2026-01-21 Zining Liang , Runbang Wang , Xuzhou Ye , Qiuqiang Kong

Data from diffusion magnetic resonance imaging (dMRI) can be used to reconstruct fiber tracts, for example, in muscle and white matter. Estimation of fiber orientations (FOs) is a crucial step in the reconstruction process and these…

计算机视觉与模式识别 · 计算机科学 2016-05-17 Chuyang Ye , Jiachen Zhuo , Rao P. Gullapalli , Jerry L. Prince

Sound field reconstruction refers to the problem of estimating the acoustic pressure field over an arbitrary region of space, using only a limited set of measurements. Physics-informed neural networks have been adopted to solve the problem…

音频与语音处理 · 电气工程与系统科学 2025-06-05 Stefano Damiano , Toon van Waterschoot

Speech representation and modelling in high-dimensional spaces of acoustic waveforms, or a linear transformation thereof, is investigated with the aim of improving the robustness of automatic speech recognition to additive noise. The…

计算与语言 · 计算机科学 2015-03-31 Matthew Ager , Zoran Cvetkovic , Peter Sollich

We study associative memory based on temporal coding in which successful retrieval is realized as an entrainment in a network of simple phase oscillators with distributed natural frequencies under the influence of white noise. The memory…

无序系统与神经网络 · 物理学 2009-10-31 Masahiko Yoshioka , Masatoshi Shiino

Ambisonics is an established framework to capture, process, and reproduce spatial sound fields based on its spherical harmonics representation. We propose a generalization of conventional spherical ambisonics to the spheroidal coordinate…

声音 · 计算机科学 2023-01-06 Shoken Kaneko

The acceptance rate in Woodcock tracking algorithm is generalized to an arbitrary position-dependent variable $q(x)$. A neural network is used to optimize $q(x)$, and the FOM value is used as the loss function. This idea comes from physics…

计算物理 · 物理学 2025-02-20 Bingnan Zhang

Achieving immersive auditory experiences in virtual environments requires flexible sound modeling that supports dynamic source positions. In this paper, we introduce a task called resounding, which aims to estimate room impulse responses at…

声音 · 计算机科学 2025-10-24 Zitong Lan , Yiduo Hao , Mingmin Zhao

This work presents a novel probabilistic interpretation of Slow Feature Analysis (SFA) through the lens of variational inference. Unlike prior formulations that recover linear SFA from Gaussian state-space models with linear emissions, this…

机器学习 · 计算机科学 2025-06-03 Merlin Schüler , Laurenz Wiskott