中文
相关论文

相关论文: Computationally efficient spatial rendering of lat…

200 篇论文

Many spatial filtering algorithms used for voice capture in, e.g., teleconferencing applications, can benefit from or even rely on knowledge of Relative Transfer Functions (RTFs). Accordingly, many RTF estimators have been proposed which,…

音频与语音处理 · 电气工程与系统科学 2021-10-06 Andreas Brendel , Johannes Zeitler , Walter Kellermann

Room geometry inference algorithms rely on the localization of acoustic reflectors to identify boundary surfaces of an enclosure. Rooms with highly absorptive walls or walls at large distances from the measurement setup pose challenges for…

音频与语音处理 · 电气工程与系统科学 2024-02-12 H. Nazim Bicer , Cagdas Tuna , Andreas Walther , Emanuël A. P. Habets

Recently, deep representation learning has shown strong performance in multiple audio tasks. However, its use for learning spatial representations from multichannel audio is underexplored. We investigate the use of a pretraining stage based…

The recently introduced acoustic ray-tracing semiclassical (RTS) method is validated for a set of practically relevant boundary conditions. RTS is a frequency domain geometrical method which directly reproduces the acoustic Green's…

计算物理 · 物理学 2018-10-17 Rok Prislan , Daniel Svenšek

The use of spatial information with multiple microphones can improve far-field automatic speech recognition (ASR) accuracy. However, conventional microphone array techniques degrade speech enhancement performance when there is an array…

音频与语音处理 · 电气工程与系统科学 2021-12-23 Kenichi Kumatani , Minhua Wu , Shiva Sundaram , Nikko Strom , Bjorn Hoffmeister

Previous research on late-reverberation modeling has mainly focused on exponentially decaying room impulse responses, whereas methods for accurately modeling non-exponential reverberation remain challenging. This paper extends the…

音频与语音处理 · 电气工程与系统科学 2024-04-02 Jon Fagerström , Sebastian J. Schlecht , Vesa Välimäki

State-of-the-art deep-learning-based voice activity detectors (VADs) are often trained with anechoic data. However, real acoustic environments are generally reverberant, which causes the performance to significantly deteriorate. To mitigate…

声音 · 计算机科学 2021-06-28 Amir Ivry , Israel Cohen , Baruch Berdugo

Visual speech recognition (VSR) systems decode spoken words from an input sequence using only the video data. Practical applications of such systems include medical assistance as well as human-machine interactions. A VSR system is typically…

计算机视觉与模式识别 · 计算机科学 2025-08-26 Iason Ioannis Panagos , Giorgos Sfikas , Christophoros Nikou

Implicit neural representations (INRs) have emerged as a powerful tool for compressing large-scale volume data. This opens up new possibilities for in situ visualization. However, the efficient application of INRs to distributed data…

分布式、并行与集群计算 · 计算机科学 2024-07-23 Qi Wu , Joseph A. Insley , Victor A. Mateevitsi , Silvio Rizzi , Michael E. Papka , Kwan-Liu Ma

Underground stations are a common communication situation in towns: we talk with friends or colleagues, listen to announcements or shop for titbits while background noise and reverberation are challenging communication. Here, we perform an…

声音 · 计算机科学 2025-09-15 Ľuboš Hládek , Stephan D. Ewert , Bernhard U. Seeber

Modeling late reverberation in real-time interactive applications is a challenging task when multiple sound sources and listeners are present in the same environment. This is especially problematic when the environment is geometrically…

声音 · 计算机科学 2025-10-14 Matteo Scerbo , Sebastian J. Schlecht , Randall Ali , Lauri Savioja , Enzo De Sena

This paper proposes a speech enhancement method which exploits the high potential of residual connections in a Wide Residual Network architecture. This is supported on single dimensional convolutions computed alongside the time domain,…

音频与语音处理 · 电气工程与系统科学 2019-04-11 Jorge Llombart , Dayana Ribas , Antonio Miguel , Luis Vicente , Alfonso Ortega , Eduardo Lleida

Porous acoustic absorbers have excellent properties in the low-frequency range when positioned in room edges, therefore they are a common method for reducing low-frequency reverberation. However, standard room acoustic simulation methods…

Modeling room acoustics in a field setting involves some degree of blind parameter estimation from noisy and reverberant audio. Modern approaches leverage convolutional neural networks (CNNs) in tandem with time-frequency representation.…

音频与语音处理 · 电气工程与系统科学 2023-03-15 Christopher Ick , Adib Mehrabi , Wenyu Jin

The two-dimensional (2D) numerical approaches for vocal tract (VT) modelling can afford a better balance between the low computational cost and accurate rendering of acoustic wave propagation. However, they require a high spatio-temporal…

声音 · 计算机科学 2021-02-10 Debasish Ray Mohapatra , Victor Zappi , Sidney Fels

This article addresses the modeling of reverberant recording environments in the context of under-determined convolutive blind source separation. We model the contribution of each source to all mixture channels in the time-frequency domain…

机器学习 · 统计学 2009-12-14 Ngoc Duong , Emmanuel Vincent , Remi Gribonval

Multimodal research and applications are becoming more commonplace as Virtual Reality (VR) technology integrates different sensory feedback, enabling the recreation of real spaces in an audio-visual context. Within VR experiences, numerous…

音频与语音处理 · 电气工程与系统科学 2025-04-08 Mauricio Flores-Vargas , Enda Bates , Rachel McDonnell

In the context of building acoustics and the acoustic diagnosis of an existing room, this paper introduces and investigates a new approach to estimate mean absorption coefficients solely from a room impulse response (RIR). This inverse…

神经与进化计算 · 计算机科学 2021-09-02 Cédric Foy , Antoine Deleforge , Diego Di Carlo

Most speech enhancement algorithms make use of the short-time Fourier transform (STFT), which is a simple and flexible time-frequency decomposition that estimates the short-time spectrum of a signal. However, the duration of short STFT…

声音 · 计算机科学 2015-09-03 Scott Wisdom , Thomas Powers , Les Atlas , James Pitton

We address the issue of the exploding computational requirements of recent State-of-the-art (SOTA) open set multimodel 3D mapping (dense 3D mapping) algorithms and present Voxel-Aggregated Feature Synthesis (VAFS), a novel approach to dense…

计算机视觉与模式识别 · 计算机科学 2025-01-10 Owen Burns , Rizwan Qureshi