中文
相关论文

相关论文: Acoustic Room Compensation Using Local PCA-based R…

200 篇论文

We introduce a novel algorithm for online estimation of acoustic impulse responses (AIRs) which allows for fast convergence by exploiting prior knowledge about the fundamental structure of AIRs. The proposed method assumes that the…

音频与语音处理 · 电气工程与系统科学 2021-05-10 Thomas Haubner , Andreas Brendel , Walter Kellermann

We present a method to maintain the subjective perception of volume of audio signals and, at the same time, reduce their absolute peak value. We focus on achieving this without compromising the perceived audio quality. This is specially…

信号处理 · 电气工程与系统科学 2022-02-17 A. Jeannerot , N. de Koeijer , P. Martínez-Nuevo , M. B. Møller , J. Dyreby , P. Prandoni

Acoustic Echo Cancellation (AEC) plays a key role in voice interaction. Due to the explicit mathematical principle and intelligent nature to accommodate conditions, adaptive filters with different types of implementations are always used…

声音 · 计算机科学 2020-05-20 Lu Ma , Hua Huang , Pei Zhao , Tengrong Su

Prediction of room impulse responses (RIRs) is essential for room acoustics, spatial audio, and immersive applications, yet conventional simulations and measurements remain computationally expensive and time-consuming. This work proposes a…

音频与语音处理 · 电气工程与系统科学 2025-09-30 Imran Muhammad , Gerald Schuller

Automatic speech recognition (ASR) systems have been widely deployed in modern smart devices to provide convenient and diverse voice-controlled services. Since ASR systems are vulnerable to audio replay attacks that can spoof and mislead…

密码学与安全 · 计算机科学 2022-08-05 Shu Wang , Jiahao Cao , Xu He , Kun Sun , Qi Li

In this paper, we propose a method to estimate the proximity of an acoustic reflector, e.g., a wall, using ego-noise, i.e., the noise produced by the moving parts of a listening robot. This is achieved by estimating the times of arrival of…

声音 · 计算机科学 2021-11-17 Usama Saqib , Antoine Deleforge , Jesper Jensen

Voice assistants overhear conversations and a consent management mechanism is required. Consent management can be implemented using speaker recognition. Users that do not give consent enrol their voice and all their further recordings are…

声音 · 计算机科学 2024-10-28 Arash Shahmansoori , Utz Roedig

Sound in indoor spaces forms a complex wavefield due to multiple scattering encountered by the sound. Indoor acoustic communication involving multiple sources and receivers thus inevitably suffers from cross-talks. Here, we demonstrate the…

声音 · 计算机科学 2024-02-13 Hongkuan Zhang , Qiyuan Wang , Mathias Fink , Guancong Ma

Active noise control (ANC) over a sizeable space requires a large number of reference and error microphones to satisfy the spatial Nyquist sampling criterion, which limits the feasibility of practical realization of such systems. This paper…

声音 · 计算机科学 2018-03-02 Yu Maeno , Yuki Mitsufuji , Thushara D. Abhayapala

In their everyday life, the speech recognition performance of human listeners is influenced by diverse factors, such as the acoustic environment, the talker and listener positions, possibly impaired hearing, and optional hearing devices.…

音频与语音处理 · 电气工程与系统科学 2021-04-02 Marc René Schädler

Given an input sound signal and a target virtual sound source, sound spatialisation algorithms manipulate the signal so that a listener perceives it as though it were emitted from the target source. There exist several established…

声音 · 计算机科学 2017-11-28 Ali Tarzan , Marco Alunno , Paolo Bientinesi

Transformer models have been used in automatic speech recognition (ASR) successfully and yields state-of-the-art results. However, its performance is still affected by speaker mismatch between training and test data. Further finetuning a…

音频与语音处理 · 电气工程与系统科学 2021-10-19 Yingzhu Zhao , Chongjia Ni , Cheung-Chi Leung , Shafiq Joty , Eng Siong Chng , Bin Ma

This paper addresses the prevalent issue of incorrect speech output in audio-visual speech enhancement (AVSE) systems, which is often caused by poor video quality and mismatched training and test data. We introduce a post-processing…

音频与语音处理 · 电气工程与系统科学 2024-10-01 Wenze Ren , Kuo-Hsuan Hung , Rong Chao , YouJin Li , Hsin-Min Wang , Yu Tsao

Approximate computing is an emerging computing paradigm that offers improved power consumption by relaxing the requirement for full accuracy. Since real-world applications may have different requirements for design accuracy, one trend of…

硬件体系结构 · 计算机科学 2022-07-04 Jingxiao Ma , Sherief Reda

Acoustic-to-articulatory inversion (AAI) is to obtain the movement of articulators from speech signals. Until now, achieving a speaker-independent AAI remains a challenge given the limited data. Besides, most current works only use audio…

声音 · 计算机科学 2022-04-05 Jianrong Wang , Jinyu Liu , Longxuan Zhao , Shanyu Wang , Ruiguo Yu , Li Liu

Automatic Speaker Verification (ASV) is the process of identifying a person based on the voice presented to a system. Different synthetic approaches allow spoofing to deceive ASV systems (ASVs), whether using techniques to imitate a voice…

音频与语音处理 · 电气工程与系统科学 2019-10-30 Mohammad Adiban , Hossein Sameti , Saeedreza Shehnepoor

A new device, the opto-acoustic parametric amplifier (OAPA), is introduced, which is analogous to an optical parametric amplifier (OPA). The radiation pressure reaction of light on the reflective surface of an acoustic resonator provides…

广义相对论与量子宇宙学 · 物理学 2008-02-28 Chunnong Zhao , Li Ju , Haixing Miao , Slawomir Gras , Yaohui Fan , David G. Blair

The study of spatial audio and room acoustics aims to create immersive audio experiences by modeling the physics and psychoacoustics of how sound behaves in space. In the long history of this research area, various key technologies have…

音频与语音处理 · 电气工程与系统科学 2025-03-18 Shoichi Koyama , Enzo De Sena , Prasanga Samarasinghe , Mark R. P. Thomas , Fabio Antonacci

Installed sound applications typically involve a large number of audio processors, amplifiers and speaker systems spread across the venue. They could be spatially distributed at the venue across different rack rooms and floors. These…

网络与互联网体系结构 · 计算机科学 2016-11-17 Shubham Sharma , Aditya Sridhar , Jai Prakash Krishnia

Immersion in virtual and augmented reality solutions is reliant on plausible spatial audio. However, plausibly representing a space for immersive audio often requires many individual acoustic measurements of source-microphone pairs with…

音频与语音处理 · 电气工程与系统科学 2025-10-07 Ben Heritage , Fiona Ryder , Michael McLoughlin , Karolina Prawda