中文
相关论文

相关论文: The feasibility of sound zone control using an arr…

200 篇论文

Spatial audio formats like Ambisonics are playback device layout-agnostic and well-suited for applications such as teleconferencing and virtual reality. Conventional Ambisonic encoding methods often rely on spherical microphone arrays for…

音频与语音处理 · 电气工程与系统科学 2024-09-17 Yue Qiao , Vinay Kothapally , Meng Yu , Dong Yu

Multi-channel acoustic signal processing is a well-established and powerful tool to exploit the spatial diversity between a target signal and non-target or noise sources for signal enhancement. However, the textbook solutions for optimal…

音频与语音处理 · 电气工程与系统科学 2025-01-14 Reinhold Haeb-Umbach , Tomohiro Nakatani , Marc Delcroix , Christoph Boeddeker , Tsubasa Ochiai

Personal sound zone (PSZ) reproduction system, which attempts to create distinct virtual acoustic scenes for different listeners at their respective positions within the same spatial area using one loudspeaker array, is a fundamental…

声音 · 计算机科学 2025-12-12 Wenye Zhu , Jun Tang , Xiaofei Li

An acoustic metamaterial made of a two-dimensional (2D) periodic array of multi-resonant acoustic scatterers is analysed in this paper. The building blocks consist of a combination of elastic beams of Low-Density Polyethylene Foam (LDPF)…

材料科学 · 物理学 2011-04-01 V. Romero-García , A. Krynkin , L. M. Garcia-Raffi , O. Umnova , J. V. Sánchez-Pérez

With the recent surge of video conferencing tools usage, providing high-quality speech signals and accurate captions have become essential to conduct day-to-day business or connect with friends and families. Single-channel personalized…

音频与语音处理 · 电气工程与系统科学 2021-10-22 Hassan Taherian , Sefik Emre Eskimez , Takuya Yoshioka , Huaming Wang , Zhuo Chen , Xuedong Huang

Extremely large-scale array (XL-array) has emerged as a promising technology to enable near-field communications for achieving enhanced spectrum efficiency and spatial resolution, by drastically increasing the number of antennas. However,…

信息论 · 计算机科学 2026-01-21 Cong Zhou , Changsheng You , Haodong Zhang , Li Chen , Shuo Shi

Spatial analysis of room acoustics is an ongoing research topic. Microphone arrays have been employed for spatial analyses with an important objective being the estimation of the direction-of-arrival (DOA) of direct sound and early room…

音频与语音处理 · 电气工程与系统科学 2024-01-09 Hai Morgenstern , Boaz Rafaely

One of the biggest challenges of acoustic scene classification (ASC) is to find proper features to better represent and characterize environmental sounds. Environmental sounds generally involve more sound sources while exhibiting less…

声音 · 计算机科学 2019-04-11 Hongwei Song , Jiqing Han , Shiwen Deng

Methods are proposed for modifying the reverberation characteristics of sound fields in rooms by employing a loudspeaker with adjustable directivity, realized with a compact spherical loudspeaker array (SLA). These methods are based on…

音频与语音处理 · 电气工程与系统科学 2024-01-09 Hai Morgenstern , Boaz Rafaely

Acoustic local positioning systems (ALPSs) are an interesting alternative for indoor positioning due to certain advantages over other approaches, including their relatively high accuracy, low cost, and room-level signal propagation.…

A combination of cloud-based deep learning (DL) algorithms with portable/wearable (P/W) devices has been developed as a smart heath care system to support automatic cardiac arrhythmias (CAs) classification using electrocardiography (ECG).…

信号处理 · 电气工程与系统科学 2023-08-22 Tsai-Min Chen , Yuan-Hong Tsai , Huan-Hsin Tseng , Kai-Chun Liu , Jhih-Yu Chen , Chih-Han Huang , Guo-Yuan Li , Chun-Yen Shen , Yu Tsao

Far-infrared detectors for future cooled space telescopes require ultra-sensitive detectors with optical noise equivalent powers of order 0.2 aW/\sqrt Hz. This performance has already been demonstrated in arrays of transition edge sensors.…

天体物理仪器与方法 · 物理学 2022-07-01 D J Goldie , S. Withington , C. N. Thomas , P. A. R. Ade , R. V. Sudiwala

Wearable devices like smart glasses are approaching the compute capability to seamlessly generate real-time closed captions for live conversations. We build on our recently introduced directional Automatic Speech Recognition (ASR) for smart…

音频与语音处理 · 电气工程与系统科学 2024-01-22 Ju Lin , Niko Moritz , Yiteng Huang , Ruiming Xie , Ming Sun , Christian Fuegen , Frank Seide

Active speaker detection (ASD) is a multi-modal task that aims to identify who, if anyone, is speaking from a set of candidates. Current audio-visual approaches for ASD typically rely on visually pre-extracted face tracks (sequences of…

音频与语音处理 · 电气工程与系统科学 2022-03-08 Davide Berghi , Adrian Hilton , Philip J. B. Jackson

Acoustic scene classification (ASC) aims to identify the type of scene (environment) in which a given audio signal is recorded. The log-mel feature and convolutional neural network (CNN) have recently become the most popular time-frequency…

声音 · 计算机科学 2021-08-12 Yuzhong Wu , Tan Lee

We propose an enhanced spatial modulation (SM)-based scheme for indoor visible light communication systems. This scheme enhances the achievable throughput of conventional SM schemes by transmitting higher order complex modulation symbol,…

信号处理 · 电气工程与系统科学 2022-01-21 Shimaa Naser , Lina Bariah , Sami Muhaidat , Mahmoud Al-Qutayri , Paschalis C. Sofotasios

Acoustic lenses are employed in a variety of applications, from biomedical imaging and surgery, to defense systems, but their performance is limited by their linear operational envelope and complexity. Here we show a dramatic focusing…

软凝聚态物质 · 物理学 2015-05-14 Alessandro Spadoni , Chiara Daraio

We introduce BANC, a neural binaural audio codec designed for efficient speech compression in single and two-speaker scenarios while preserving the spatial location information of each speaker. Our key contributions are as follows: 1) The…

声音 · 计算机科学 2024-11-26 Anton Ratnarajah , Shi-Xiong Zhang , Dong Yu

The use of spatial information with multiple microphones can improve far-field automatic speech recognition (ASR) accuracy. However, conventional microphone array techniques degrade speech enhancement performance when there is an array…

音频与语音处理 · 电气工程与系统科学 2021-12-23 Kenichi Kumatani , Minhua Wu , Shiva Sundaram , Nikko Strom , Bjorn Hoffmeister

Contrastive self-supervised learning (CSL) for speaker verification (SV) has drawn increasing interest recently due to its ability to exploit unlabeled data. Performing data augmentation on raw waveforms, such as adding noise or…

音频与语音处理 · 电气工程与系统科学 2024-03-12 Chong-Xin Gan , Man-Wai Mak , Weiwei Lin , Jen-Tzung Chien