中文
相关论文

相关论文: Ambisonics Encoder for Wearable Array with Improve…

200 篇论文

Upcoming ground-based cosmic microwave background experiments will provide CMB maps with high sensitivity and resolution that can be used for high fidelity lensing reconstruction. However, the sky coverage will be incomplete and the noise…

宇宙学与河外天体物理 · 物理学 2019-12-11 Mark Mirmelstein , Julien Carron , Antony Lewis

Generating radiology reports automatically reduces the workload of radiologists and helps the diagnoses of specific diseases. Many existing methods take this task as modality transfer process. However, since the key information related to…

计算机视觉与模式识别 · 计算机科学 2024-04-02 Yitian Tao , Liyan Ma , Jing Yu , Han Zhang

Spatially selective active noise control (SSANC) hearables aim to attenuate noise from certain directions at the eardrum while preserving desired speech arriving from selected directions. Existing SSANC systems typically assume an accurate…

音频与语音处理 · 电气工程与系统科学 2026-05-19 Tong Xiao , Reinhild Roden , Matthias Blau , Simon Doclo

In embedding-matching acoustic-to-word (A2W) ASR, every word in the vocabulary is represented by a fixed-dimension embedding vector that can be added or removed independently of the rest of the system. The approach is potentially an elegant…

音频与语音处理 · 电气工程与系统科学 2023-02-21 Hao Yen , Woojay Jeon

Earables, such as True Wireless Stereo earphones and VR/AR headsets, are increasingly popular, yet their compact design poses challenges for robust voice-related applications like telecommunication and voice assistant interactions in noisy…

声音 · 计算机科学 2025-12-03 Lixing He , Yunqi Guo , Haozheng Hou , Zhenyu Yan

Multichannel frequency estimation with incomplete data and miscellaneous noises arises in array signal processing, modal analysis, wireless communications, and so on. In this paper, we consider maximum-likelihood(-like) optimization methods…

信号处理 · 电气工程与系统科学 2023-07-18 Xunmeng Wu , Zai Yang , Zongben Xu

Integrated Sensing and Communication (ISAC) emerges as a promising technology for B5G/6G, particularly in the millimeter-wave (mmWave) band. However, the widespread adoption of hybrid architecture in mmWave systems compromises multiplexing…

信号处理 · 电气工程与系统科学 2024-03-18 Boxun Liu , Shijian Gao , Zonghui Yang , Xiang Cheng

Speech enhancement in hearing aids remains a difficult task in nonstationary acoustic environments, mainly because current signal processing algorithms rely on fixed, manually tuned parameters that cannot adapt in situ to different users or…

In this paper, we propose a novel separation system for extracting two speech signals from two microphone recordings. Our system combines the blind source separation technique with cepstral smoothing of binary time-frequency masks. The last…

声音 · 计算机科学 2026-03-17 Ibrahim Missaoui , Zied Lachiri

We propose a new method, {\it binary fused compressive sensing} (BFCS), to recover sparse piece-wise smooth signals from 1-bit compressive measurements. The proposed algorithm is a modification of the previous {\it binary iterative hard…

计算机视觉与模式识别 · 计算机科学 2014-02-21 Xiangrong Zeng , Mário A. T. Figueiredo

Sound field decomposition predicts waveforms in arbitrary directions using signals from a limited number of microphones as inputs. Sound field decomposition is fundamental to downstream tasks, including source localization, source…

声音 · 计算机科学 2022-10-25 Qiuqiang Kong , Shilei Liu , Junjie Shi , Xuzhou Ye , Yin Cao , Qiaoxi Zhu , Yong Xu , Yuxuan Wang

Signal decomposition is a classical problem in signal processing, which aims to separate an observed signal into two or more components each with its own property. Usually each component is described by its own subspace or dictionary.…

计算机视觉与模式识别 · 计算机科学 2018-12-27 Shervin Minaee , Yao Wang

Spherical Harmonics ROOM), an open-source Python library for room acoustics simulation using Ambisonics, available at https://github.com/Yhonatangayer/shroom and installable via \texttt{pip install pyshroom}. \textbf{shroom} projects…

音频与语音处理 · 电气工程与系统科学 2026-03-31 Yhonatan Gayer

In this paper, the concept of an adaptation algorithm is proposed, which can be used to blindly adapt the microphone array geometry of a humanoid robot such that the performance of the underlying signal separation algorithm is improved. As…

声音 · 计算机科学 2014-04-29 Hendrik Barfuss , Walter Kellermann

Automatic speech recognition (ASR) in the cloud allows the use of larger models and more powerful multi-channel signal processing front-ends compared to on-device processing. However, it also adds an inherent latency due to the transmission…

音频与语音处理 · 电气工程与系统科学 2021-06-16 Lukas Drude , Jahn Heymann , Andreas Schwarz , Jean-Marc Valin

Spherical microphone arrays (SMAs) are widely used for sound field analysis, and sparse recovery (SR) techniques can significantly enhance their spatial resolution by modeling the sound field as a sparse superposition of dominant plane…

音频与语音处理 · 电气工程与系统科学 2025-09-05 Shunxi Xu , Craig T. Jin

Recovering texture information from the aliasing regions has always been a major challenge for Single Image Super Resolution (SISR) task. These regions are often submerged in noise so that we have to restore texture details while…

图像与视频处理 · 电气工程与系统科学 2021-09-13 Fanyi Wang , Haotian Hu , Cheng Shen

Active noise control (ANC) systems are commonly designed to achieve maximal sound reduction regardless of the incident direction of the sound. When desired sound is present, the state-of-the-art methods add a separate system to reconstruct…

音频与语音处理 · 电气工程与系统科学 2023-05-15 Tong Xiao , Buye Xu , Chuming Zhao

Synthetic aperture sonar (SAS) measures a scene from multiple views in order to increase the resolution of reconstructed imagery. Image reconstruction methods for SAS coherently combine measurements to focus acoustic energy onto the scene.…

图像与视频处理 · 电气工程与系统科学 2023-06-19 Albert W. Reed , Juhyeon Kim , Thomas Blanford , Adithya Pediredla , Daniel C. Brown , Suren Jayasuriya

Subword units are commonly used for end-to-end automatic speech recognition (ASR), while a fully acoustic-oriented subword modeling approach is somewhat missing. We propose an acoustic data-driven subword modeling (ADSM) approach that…

计算与语言 · 计算机科学 2023-10-24 Wei Zhou , Mohammad Zeineldeen , Zuoyun Zheng , Ralf Schlüter , Hermann Ney