中文
相关论文

相关论文: Localizing Spatial Information in Neural Spatiospe…

200 篇论文

This paper proposes a neural network based speech separation method using spatially distributed microphones. Unlike with traditional microphone array settings, neither the number of microphones nor their spatial arrangement is known in…

音频与语音处理 · 电气工程与系统科学 2020-05-01 Dongmei Wang , Zhuo Chen , Takuya Yoshioka

Automotive radar perception pipelines commonly construct angle-domain representations via beamforming before applying learning-based models. This work instead investigates a representational question: can meaningful spatial structure be…

计算机视觉与模式识别 · 计算机科学 2026-04-03 George Sebastian , Philipp Berthold , Bianca Forkel , Leon Pohl , Mirko Maehlisch

Speech separation with several speakers is a challenging task because of the non-stationarity of the speech and the strong signal similarity between interferent sources. Current state-of-the-art solutions can separate well the different…

信号处理 · 电气工程与系统科学 2021-02-09 Nicolas Furnon , Romain Serizel , Irina Illina , Slim Essid

Vision Transformers face a fundamental limitation: standard self-attention jointly processes spatial and channel dimensions, leading to entangled representations that prevent independent modeling of structural and semantic dependencies.…

计算机视觉与模式识别 · 计算机科学 2025-12-05 Jiashu Liao , Pietro Liò , Marc de Kamps , Duygu Sarikaya

Multi-channel speech separation in dynamic environments is challenging as time-varying spatial and spectral features evolve at different temporal scales. Existing methods typically employ sequential architectures, forcing a single network…

音频与语音处理 · 电气工程与系统科学 2026-02-27 Yuzhu Wang , Archontis Politis , Konstantinos Drossos , Tuomas Virtanen

Fine-grained high-resolution remote sensing mapping typically relies on localized visual features, which restricts cross-domain generalizability and often leads to fragmented predictions of large-scale land covers. While global geospatial…

计算机视觉与模式识别 · 计算机科学 2026-04-23 Jienan Lyu , Miao Yang , Jinchen Cai , Yiwen Hu , Guanyi Lu , Junhao Qiu , Runmin Dong

Deep learning-based techniques for automatic dysarthric speech detection have recently attracted interest in the research community. State-of-the-art techniques typically learn neurotypical and dysarthric discriminative representations by…

音频与语音处理 · 电气工程与系统科学 2021-10-04 Ina Kodrasi

We propose a method to improve subject transfer in motor imagery BCIs by aligning covariance matrices on a Riemannian manifold, followed by computing a new common spatial patterns (CSP) based spatial filter. We explore various ways to…

计算机视觉与模式识别 · 计算机科学 2025-04-25 Tekin Gunasar , Virginia de Sa

The paper studies the problem of designing the Intelligent Reflecting Surface (IRS) phase shifters for Multiple Input Single Output (MISO) communication systems in spatiotemporally correlated channel environments, where the destination can…

信息论 · 计算机科学 2022-11-18 Spilios Evmorfos , Athina P. Petropulu , H. Vincent Poor

Cognitive beamforming (CB) is a multi-antenna technique for efficient spectrum sharing between primary users (PUs) and secondary users (SUs) in a cognitive radio network. Specifically, a multi-antenna SU transmitter applies CB to suppress…

信息论 · 计算机科学 2010-10-18 Kaibin Huang , Rui Zhang

This work addresses channel estimation in multiple antenna multicell interference-limited networks. Channel state information (CSI) acquisition is vital for interference mitigation. Wireless networks often suffer from multicell…

信息论 · 计算机科学 2015-05-11 Maha Alodeh , Symeon Chatzinotas , Bjorn Ottersten

Personalized speech enhancement has been a field of active research for suppression of speechlike interferers such as competing speakers or TV dialogues. Compared with single channel approaches, multichannel PSE systems can be more…

音频与语音处理 · 电气工程与系统科学 2022-11-17 Yicheng Hsu , Yonghan Lee , Mingsian R. Bai

The method of Common Spatial Patterns (CSP) is widely used for feature extraction of electroencephalography (EEG) data, such as in motor imagery brain-computer interface (BCI) systems. It is a data-driven method estimating a set of spatial…

信号处理 · 电气工程与系统科学 2022-02-10 Mahta Mousavi , Eric Lybrand , Shuangquan Feng , Shuai Tang , Rayan Saab , Virginia de Sa

Providing guaranteed quality of service for cell-edge users remains a longstanding challenge in wireless networks. While coordinated interference management was proposed decades ago, its potential has been limited by computational…

信号处理 · 电气工程与系统科学 2026-03-31 Tenghao Cai , Lei Li , Shutao Zhang , Tsung-Hui Chang

Mesh is a powerful data structure for 3D shapes. Representation learning for 3D meshes is important in many computer vision and graphics applications. The recent success of convolutional neural networks (CNNs) for structured data (e.g.,…

计算机视觉与模式识别 · 计算机科学 2021-11-29 Zhongpai Gao , Junchi Yan , Guangtao Zhai , Juyong Zhang , Yiyan Yang , Xiaokang Yang

Speaker tracking methods often rely on spatial observations to assign coherent track identities over time. This raises limits in scenarios with intermittent and moving speakers, i.e., speakers that may change position when they are…

音频与语音处理 · 电气工程与系统科学 2025-06-26 Taous Iatariene , Can Cui , Alexandre Guérin , Romain Serizel

This paper presents an improved deep embedding learning method based on convolutional neural network (CNN) for text-independent speaker verification. Two improvements are proposed for x-vector embedding learning: (1) Multi-scale convolution…

音频与语音处理 · 电气工程与系统科学 2020-01-15 Bin Gu , Wu Guo

End-to-end transformer-based automatic speech recognition (ASR) systems often capture multiple speech traits in their learned representations that are highly entangled, leading to a lack of interpretability. In this study, we propose the…

音频与语音处理 · 电气工程与系统科学 2024-11-28 Pu Wang , Hugo Van hamme

Deep learning architectures have made significant progress in terms of performance in many research areas. The automatic speech recognition (ASR) field has thus benefited from these scientific and technological advances, particularly for…

声音 · 计算机科学 2024-03-01 Quentin Raymondaud , Mickael Rouvier , Richard Dufour

Sound field reproduction with undistorted sound quality and precise spatial localization is desirable for automotive audio systems. However, the complexity of automotive cabin acoustic environment often necessitates a trade-off between…

音频与语音处理 · 电气工程与系统科学 2025-09-12 Yufan Qian , Tianshu Qu , Xihong Wu