中文
相关论文

相关论文: Localization based on enhanced low frequency inter…

200 篇论文

Microphone arrays are usually assumed to have rigid geometries: the microphones may move with respect to the sound field but remain fixed relative to each other. However, many useful arrays, such as those in wearable devices, have sensors…

音频与语音处理 · 电气工程与系统科学 2019-12-12 Ryan M. Corey , Andrew C. Singer

Indoor localization is a long-standing challenge in mobile computing, with significant implications for enabling location-aware and intelligent applications within smart environments such as homes, offices, and retail spaces. As AI…

音频与语音处理 · 电气工程与系统科学 2025-08-26 Amod K. Agrawal

Deep learning has dramatically improved the performance of speech recognition systems through learning hierarchies of features optimized for the task at hand. However, true end-to-end learning, where features are learned directly from…

计算与语言 · 计算机科学 2016-04-06 Zhenyao Zhu , Jesse H. Engel , Awni Hannun

Networked integrated sensing and communication (ISAC) has gained significant attention as a promising technology for enabling next-generation wireless systems. To further enhance networked ISAC, delegating the reception of sensing signals…

信号处理 · 电气工程与系统科学 2025-10-14 Meidong Xia , Zhenyao He , Wei Xu , Yongming Huang , Derrick Wing Kwan Ng , Naofal Al-Dhahir

Enabling multi-target sensing in near-field integrated sensing and communication (ISAC) systems is a key challenge, particularly when line-of-sight paths are blocked. This paper proposes a beamforming framework that leverages a…

信号处理 · 电气工程与系统科学 2025-09-11 Hang Ruan , Homa Nikbakht , Ruizhi Zhang , Honglei Chen , Yonina C. Eldar

Transformer-based visual object tracking has been utilized extensively. However, the Transformer structure is lack of enough inductive bias. In addition, only focusing on encoding the global feature does harm to modeling local details,…

计算机视觉与模式识别 · 计算机科学 2022-08-09 Changhong Fu , Weiyu Peng , Sihang Li , Junjie Ye , Ziang Cao

A novel end-to-end binaural sound localisation approach is proposed which estimates the azimuth of a sound source directly from the waveform. Instead of employing hand-crafted features commonly employed for binaural sound localisation, such…

声音 · 计算机科学 2019-04-04 Paolo Vecchiotti , Ning Ma , Stefano Squartini , Guy J. Brown

To meet the requirements for high data rates and ubiquitous connectivity in 6G networks, higher frequencies and larger array apertures are employed to enhance spatial resolution and spectral efficiency. This evolution leads to an expansion…

信号处理 · 电气工程与系统科学 2026-04-03 Shupei Zhang , Boya Di , Lingyang Song

In this letter, we investigate a robust beamforming problem for localization-aided millimeter wave (mmWave) communication systems. To handle this problem, we propose a novel restriction and relaxation (R&R) method. The proposed R&R method…

信息论 · 计算机科学 2022-04-01 Junchang Sun , Shuai Ma , Shiyin Li , Ruixin Yang , Minghui Min , Gonzalo Seco-Granados

This paper investigates a practical partially-connected hybrid beamforming transmitter for integrated sensing and communication (ISAC) with distortion from nonlinear power amplification. For this ISAC system, we formulate a communication…

信号处理 · 电气工程与系统科学 2025-07-21 Zeyuan Zhang , Yue Xiu , Phee Lep Yeoh , Guangyi Liu , Zixing Wu , Ning Wei

This letter introduces a novel speech enhancement method in the Hilbert-Huang Transform domain to mitigate the effects of acoustic impulsive noises. The estimation and selection of noise components is based on the impulsiveness index of…

音频与语音处理 · 电气工程与系统科学 2019-10-08 C. Medina , R. Coelho

Modern day audio signal classification techniques lack the ability to classify low feature audio signals in the form of spectrographic temporal frequency data representations. Additionally, currently utilized techniques rely on full diverse…

声音 · 计算机科学 2024-10-30 Noel Elias

Using more test-time computation during language model inference, such as generating more intermediate thoughts or sampling multiple candidate answers, has proven effective in significantly improving model performance. This paper takes an…

机器学习 · 计算机科学 2025-08-20 Xingwu Chen , Miao Lu , Beining Wu , Difan Zou

Binaural audio gives the listener the feeling of being in the recording place and enhances the immersive experience if coupled with AR/VR. But the problem with binaural audio recording is that it requires a specialized setup which is not…

声音 · 计算机科学 2021-08-12 Kranti Kumar Parida , Siddharth Srivastava , Neeraj Matiyali , Gaurav Sharma

Recently, our proposed recurrent neural network (RNN) based all deep learning minimum variance distortionless response (ADL-MVDR) beamformer method yielded superior performance over the conventional MVDR by replacing the matrix inversion…

声音 · 计算机科学 2021-04-27 Xiyun Li , Yong Xu , Meng Yu , Shi-Xiong Zhang , Jiaming Xu , Bo Xu , Dong Yu

It is a well-established principle that cross-correlating seismic observations at different receiver locations can yield estimates of band-limited inter-receiver Green's functions. This principle, known as seismic interferometry, is a…

地球物理 · 物理学 2021-08-11 Daniella Ayala-Garcia , Andrew Curtis , Michal Branicki

Speech enhancement in hearing aids is a challenging task since the hardware limits the number of possible operations and the latency needs to be in the range of only a few milliseconds. We propose a deep-learning model compatible with these…

音频与语音处理 · 电气工程与系统科学 2023-07-19 Nils L. Westhausen , Bernd T. Meyer

Binaural Audio Telepresence (BAT) aims to encode the acoustic scene at the far end into binaural signals for the user at the near end. BAT encompasses an immense range of applications that can vary between two extreme modes of Immersive BAT…

音频与语音处理 · 电气工程与系统科学 2024-05-15 Yicheng Hsu , Mingsian R. Bai

Transformer-based models have gained increasing popularity achieving state-of-the-art performance in many research fields including speech translation. However, Transformer's quadratic complexity with respect to the input sequence length…

计算与语言 · 计算机科学 2023-10-19 Sara Papi , Marco Gaido , Matteo Negri , Marco Turchi

In multi-channel speech enhancement and robust automatic speech recognition (ASR), beamforming can typically improve the signal-to-noise ratio (SNR) of the target speaker and produce reliable enhancement with little distortion to target…

音频与语音处理 · 电气工程与系统科学 2025-07-22 Zhong-Qiu Wang , Ruizhe Pang