English
Related papers

Related papers: Improved in-car sound pick-up using multichannel W…

200 papers

This work describes a speech denoising system for machine ears that aims to improve speech intelligibility and the overall listening experience in noisy environments. We recorded approximately 100 hours of audio data with reverberation and…

Audio and Speech Processing · Electrical Eng. & Systems 2022-02-18 Cong Han , E. Merve Kaya , Kyle Hoefer , Malcolm Slaney , Simon Carlile

In a world increasingly dependent on road-based transportation, it is essential to understand vehicles. We introduce the AI mechanic, an acoustic vehicle characterization deep learning system, as an integrated approach using sound captured…

Sound · Computer Science 2022-05-20 Adam M. Terwilliger , Joshua E. Siegel

In this paper, a practical model for non-stationary Vehicle-to-Vehicle (V2V) multiple-input multiple-output (MIMO) channels is proposed. The new model considers more accurate output phase of Doppler frequency and is simplified by the Taylor…

Signal Processing · Electrical Eng. & Systems 2020-02-04 Weidong Li , Qiuming Zhu , Cheng-Xiang Wang , Fei Bai , Xiaomin Chen , Dazhuan Xu

Current disfluency detection models focus on individual utterances each from a single speaker. However, numerous discontinuity phenomena in spoken conversational transcripts occur across multiple turns, hampering human readability and the…

Computation and Language · Computer Science 2023-10-30 Hua Shen , Vicky Zayats , Johann C. Rocholl , Daniel D. Walker , Dirk Padfield

This paper proposes a model that integrates sub-band processing and deep filtering to fully exploit information from the target time-frequency (TF) bin and its surrounding TF bins for single-channel speech enhancement. The sub-band module…

Sound · Computer Science 2025-06-03 Shenghui Lu , Hukai Huang , Jinanglong Yao , Kaidi Wang , Qingyang Hong , Lin Li

Sophisticated user interaction in the automotive industry is a fast emerging topic. Mid-air gestures and speech already have numerous applications for driver-car interaction. Additionally, multimodal approaches are being developed to…

Human-Computer Interaction · Computer Science 2020-12-29 Abdul Rafey Aftab , Michael von der Beeck , Michael Feld

With the proliferation of video platforms on the internet, recording musical performances by mobile devices has become commonplace. However, these recordings often suffer from degradation such as noise and reverberation, which negatively…

Sound · Computer Science 2023-08-25 Yunkee Chae , Junghyun Koo , Sungho Lee , Kyogu Lee

Microphone array techniques are widely used in sound source localization and smart city acoustic-based traffic monitoring, but these applications face significant challenges due to the scarcity of labeled real-world traffic audio data and…

Audio and Speech Processing · Electrical Eng. & Systems 2024-12-30 Shitong Fan , Feiyang Xiao , Wenbo Wang , Shuhan Qi , Qiaoxi Zhu , Wenwu Wang , Jian Guan

Consumer electronic (CE) devices increasingly rely on wireless local area networks (WLANs). Next generation WLANs will continue to exploit multiple antenna systems to satisfy the growing need for WLAN system capacity. Multiple-input…

Signal Processing · Electrical Eng. & Systems 2018-12-05 Xiaofu Ma , Qinghai Gao , Ji Wang , Vuk Marojevic , Jeffrey H. Reed

Inspired by the fact that humans use diverse sensory organs to perceive the world, sensors with different modalities are deployed in end-to-end driving to obtain the global context of the 3D scene. In previous works, camera and LiDAR inputs…

Computer Vision and Pattern Recognition · Computer Science 2022-08-04 Qingwen Zhang , Mingkai Tang , Ruoyu Geng , Feiyi Chen , Ren Xin , Lujia Wang

We present bounds and a closed-form high-SNR expression for the capacity of multiple-antenna systems affected by Wiener phase noise. Our results are developed for the scenario where a single oscillator drives all the radio-frequency…

Information Theory · Computer Science 2016-11-17 Giuseppe Durisi , Alberto Tarable , Christian Camarda , Rahul Devassy , Guido Montorsi

This paper presents an acoustic impedance control architecture for an electroacoustic absorber combining both a feedforward and a feedback microphone-based strategies on a current-driven loudspeaker. Feedforward systems enable good…

Systems and Control · Electrical Eng. & Systems 2026-01-08 Maxime Volery , Xinxin Guo , Hervé Lissek

Separating different speaker properties from a multi-speaker environment is challenging. Instead of separating a two-speaker signal in signal space like speech source separation, a speaker embedding de-mixing approach is proposed. The…

Sound · Computer Science 2021-02-08 Yanpei Shi , Thomas Hain

The present work deals with a new passive system for real-time detection, classification and direction of arrival estimator of Unmanned Aerial Vehicles (UAVs). The proposed system composed of a very low cost hardware components, comprises…

Audio and Speech Processing · Electrical Eng. & Systems 2019-03-01 Konstantinos Polyzos , Evangelos Dermatas

Current multichannel speech enhancement algorithms typically assume a stationary sound source, a common mismatch with reality that limits their performance in real-world scenarios. This paper focuses on attention-driven spatial filtering…

Audio and Speech Processing · Electrical Eng. & Systems 2023-12-19 Yuzhu Wang , Archontis Politis , Tuomas Virtanen

The success of nonlinear noise reduction applied to a single channel recording of human voice is measured in terms of the recognition rate of a commercial speech recognition program in comparison to the optimal linear filter. The overall…

Data Analysis, Statistics and Probability · Physics 2007-06-20 Krzysztof Urbanowicz , Holger Kantz

Target speech separation refers to extracting the target speaker's speech from mixed signals. Despite the recent advances in deep learning based close-talk speech separation, the applications to real-world are still an open issue. Two main…

Sound · Computer Science 2020-01-03 Rongzhi Gu , Yuexian Zou

Speech pre-processing techniques such as denoising, de-reverberation, and separation, are commonly employed as front-ends for various downstream speech processing tasks. However, these methods can sometimes be inadequate, resulting in…

Audio and Speech Processing · Electrical Eng. & Systems 2025-06-17 Sirui Li , Shuai Wang , Zhijun Liu , Zhongjie Jiang , Yannan Wang , Haizhou Li

Speaker counting is the task of estimating the number of people that are simultaneously speaking in an audio recording. For several audio processing tasks such as speaker diarization, separation, localization and tracking, knowing the…

Sound · Computer Science 2021-01-07 Pierre-Amaury Grumiaux , Srdan Kitic , Laurent Girin , Alexandre Guérin

Recent studies have demonstrated that incorporating auxiliary information, such as speaker voiceprint or visual cues, can substantially improve Speech Enhancement (SE) performance. However, single-channel methods often yield suboptimal…

Audio and Speech Processing · Electrical Eng. & Systems 2026-03-06 Chihyun Liu , Jiaxuan Fan , Mingtung Sun , Michael Anthony , Mingsian R. Bai , Yu Tsao