English
Related papers

Related papers: F-T-LSTM based Complex Network for Joint Acoustic …

200 papers

Despite the potential of diffusion models in speech enhancement, their deployment in Acoustic Echo Cancellation (AEC) has been restricted. In this paper, we propose DI-AEC, pioneering a diffusion-based stochastic regeneration approach…

Audio and Speech Processing · Electrical Eng. & Systems 2024-01-10 Yang Liu , Li Wan , Yun Li , Yiteng Huang , Ming Sun , James Luan , Yangyang Shi , Xin Lei

The successful deployment of deep learning-based acoustic echo and noise reduction (AENR) methods in consumer devices has spurred interest in developing low-complexity solutions, while emphasizing the need for robust performance in…

Audio and Speech Processing · Electrical Eng. & Systems 2025-08-05 Shrishti Saha Shetu , Naveen Kumar Desiraju , Wolfgang Mack , Emanuël A. P. Habets

Acoustic echo cancellation (AEC) aims to remove interference signals while leaving near-end speech least distorted. As the indistinguishable patterns between near-end speech and interference signals, near-end speech can't be separated…

Audio and Speech Processing · Electrical Eng. & Systems 2023-07-27 Chang Han , Xinmeng Xu , Weiping Tu , Yuhong Yang , Yajie Liu

This paper introduces a dual-signal transformation LSTM network (DTLN) for real-time speech enhancement as part of the Deep Noise Suppression Challenge (DNS-Challenge). This approach combines a short-time Fourier transform (STFT) and a…

Audio and Speech Processing · Electrical Eng. & Systems 2020-10-23 Nils L. Westhausen , Bernd T. Meyer

Time delay estimation (TDE) plays a key role in acoustic echo cancellation (AEC) using adaptive filter method. Considerable residual echo will be left if estimation error arises. Here, in this paper, we proposed an adaptive filter bank…

Sound · Computer Science 2025-02-11 Lu Ma

With the development of society, time series anomaly detection plays an important role in network and IoT services. However, most existing anomaly detection methods directly analyze time series in the time domain and cannot distinguish some…

Artificial Intelligence · Computer Science 2024-12-04 Yi-Xiang Lu , Xiao-Bo Jin , Jian Chen , Dong-Jie Liu , Guang-Gang Geng

Deep neural networks (DNNs) have shown promising results for acoustic echo cancellation (AEC). But the DNN-based AEC models let through all near-end speakers including the interfering speech. In light of recent studies on personalized…

Sound · Computer Science 2022-07-01 Shimin Zhang , Ziteng Wang , Yukai Ju , Yihui Fu , Yueyue Na , Qiang Fu , Lei Xie

End-to-end Automatic Speech Recognition (ASR) systems based on neural networks have seen large improvements in recent years. The availability of large scale hand-labeled datasets and sufficient computing resources made it possible to train…

Computer Vision and Pattern Recognition · Computer Science 2023-01-05 Maxime Burchi , Radu Timofte

Acoustic echo cancellation (AEC), noise suppression (NS) and dereverberation (DR) are an integral part of modern full-duplex communication systems. As the demand for teleconferencing systems increases, addressing these tasks is required for…

In this paper, we propose a residual echo suppression method using a UNet neural network that directly maps the outputs of a linear acoustic echo canceler to the desired signal in the spectral domain. This system embeds a design parameter…

Sound · Computer Science 2021-06-28 Amir Ivry , Israel Cohen , Baruch Berdugo

Multivariate Time-Series (MTS) clustering is crucial for signal processing and data analysis. Although deep learning approaches, particularly those leveraging Contrastive Learning (CL), are prominent for MTS representation, existing…

Machine Learning · Computer Science 2026-01-13 Zexi Tan , Tao Xie , Haoyi Xiao , Baoyao Yang , Yuzhu Ji , An Zeng , Xiang Zhang , Yiqun Zhang

Building on the deep learning based acoustic echo cancellation (AEC) in the single-loudspeaker (single-channel) and single-microphone setup, this paper investigates multi-channel AEC (MCAEC) and multi-microphone AEC (MMAEC). We train a deep…

Audio and Speech Processing · Electrical Eng. & Systems 2021-03-04 Hao Zhang , DeLiang Wang

Although today's speech communication systems support various bandwidths from narrowband to super-wideband and beyond, state-of-the art DNN methods for acoustic echo cancellation (AEC) are lacking modularity and bandwidth scalability. Our…

Audio and Speech Processing · Electrical Eng. & Systems 2022-11-08 Ernst Seidel , Rasmus Kongsgaard Olsson , Karim Haddad , Zhengyang Li , Pejman Mowlaee , Tim Fingscheidt

To achieve robust far-field automatic speech recognition (ASR), existing techniques typically employ an acoustic front end (AFE) cascaded with a neural transducer (NT) ASR model. The AFE output, however, could be unreliable, as the…

This work presents a statistical analysis of a class of jointly optimized beamformer-assisted acoustic echo cancelers (AEC) with the beamformer (BF) implemented in the Generalized Sidelobe Canceler (GSC) form and using the least-mean square…

Statistics Theory · Mathematics 2015-03-06 Marcos H. Maruo , José C. M. Bermudez , Leonardo S. Resende

Automatic speech recognition (ASR) tasks are resolved by end-to-end deep learning models, which benefits us by less preparation of raw data, and easier transformation between languages. We propose a novel end-to-end deep learning model…

Audio and Speech Processing · Electrical Eng. & Systems 2018-10-31 Xinpei Zhou , Jiwei Li , Xi Zhou

In this paper, we propose long short term memory speech enhancement network (LSTMSE-Net), an audio-visual speech enhancement (AVSE) method. This innovative method leverages the complementary nature of visual and audio information to boost…

Fundamental frequency is one of the most important parameters of human speech, of importance for the classification of accent, gender, speaking styles, speaker identification, age, among others. The proper detection of this parameter…

Sound · Computer Science 2019-11-13 Marvin Coto-Jimenez

Most digital audio tampering detection methods based on electrical network frequency (ENF) only utilize the static spatial information of ENF, ignoring the variation of ENF in time series, which limit the ability of ENF feature…

Sound · Computer Science 2022-08-26 Chunyan Zeng , Shuai Kong , Zhifeng Wang , Xiangkui Wan , Yunfan Chen

Deep learning-based methods that jointly perform the task of acoustic echo and noise reduction (AENR) often require high memory and computational resources, making them unsuitable for real-time deployment on low-resource platforms such as…

Audio and Speech Processing · Electrical Eng. & Systems 2024-08-29 Shrishti Saha Shetu , Naveen Kumar Desiraju , Jose Miguel Martinez Aponte , Emanuël A. P. Habets , Edwin Mabande