中文
相关论文

相关论文: Incremental Averaging Method to Improve Graph-Base…

200 篇论文

The voting method, an ensemble approach for fundamental frequency estimation, is empirically known for its robustness but lacks thorough investigation. This paper provides a principled analysis and improvement of this technique. First, we…

声音 · 计算机科学 2026-02-03 Junya Koguchi , Tomoki Koriyama

Speech dereverberation aims to alleviate the negative impact of late reverberant reflections. The weighted prediction error (WPE) method is a well-established technique known for its superior performance in dereverberation. However, in…

音频与语音处理 · 电气工程与系统科学 2023-12-07 Ziye Yang , Mengfei Zhang , Jie Chen

Direction of arrival (DoA) estimation is a fundamental task in array processing. A popular family of DoA estimation algorithms are subspace methods, which operate by dividing the measurements into distinct signal and noise subspaces.…

信号处理 · 电气工程与系统科学 2024-07-12 Dor H. Shmuel , Julian P. Merkofer , Guy Revach , Ruud J. G. van Sloun , Nir Shlezinger

The large-scale integration of inverter-based resources (IBRs), particularly distributed photovoltaics (DPVs), into distribution networks increases the need for integrated transmission and distribution (T&D) co-simulation. A key challenge…

系统与控制 · 电气工程与系统科学 2026-04-23 Jong Ha Woo , Qi Xiao , Yu Ma , Zishuo Yang , Victor Daldegan Paduani , Ning Lu

Speech communication systems are prone to performance degradation in reverberant and noisy acoustic environments. Dereverberation and noise reduction algorithms typically require several model parameters, e.g. the speech, reverberation and…

音频与语音处理 · 电气工程与系统科学 2020-01-28 Yaron Laufer , Bracha Laufer-Goldshtein , Sharon Gannot

We consider time-of-arrival (ToA) estimation of a first arrival-path for a device working in narrowband Internet-of-Things (NB-IoT) systems. Due to a limited 180 KHz bandwidth used in NB-IoT, the time-domain auto-correlation function (ACF)…

信息论 · 计算机科学 2017-11-13 Sha Hu , Xuhong Li , Fredrik Rusek

The objective of this work is effective speaker diarisation using multi-scale speaker embeddings. Typically, there is a trade-off between the ability to recognise short speaker segments and the discriminative power of the embedding,…

音频与语音处理 · 电气工程与系统科学 2021-10-11 Youngki Kwon , Hee-Soo Heo , Jee-weon Jung , You Jin Kim , Bong-Jin Lee , Joon Son Chung

Artefacts that serve to distinguish bona fide speech from spoofed or deepfake speech are known to reside in specific subbands and temporal segments. Various approaches can be used to capture and model such artefacts, however, none works…

音频与语音处理 · 电气工程与系统科学 2021-08-24 Hemlata Tak , Jee-weon Jung , Jose Patino , Madhu Kamble , Massimiliano Todisco , Nicholas Evans

Gaussian processes regression is applied to augment experimental data of transfer-path analysis (TPA) by known information about the underlying physical properties of the system under investigation. The approach can be used as an…

经典物理 · 物理学 2019-05-21 Christopher Albert

The estimation of direction of arrival (DOA) is a crucial issue in conventional radar, wireless communication, and integrated sensing and communication (ISAC) systems. However, low-cost systems often suffer from imperfect factors, such as…

信号处理 · 电气工程与系统科学 2024-03-21 Peng Chen , Zhimin Chen , Liang Liu , Yun Chen , Xianbin Wang

Text-to-speech (TTS) methods have shown promising results in voice cloning, but they require a large number of labeled text-speech pairs. Minimally-supervised speech synthesis decouples TTS by combining two types of discrete speech…

声音 · 计算机科学 2023-12-19 Chunyu Qiang , Hao Li , Yixin Tian , Yi Zhao , Ying Zhang , Longbiao Wang , Jianwu Dang

Recent studies demonstrate that diffusion models can serve as a strong prior for solving inverse problems. A prominent example is Diffusion Posterior Sampling (DPS), which approximates the posterior distribution of data given the measure…

机器学习 · 统计学 2024-09-16 Yaxuan Zhu , Zehao Dou , Haoxin Zheng , Yasi Zhang , Ying Nian Wu , Ruiqi Gao

Many multi-microphone speech enhancement algorithms require the relative transfer function (RTF) vector of the desired speech source, relating the acoustic transfer functions of all array microphones to a reference microphone. In this…

音频与语音处理 · 电气工程与系统科学 2022-11-22 N. Gößling , S. Doclo

We propose a direction of arrival (DOA) estimation method that combines sound-intensity vector (IV)-based DOA estimation and DNN-based denoising and dereverberation. Since the accuracy of IV-based DOA estimation degrades due to…

音频与语音处理 · 电气工程与系统科学 2019-10-11 Masahiro Yasuda , Yuma Koizumi , Luca Mazzon , Shoichiro Saito , Hisashi Uematsu

The aim of speech enhancement is to improve speech signal quality and intelligibility from a noisy microphone signal. In many applications, it is crucial to enable processing with small computational complexity and minimal requirements…

音频与语音处理 · 电气工程与系统科学 2023-09-08 Julitta Bartolewska , Stanisław Kacprzak , Konrad Kowalczyk

In speaker-independent speech emotion recognition, the training and testing samples are collected from diverse speakers, leading to a multi-domain shift challenge across the feature distributions of data from different speakers.…

声音 · 计算机科学 2024-01-19 Cheng Lu , Yuan Zong , Hailun Lian , Yan Zhao , Björn Schuller , Wenming Zheng

This paper presents a speech enhancement method, where an adaptive threshold is statistically determined based on Gaussian modeling of Teager energy (TE) operated perceptual wavelet packet (PWP) coefficients of noisy speech. In order to…

音频与语音处理 · 电气工程与系统科学 2018-03-07 Md Tauhidul Islam , Celia Shahnaz

Current backscatter channel estimators employ an inefficient silent pilot transmission protocol, where tags alternate between silent and active states. To enhance performance, we propose a novel approach where tags remain active…

信号处理 · 电气工程与系统科学 2023-10-04 Fatemeh Rezaei , Diluka Galappaththige , Chintha Tellambura , Amine Maaref

We consider the problem of estimating the direction-of-arrival (DoA) of a desired source located in a known region of interest in the presence of interfering sources and multipath. We propose an approach that precedes the DoA estimation and…

信号处理 · 电气工程与系统科学 2024-09-13 Amitay Bar , Joseph S. Picard , Israel Cohen , Ronen Talmon

The performance of speaker verification degrades significantly in adverse acoustic environments with strong reverberation and noise. To address this issue, this paper proposes a spatial-temporal graph convolutional network (GCN) method for…

声音 · 计算机科学 2023-07-06 Yijiang Chen , Chengdong Liang , Xiao-Lei Zhang
‹ 上一页 1 8 9 10 下一页 ›