English

Noise-Conditioned Mixture-of-Experts Framework for Robust Speaker Verification

Sound 2026-03-11 v3 Multimedia Audio and Speech Processing

Abstract

Robust speaker verification under noisy conditions remains an open challenge. Conventional deep learning methods learn a robust unified speaker representation space against diverse background noise and achieve significant improvement. In contrast, this paper presents a noise-conditioned mixture-ofexperts framework that decomposes the feature space into specialized noise-aware subspaces for speaker verification. Specifically, we propose a noise-conditioned expert routing mechanism, a universal model based expert specialization strategy, and an SNR-decaying curriculum learning protocol, collectively improving model robustness and generalization under diverse noise conditions. The proposed method can automatically route inputs to expert networks based on noise information derived from the inputs, where each expert targets distinct noise characteristics while preserving speaker identity information. Comprehensive experiments demonstrate consistent superiority over baselines

Keywords

Cite

@article{arxiv.2510.18533,
  title  = {Noise-Conditioned Mixture-of-Experts Framework for Robust Speaker Verification},
  author = {Bin Gu and Haitao Zhao and Jibo Wei},
  journal= {arXiv preprint arXiv:2510.18533},
  year   = {2026}
}

Comments

Accepted by Signal Processing Letters

R2 v1 2026-07-01T06:57:41.213Z