English

Dual-LoRA: Parameter-Efficient Adversarial Disentanglement for Cross-Lingual Speaker Verification

Audio and Speech Processing 2026-05-01 v2

Abstract

Cross-lingual speaker verification suffers from severe language-speaker entanglement. This causes systematic degradation in the hardest scenario: correctly accepting utterances from the same speaker across different languages while rejecting those from different speakers sharing the same language. Standard adversarial disentanglement degrades speaker discriminability; blind discriminators inadvertently penalize speaker-discriminative traits that merely correlate with language. To address this, we propose Dual-LoRA, injecting trainable task-factorized LoRA adapters into a frozen pre-trained backbone. Our core innovation is a Language-Anchored Adversary: by grounding the discriminator with an explicit language branch, adversarial gradients target true linguistic cues rather than arbitrary correlations, preserving essential speaker characteristics. Evaluated on the TidyVoice benchmark, our system achieves a 0.91% validation EER and achieves 3rd place in the official challenge.

Keywords

Cite

@article{arxiv.2604.26327,
  title  = {Dual-LoRA: Parameter-Efficient Adversarial Disentanglement for Cross-Lingual Speaker Verification},
  author = {Qituan Shangguan and Junhao Du and Kunyang Peng and Feng Xue and Hui Zhang and Xinsheng Wang and Kai Yu and Shuai Wang},
  journal= {arXiv preprint arXiv:2604.26327},
  year   = {2026}
}

Comments

Submitted to Interspeech 2026; 5 pages

R2 v1 2026-07-01T12:40:33.943Z