English

Clova Baseline System for the VoxCeleb Speaker Recognition Challenge 2020

Audio and Speech Processing 2020-09-30 v1 Sound

Abstract

This report describes our submission to the VoxCeleb Speaker Recognition Challenge (VoxSRC) at Interspeech 2020. We perform a careful analysis of speaker recognition models based on the popular ResNet architecture, and train a number of variants using a range of loss functions. Our results show significant improvements over most existing works without the use of model ensemble or post-processing. We release the training code and pre-trained models as unofficial baselines for this year's challenge.

Keywords

Cite

@article{arxiv.2009.14153,
  title  = {Clova Baseline System for the VoxCeleb Speaker Recognition Challenge 2020},
  author = {Hee Soo Heo and Bong-Jin Lee and Jaesung Huh and Joon Son Chung},
  journal= {arXiv preprint arXiv:2009.14153},
  year   = {2020}
}