Clova Baseline System for the VoxCeleb Speaker Recognition Challenge 2020
Audio and Speech Processing
2020-09-30 v1 Sound
Abstract
This report describes our submission to the VoxCeleb Speaker Recognition Challenge (VoxSRC) at Interspeech 2020. We perform a careful analysis of speaker recognition models based on the popular ResNet architecture, and train a number of variants using a range of loss functions. Our results show significant improvements over most existing works without the use of model ensemble or post-processing. We release the training code and pre-trained models as unofficial baselines for this year's challenge.
Cite
@article{arxiv.2009.14153,
title = {Clova Baseline System for the VoxCeleb Speaker Recognition Challenge 2020},
author = {Hee Soo Heo and Bong-Jin Lee and Jaesung Huh and Joon Son Chung},
journal= {arXiv preprint arXiv:2009.14153},
year = {2020}
}