English

TV100: A TV Series Dataset that Pre-Trained CLIP Has Not Seen

Computer Vision and Pattern Recognition 2024-04-22 v1 Machine Learning

Abstract

The era of pre-trained models has ushered in a wealth of new insights for the machine learning community. Among the myriad of questions that arise, one of paramount importance is: 'Do pre-trained models possess comprehensive knowledge?' This paper seeks to address this crucial inquiry. In line with our objective, we have made publicly available a novel dataset comprised of images from TV series released post-2021. This dataset holds significant potential for use in various research areas, including the evaluation of incremental learning, novel class discovery, and long-tailed learning, among others. Project page: https://tv-100.github.io/

Keywords

Cite

@article{arxiv.2404.12407,
  title  = {TV100: A TV Series Dataset that Pre-Trained CLIP Has Not Seen},
  author = {Da-Wei Zhou and Zhi-Hong Qi and Han-Jia Ye and De-Chuan Zhan},
  journal= {arXiv preprint arXiv:2404.12407},
  year   = {2024}
}

Comments

Project page: https://tv-100.github.io/