English

Bootstrap your own latent: A new approach to self-supervised Learning

Machine Learning 2020-09-11 v3 Computer Vision and Pattern Recognition Machine Learning

Abstract

We introduce Bootstrap Your Own Latent (BYOL), a new approach to self-supervised image representation learning. BYOL relies on two neural networks, referred to as online and target networks, that interact and learn from each other. From an augmented view of an image, we train the online network to predict the target network representation of the same image under a different augmented view. At the same time, we update the target network with a slow-moving average of the online network. While state-of-the art methods rely on negative pairs, BYOL achieves a new state of the art without them. BYOL reaches 74.3%74.3\% top-1 classification accuracy on ImageNet using a linear evaluation with a ResNet-50 architecture and 79.6%79.6\% with a larger ResNet. We show that BYOL performs on par or better than the current state of the art on both transfer and semi-supervised benchmarks. Our implementation and pretrained models are given on GitHub.

Keywords

Cite

@article{arxiv.2006.07733,
  title  = {Bootstrap your own latent: A new approach to self-supervised Learning},
  author = {Jean-Bastien Grill and Florian Strub and Florent Altché and Corentin Tallec and Pierre H. Richemond and Elena Buchatskaya and Carl Doersch and Bernardo Avila Pires and Zhaohan Daniel Guo and Mohammad Gheshlaghi Azar and Bilal Piot and Koray Kavukcuoglu and Rémi Munos and Michal Valko},
  journal= {arXiv preprint arXiv:2006.07733},
  year   = {2020}
}
R2 v1 2026-06-23T16:18:13.918Z