English

Masked Autoencoders for Egocentric Video Understanding @ Ego4D Challenge 2022

Computer Vision and Pattern Recognition 2022-11-29 v1

Abstract

In this report, we present our approach and empirical results of applying masked autoencoders in two egocentric video understanding tasks, namely, Object State Change Classification and PNR Temporal Localization, of Ego4D Challenge 2022. As team TheSSVL, we ranked 2nd place in both tasks. Our code will be made available.

Keywords

Cite

@article{arxiv.2211.15286,
  title  = {Masked Autoencoders for Egocentric Video Understanding @ Ego4D Challenge 2022},
  author = {Jiachen Lei and Shuang Ma and Zhongjie Ba and Sai Vemprala and Ashish Kapoor and Kui Ren},
  journal= {arXiv preprint arXiv:2211.15286},
  year   = {2022}
}

Comments

5 pages