English

Multi-model learning by sequential reading of untrimmed videos for action recognition

Computer Vision and Pattern Recognition 2024-01-29 v1

Abstract

We propose a new method for learning videos by aggregating multiple models by sequentially extracting video clips from untrimmed video. The proposed method reduces the correlation between clips by feeding clips to multiple models in turn and synchronizes these models through federated learning. Experimental results show that the proposed method improves the performance compared to the no synchronization.

Keywords

Cite

@article{arxiv.2401.14675,
  title  = {Multi-model learning by sequential reading of untrimmed videos for action recognition},
  author = {Kodai Kamiya and Toru Tamaki},
  journal= {arXiv preprint arXiv:2401.14675},
  year   = {2024}
}

Comments

The International Workshop on Frontiers of Computer Vision (IW-FCV2024)