English

Learning to detect video events from zero or very few video examples

Machine Learning 2015-11-26 v1 Computer Vision and Pattern Recognition

Abstract

In this work we deal with the problem of high-level event detection in video. Specifically, we study the challenging problems of i) learning to detect video events from solely a textual description of the event, without using any positive video examples, and ii) additionally exploiting very few positive training samples together with a small number of ``related'' videos. For learning only from an event's textual description, we first identify a general learning framework and then study the impact of different design choices for various stages of this framework. For additionally learning from example videos, when true positive training samples are scarce, we employ an extension of the Support Vector Machine that allows us to exploit ``related'' event videos by automatically introducing different weights for subsets of the videos in the overall training set. Experimental evaluations performed on the large-scale TRECVID MED 2014 video dataset provide insight on the effectiveness of the proposed methods.

Keywords

Cite

@article{arxiv.1511.08032,
  title  = {Learning to detect video events from zero or very few video examples},
  author = {Christos Tzelepis and Damianos Galanopoulos and Vasileios Mezaris and Ioannis Patras},
  journal= {arXiv preprint arXiv:1511.08032},
  year   = {2015}
}

Comments

Image and Vision Computing Journal, Elsevier, 2015, accepted for publication

R2 v1 2026-06-22T11:54:00.209Z