English

DASZL: Dynamic Action Signatures for Zero-shot Learning

Computer Vision and Pattern Recognition 2020-11-19 v3

Abstract

There are many realistic applications of activity recognition where the set of potential activity descriptions is combinatorially large. This makes end-to-end supervised training of a recognition system impractical as no training set is practically able to encompass the entire label set. In this paper, we present an approach to fine-grained recognition that models activities as compositions of dynamic action signatures. This compositional approach allows us to reframe fine-grained recognition as zero-shot activity recognition, where a detector is composed "on the fly" from simple first-principles state machines supported by deep-learned components. We evaluate our method on the Olympic Sports and UCF101 datasets, where our model establishes a new state of the art under multiple experimental paradigms. We also extend this method to form a unique framework for zero-shot joint segmentation and classification of activities in video and demonstrate the first results in zero-shot decoding of complex action sequences on a widely-used surgical dataset. Lastly, we show that we can use off-the-shelf object detectors to recognize activities in completely de-novo settings with no additional training.

Keywords

Cite

@article{arxiv.1912.03613,
  title  = {DASZL: Dynamic Action Signatures for Zero-shot Learning},
  author = {Tae Soo Kim and Jonathan D. Jones and Michael Peven and Zihao Xiao and Jin Bai and Yi Zhang and Weichao Qiu and Alan Yuille and Gregory D. Hager},
  journal= {arXiv preprint arXiv:1912.03613},
  year   = {2020}
}

Comments

10 pages, 4 figures, 3 tables, AAAI2021 submission

R2 v1 2026-06-23T12:39:08.248Z