English

Spatiotemporal Networks for Video Emotion Recognition

Computer Vision and Pattern Recognition 2017-04-13 v3

Abstract

Our experiment adapts several popular deep learning methods as well as some traditional methods on the problem of video emotion recognition. In our experiment, we use the CNN-LSTM architecture for visual information extraction and classification and utilize traditional methods such as for audio feature classification. For multimodal fusion, we use the traditional Support Vector Machine. Our experiment yields a good result on the AFEW 6.0 Dataset.

Keywords

Cite

@article{arxiv.1704.00570,
  title  = {Spatiotemporal Networks for Video Emotion Recognition},
  author = {Lijie Fan and Yunjie Ke},
  journal= {arXiv preprint arXiv:1704.00570},
  year   = {2017}
}

Comments

The reason is that the article is just an experimental report without being reviewed carefully. There are some fatal drawbacks in this article and it may not be suitable for being published

R2 v1 2026-06-22T19:05:44.432Z