English

Analyzing the Influence of Dataset Composition for Emotion Recognition

Machine Learning 2021-03-08 v1

Abstract

Recognizing emotions from text in multimodal architectures has yielded promising results, surpassing video and audio modalities under certain circumstances. However, the method by which multimodal data is collected can be significant for recognizing emotional features in language. In this paper, we address the influence data collection methodology has on two multimodal emotion recognition datasets, the IEMOCAP dataset and the OMG-Emotion Behavior dataset, by analyzing textual dataset compositions and emotion recognition accuracy. Experiments with the full IEMOCAP dataset indicate that the composition negatively influences generalization performance when compared to the OMG-Emotion Behavior dataset. We conclude by discussing the impact this may have on HRI experiments.

Keywords

Cite

@article{arxiv.2103.03700,
  title  = {Analyzing the Influence of Dataset Composition for Emotion Recognition},
  author = {A. Sutherland and S. Magg and C. Weber and S. Wermter},
  journal= {arXiv preprint arXiv:2103.03700},
  year   = {2021}
}

Comments

2 pages, 2 figures, presented at IROS 2018 Workshop on Language and Robotics

R2 v1 2026-06-23T23:48:18.487Z