English

UG18 at SemEval-2018 Task 1: Generating Additional Training Data for Predicting Emotion Intensity in Spanish

Computation and Language 2018-05-29 v1

Abstract

The present study describes our submission to SemEval 2018 Task 1: Affect in Tweets. Our Spanish-only approach aimed to demonstrate that it is beneficial to automatically generate additional training data by (i) translating training data from other languages and (ii) applying a semi-supervised learning method. We find strong support for both approaches, with those models outperforming our regular models in all subtasks. However, creating a stepwise ensemble of different models as opposed to simply averaging did not result in an increase in performance. We placed second (EI-Reg), second (EI-Oc), fourth (V-Reg) and fifth (V-Oc) in the four Spanish subtasks we participated in.

Keywords

Cite

@article{arxiv.1805.10824,
  title  = {UG18 at SemEval-2018 Task 1: Generating Additional Training Data for Predicting Emotion Intensity in Spanish},
  author = {Marloes Kuijper and Mike van Lenthe and Rik van Noord},
  journal= {arXiv preprint arXiv:1805.10824},
  year   = {2018}
}

Comments

Accepted at SemEval 2018

R2 v1 2026-06-23T02:10:10.762Z