English

Deep Learning Brasil -- NLP at SemEval-2020 Task 9: Overview of Sentiment Analysis of Code-Mixed Tweets

Computation and Language 2020-08-05 v1 Information Retrieval Machine Learning

Abstract

In this paper, we describe a methodology to predict sentiment in code-mixed tweets (hindi-english). Our team called verissimo.manoel in CodaLab developed an approach based on an ensemble of four models (MultiFiT, BERT, ALBERT, and XLNET). The final classification algorithm was an ensemble of some predictions of all softmax values from these four models. This architecture was used and evaluated in the context of the SemEval 2020 challenge (task 9), and our system got 72.7% on the F1 score.

Keywords

Cite

@article{arxiv.2008.01544,
  title  = {Deep Learning Brasil -- NLP at SemEval-2020 Task 9: Overview of Sentiment Analysis of Code-Mixed Tweets},
  author = {Manoel Veríssimo dos Santos Neto and Ayrton Denner da Silva Amaral and Nádia Félix Felipe da Silva and Anderson da Silva Soares},
  journal= {arXiv preprint arXiv:2008.01544},
  year   = {2020}
}