In this paper, we describe the team \textit{BRUMS} entry to OffensEval 2: Multilingual Offensive Language Identification in Social Media in SemEval-2020. The OffensEval organizers provided participants with annotated datasets containing posts from social media in Arabic, Danish, English, Greek and Turkish. We present a multilingual deep learning model to identify offensive language in social media. Overall, the approach achieves acceptable evaluation scores, while maintaining flexibility between languages.
@article{arxiv.2010.06278,
title = {BRUMS at SemEval-2020 Task 12 : Transformer based Multilingual Offensive Language Identification in Social Media},
author = {Tharindu Ranasinghe and Hansi Hettiarachchi},
journal= {arXiv preprint arXiv:2010.06278},
year = {2020}
}
Comments
Accepted to SemEval-2020 (International Workshop on Semantic Evaluation) at COLING 2020