English

Examining a hate speech corpus for hate speech detection and popularity prediction

Computation and Language 2018-05-15 v1 Artificial Intelligence Computers and Society

Abstract

As research on hate speech becomes more and more relevant every day, most of it is still focused on hate speech detection. By attempting to replicate a hate speech detection experiment performed on an existing Twitter corpus annotated for hate speech, we highlight some issues that arise from doing research in the field of hate speech, which is essentially still in its infancy. We take a critical look at the training corpus in order to understand its biases, while also using it to venture beyond hate speech detection and investigate whether it can be used to shed light on other facets of research, such as popularity of hate tweets.

Keywords

Cite

@article{arxiv.1805.04661,
  title  = {Examining a hate speech corpus for hate speech detection and popularity prediction},
  author = {Filip Klubička and Raquel Fernández},
  journal= {arXiv preprint arXiv:1805.04661},
  year   = {2018}
}

Comments

8 pages, 1 figure, 10 tables, published in proceedings of 4REAL2018: Workshop on Replicability and Reproducibility of Research Results in Science and Technology of Language

R2 v1 2026-06-23T01:52:43.859Z