English

CroSentiNews 2.0: A Sentence-Level News Sentiment Corpus

Computation and Language 2023-05-16 v1

Abstract

This article presents a sentence-level sentiment dataset for the Croatian news domain. In addition to the 3K annotated texts already present, our dataset contains 14.5K annotated sentence occurrences that have been tagged with 5 classes. We provide baseline scores in addition to the annotation process and inter-annotator agreement.

Cite

@article{arxiv.2305.08187,
  title  = {CroSentiNews 2.0: A Sentence-Level News Sentiment Corpus},
  author = {Gaurish Thakkar and Nives Mikelic Preradović and Marko Tadić},
  journal= {arXiv preprint arXiv:2305.08187},
  year   = {2023}
}
R2 v1 2026-06-28T10:34:05.047Z