CroSentiNews 2.0: A Sentence-Level News Sentiment Corpus
Computation and Language
2023-05-16 v1
Abstract
This article presents a sentence-level sentiment dataset for the Croatian news domain. In addition to the 3K annotated texts already present, our dataset contains 14.5K annotated sentence occurrences that have been tagged with 5 classes. We provide baseline scores in addition to the annotation process and inter-annotator agreement.
Cite
@article{arxiv.2305.08187,
title = {CroSentiNews 2.0: A Sentence-Level News Sentiment Corpus},
author = {Gaurish Thakkar and Nives Mikelic Preradović and Marko Tadić},
journal= {arXiv preprint arXiv:2305.08187},
year = {2023}
}
Related papers
View all related →
Computation and Language · Computer Science
Croatian Film Review Dataset (Cro-FiReDa): A Sentiment Annotated Dataset of Film Reviews
Gaurish Thakkar, Nives Mikelic Preradovic, Marko Tadić
2023-05-16
Computation and Language · Computer Science
Author's Sentiment Prediction
Mohaddeseh Bastan, Mahnaz Koupaee, Youngseo Son, Richard Sicoli +1
2023-01-18
Computation and Language · Computer Science
Multi-task Learning for Cross-Lingual Sentiment Analysis
Gaurish Thakkar, Nives Mikelic Preradovic, Marko Tadic
2022-12-15
Computation and Language · Computer Science
Czech News Dataset for Semantic Textual Similarity
Jakub Sido, Michal Seják, Ondřej Pražák, Miloslav Konopík +1
2022-01-24
Computation and Language · Computer Science
FinnSentiment -- A Finnish Social Media Corpus for Sentiment Polarity Annotation
Krister Lindén, Tommi Jauhiainen, Sam Hardwick
2020-12-07
Computation and Language · Computer Science
Quotations, Coreference Resolution, and Sentiment Annotations in Croatian News Articles: An Exploratory Study
Jelena Sarajlić, Gaurish Thakkar, Diego Alves, Nives Mikelic Preradović
2022-12-15
Computation and Language · Computer Science
DynaSent: A Dynamic Benchmark for Sentiment Analysis
Christopher Potts, Zhengxuan Wu, Atticus Geiger, Douwe Kiela
2021-01-01
Computation and Language · Computer Science
Building a Sentiment Corpus of Tweets in Brazilian Portuguese
Henrico Bertini Brum, Maria das Graças Volpe Nunes
2017-12-27
Computation and Language · Computer Science
The ParlaSent-BCS dataset of sentiment-annotated parliamentary debates from Bosnia-Herzegovina, Croatia, and Serbia
Michal Mochtak, Peter Rupnik, Nikola Ljubešič
2022-06-03
Computation and Language · Computer Science
GoodNewsEveryone: A Corpus of News Headlines Annotated with Emotions, Semantic Roles, and Reader Perception
Laura Bostan, Evgeny Kim, Roman Klinger
2020-03-04
Computation and Language · Computer Science
Czech Dataset for Cross-lingual Subjectivity Classification
Pavel Přibáň, Josef Steinberger
2022-05-02
Computation and Language · Computer Science
Czech Dataset for Complex Aspect-Based Sentiment Analysis Tasks
Jakub Šmíd, Pavel Přibáň, Ondřej Pražák, Pavel Král
2025-08-12
Computation and Language · Computer Science
A Corpus for Sentence-level Subjectivity Detection on English News Articles
Francesco Antici, Andrea Galassi, Federico Ruggeri, Katerina Korre +4
2024-05-27
Computation and Language · Computer Science
ArSentD-LEV: A Multi-Topic Corpus for Target-based Sentiment Analysis in Arabic Levantine Tweets
Ramy Baly, Alaa Khaddaj, Hazem Hajj, Wassim El-Hajj +1
2019-06-06
Computation and Language · Computer Science
BCSAT : A Benchmark Corpus for Sentiment Analysis in Telugu Using Word-level Annotations
Sreekavitha Parupalli, Vijjini Anvesh Rao, Radhika Mamidi
2018-07-05
Computation and Language · Computer Science
Creation of the Estonian Subjectivity Dataset: Assessing the Degree of Subjectivity on a Scale
Karl Gustav Gailit, Kadri Muischnek, Kairit Sirts
2025-12-11
Computation and Language · Computer Science
Fine-grained Czech News Article Dataset: An Interdisciplinary Approach to Trustworthiness Analysis
Matyáš Boháček, Michal Bravanský, Filip Trhlík, Václav Moravec
2022-12-19
Computation and Language · Computer Science
A Fine-Grained Sentiment Dataset for Norwegian
Lilja Øvrelid, Petter Mæhlum, Jeremy Barnes, Erik Velldal
2020-04-07
Computation and Language · Computer Science
Predicting Sentence-Level Factuality of News and Bias of Media Outlets
Francielle Vargas, Kokil Jaidka, Thiago A. S. Pardo, Fabrício Benevenuto
2024-09-16
Computation and Language · Computer Science
An English-Hindi Code-Mixed Corpus: Stance Annotation and Baseline System
Sahil Swami, Ankush Khandelwal, Vinay Singh, Syed Sarfaraz Akhtar +1
2018-05-31
Computation and Language · Computer Science
SEntFiN 1.0: Entity-Aware Sentiment Analysis for Financial News
Ankur Sinha, Satishwar Kedas, Rishu Kumar, Pekka Malo
2023-05-23
Computation and Language · Computer Science
The Causal News Corpus: Annotating Causal Relations in Event Sentences from News
Fiona Anting Tan, Ali Hürriyetoğlu, Tommaso Caselli, Nelleke Oostdijk +6
2022-04-26
Computation and Language · Computer Science
A Quality Type-aware Annotated Corpus and Lexicon for Harassment Research
Mohammadreza Rezvan, Saeedeh Shekarpour, Lakshika Balasuriya, Krishnaprasad Thirunarayan +2
2018-05-25
Computation and Language · Computer Science
iNews: A Multimodal Dataset for Modeling Personalized Affective Responses to News
Tiancheng Hu, Nigel Collier
2025-07-08
Computation and Language · Computer Science
Agreeing to Disagree: Annotating Offensive Language Datasets with Annotators' Disagreement
Elisa Leonardelli, Stefano Menini, Alessio Palmero Aprosio, Marco Guerini +1
2022-10-17