English

Sentimental LIAR: Extended Corpus and Deep Learning Models for Fake Claim Classification

Computation and Language 2020-10-23 v2 Machine Learning Social and Information Networks Machine Learning

Abstract

The rampant integration of social media in our every day lives and culture has given rise to fast and easier access to the flow of information than ever in human history. However, the inherently unsupervised nature of social media platforms has also made it easier to spread false information and fake news. Furthermore, the high volume and velocity of information flow in such platforms make manual supervision and control of information propagation infeasible. This paper aims to address this issue by proposing a novel deep learning approach for automated detection of false short-text claims on social media. We first introduce Sentimental LIAR, which extends the LIAR dataset of short claims by adding features based on sentiment and emotion analysis of claims. Furthermore, we propose a novel deep learning architecture based on the BERT-Base language model for classification of claims as genuine or fake. Our results demonstrate that the proposed architecture trained on Sentimental LIAR can achieve an accuracy of 70%, which is an improvement of ~30% over previously reported results for the LIAR benchmark.

Keywords

Cite

@article{arxiv.2009.01047,
  title  = {Sentimental LIAR: Extended Corpus and Deep Learning Models for Fake Claim Classification},
  author = {Bibek Upadhayay and Vahid Behzadan},
  journal= {arXiv preprint arXiv:2009.01047},
  year   = {2020}
}

Comments

Accepted for publication in the proceedings of IEEE ISI '20

R2 v1 2026-06-23T18:16:02.889Z