English
Related papers

Related papers: Detecting Toxicity in News Articles: Application t…

200 papers

Fake news has been coming into sight in significant numbers for numerous business and political reasons and has become frequent in the online world. People can get contaminated easily by these fake news for its fabricated words which have…

Computation and Language · Computer Science 2020-06-01 Md Gulzar Hussain , Md Rashidul Hasan , Mahmuda Rahman , Joy Protim , Sakib Al Hasan

The task of toxicity detection is still a relevant task, especially in the context of safe and fair LMs development. Nevertheless, labeled binary toxicity classification corpora are not available for all languages, which is understandable…

Computation and Language · Computer Science 2024-04-30 Daryna Dementieva , Valeriia Khylenko , Nikolay Babakov , Georg Groh

This paper presents a new training dataset for automatic genre identification GINCO, which is based on 1,125 crawled Slovenian web documents that consist of 650 thousand words. Each document was manually annotated for genre with a new…

Computation and Language · Computer Science 2022-01-12 Taja Kuzman , Peter Rupnik , Nikola Ljubešić

In the digital age, the prevalence of misleading news headlines poses a significant challenge to information integrity, necessitating robust detection mechanisms. This study explores the efficacy of Large Language Models (LLMs) in…

Computation and Language · Computer Science 2024-05-07 Md Main Uddin Rony , Md Mahfuzul Haque , Mohammad Ali , Ahmed Shatil Alam , Naeemul Hassan

WARNING: This paper contains examples of offensive materials. To address the proliferation of toxic content on social media, we introduce SMARTER, we introduce SMARTER, a data-efficient two-stage framework for explainable content moderation…

Computation and Language · Computer Science 2026-04-23 Huy Nghiem , Advik Sachdeva , Hal Daumé

As the world is becoming more dependent on the internet for information exchange, some overzealous journalists, hackers, bloggers, individuals and organizations tend to abuse the gift of free information environment by polluting it with…

Computation and Language · Computer Science 2021-04-02 Kwadwo Osei Bonsu

Some news headlines mislead readers with overrated or false information, and identifying them in advance will better assist readers in choosing proper news stories to consume. This research introduces million-scale pairs of news headline…

Computation and Language · Computer Science 2019-02-11 Seunghyun Yoon , Kunwoo Park , Joongbo Shin , Hongjun Lim , Seungpil Won , Meeyoung Cha , Kyomin Jung

Moderation is crucial to promoting healthy on-line discussions. Although several `toxicity' detection datasets and models have been published, most of them ignore the context of the posts, implicitly assuming that comments maybe judged…

Computation and Language · Computer Science 2020-06-02 John Pavlopoulos , Jeffrey Sorensen , Lucas Dixon , Nithum Thain , Ion Androutsopoulos

On the one hand, nowadays, fake news articles are easily propagated through various online media platforms and have become a grand threat to the trustworthiness of information. On the other hand, our understanding of the language of fake…

Computation and Language · Computer Science 2019-04-11 Hamid Karimi , Jiliang Tang

This paper reports on a writing style analysis of hyperpartisan (i.e., extremely one-sided) news in connection to fake news. It presents a large corpus of 1,627 articles that were manually fact-checked by professional journalists from…

Computation and Language · Computer Science 2017-02-21 Martin Potthast , Johannes Kiesel , Kevin Reinartz , Janek Bevendorff , Benno Stein

Large language models (LLMs) are increasingly popular but are also prone to generating bias, toxic or harmful language, which can have detrimental effects on individuals and communities. Although most efforts is put to assess and mitigate…

Computation and Language · Computer Science 2024-06-26 Caroline Brun , Vassilina Nikoulina

Abstract: In this paper we present an approach to develop a text-classification model which would be able to identify populist content in text. The developed BERT-based model is largely successful in identifying populist content in text and…

Computation and Language · Computer Science 2021-06-11 Jogilė Ulinskaitė , Lukas Pukelis

Over the past decade, the media landscape has seen a radical shift. As more of the public stay informed of current events via online sources, competition has grown as outlets vie for attention. This competition has prompted some online…

Human-Computer Interaction · Computer Science 2023-01-10 Marc Kydd , Lynsay A. Shepherd

To identify and classify toxic online commentary, the modern tools of data science transform raw text into key features from which either thresholding or learning algorithms can make predictions for monitoring offensive conversations. We…

Machine Learning · Computer Science 2018-10-05 David Noever

The goal of this work is to build a classifier that can identify text complexity within the context of teaching reading to English as a Second Language (ESL) learners. To present language learners with texts that are suitable to their level…

Computation and Language · Computer Science 2023-06-22 M. Zakaria Kurdi

This study explores the generation and evaluation of synthetic fake news through fact based manipulations using large language models (LLMs). We introduce a novel methodology that extracts key facts from real articles, modifies them, and…

Computation and Language · Computer Science 2025-04-10 Abdul Sittar , Luka Golob , Mateja Smiljanic

Detection of fake news is crucial to ensure the authenticity of information and maintain the news ecosystems reliability. Recently, there has been an increase in fake news content due to the recent proliferation of social media and fake…

Social and Information Networks · Computer Science 2022-07-28 Pallabi Saikia , Kshitij Gundale , Ankit Jain , Dev Jadeja , Harvi Patel , Mohendra Roy

Social media is filled with toxic content. The aim of this paper is to build a model that can detect insincere questions. We use the 'Quora Insincere Questions Classification' dataset for our analysis. The dataset is composed of sincere and…

Computation and Language · Computer Science 2019-11-05 Deepshi Mediratta , Nikhil Oswal

Over the past decade, we have witnessed the rise of misinformation on the Internet, with online users constantly falling victims of fake news. A multitude of past studies have analyzed fake news diffusion mechanics and detection and…

Social and Information Networks · Computer Science 2021-03-18 Manolis Chalkiadakis , Alexandros Kornilakis , Panagiotis Papadopoulos , Evangelos P. Markatos , Nicolas Kourtellis

Today's social media platforms enable to spread both authentic and fake news very quickly. Some approaches have been proposed to automatically detect such "fake" news based on their content, but it is difficult to agree on universal…

Social and Information Networks · Computer Science 2018-08-30 Oana Balmau , Rachid Guerraoui , Anne-Marie Kermarrec , Alexandre Maurer , Matej Pavlovic , Willy Zwaenepoel