English
Related papers

Related papers: DariMis: Harm-Aware Modeling for Dari Misinformati…

200 papers

Disinformation spreads rapidly across linguistic boundaries, yet most AI models are still benchmarked only on English. We address this gap with a systematic comparison of five multilingual transformer models: mBERT, XLM, XLM-RoBERTa,…

Computation and Language · Computer Science 2026-05-12 Zaur Gouliev , Jennifer Waters , Chengqian Wang

Propaganda is a form of persuasion that has been used throughout history with the intention goal of influencing people's opinions through rhetorical and psychological persuasion techniques for determined ends. Although Arabic ranked as the…

Computation and Language · Computer Science 2025-02-21 Lubna Al-Henaki , Hend Al-Khalifa , Abdulmalik Al-Salman , Hajar Alqubayshi , Hind Al-Twailay , Gheeda Alghamdi , Hawra Aljasim

The rapid integration of large language models into newsroom workflows has raised urgent questions about the prevalence of AI-generated content in online media. While computational studies have begun to quantify this phenomenon in…

Computation and Language · Computer Science 2026-02-17 Ozancan Ozdemir

This paper gives the overview of the first shared task at FIRE 2020 on fake news detection in the Urdu language. This is a binary classification task in which the goal is to identify fake news using a dataset composed of 900 annotated news…

Computation and Language · Computer Science 2022-07-27 Maaz Amjad , Grigori Sidorov , Alisa Zhila , Alexander Gelbukh , Paolo Rosso

Widely distributed misinformation shared across social media channels is a pressing issue that poses a significant threat to many aspects of society's well-being. Inaccurate shared information causes confusion, can adversely affect mental…

Social and Information Networks · Computer Science 2024-09-27 Juanita Zainudin , Nazlena Mohamad Ali , Alan F. Smeaton , Mohamad Taha Ijab

This paper demonstrates a two-stage method for deriving insights from social media data relating to disinformation by applying a combination of geospatial classification and embedding-based language modelling across multiple languages. In…

Computation and Language · Computer Science 2021-08-09 David Tuxworth , Dimosthenis Antypas , Luis Espinosa-Anke , Jose Camacho-Collados , Alun Preece , David Rogers

Various social media platforms, e.g., Twitter and Reddit, allow people to disseminate a plethora of information more efficiently and conveniently. However, they are inevitably full of misinformation, causing damage to diverse aspects of our…

Computation and Language · Computer Science 2024-07-30 Bing Wang , Ximing Li , Changchun Li , Bo Fu , Songwen Pei , Shengsheng Wang

Large language models (LLMs) tuned for safety often avoid acknowledging demographic differences, even when such acknowledgment is factually correct (e.g., ancestry-based disease incidence) or contextually justified (e.g., religious hiring…

Computation and Language · Computer Science 2026-04-21 Ziwen Pan , Zihan Liang , Jad Kabbara , Ali Emami

Real-world information, often multimodal, can be misinformed or potentially misleading due to factual errors, outdated claims, missing context, misinterpretation, and more. Such "misinformation" is understudied, challenging to address, and…

Computation and Language · Computer Science 2026-01-13 Xinyi Zhou , Ashish Sharma , Amy X. Zhang , Tim Althoff

The prevalence and harms of online misinformation is a perennial concern for internet platforms, institutions and society at large. Over time, information shared online has become more media-heavy and misinformation has readily adapted to…

As a leading online platform with a vast global audience, YouTube's extensive reach also makes it susceptible to hosting harmful content, including disinformation and conspiracy theories. This study explores the use of open-weight Large…

Computation and Language · Computer Science 2025-07-08 Leonardo La Rocca , Francesco Corso , Francesco Pierri

Content moderation research has recently made significant advances, but remains limited in serving the majority of the world's languages due to the lack of resources, leaving millions of vulnerable users to online hostility. This work…

Computation and Language · Computer Science 2025-10-28 Fitsum Gaim , Hoyun Song , Huije Lee , Changgeon Ko , Eui Jun Hwang , Jong C. Park

Nowadays, misinformation is widely spreading over various social media platforms and causes extremely negative impacts on society. To combat this issue, automatically identifying misinformation, especially those containing multimodal…

Computation and Language · Computer Science 2024-07-30 Bing Wang , Shengsheng Wang , Changchun Li , Renchu Guan , Ximing Li

Transcribed speech and user-generated text in Arabic typically contain a mixture of Modern Standard Arabic (MSA), the standardized language taught in schools, and Dialectal Arabic (DA), used in daily communications. To handle this…

Computation and Language · Computer Science 2023-10-24 Amr Keleg , Sharon Goldwater , Walid Magdy

In the digital age, the prevalence of misleading news headlines poses a significant challenge to information integrity, necessitating robust detection mechanisms. This study explores the efficacy of Large Language Models (LLMs) in…

Computation and Language · Computer Science 2024-05-07 Md Main Uddin Rony , Md Mahfuzul Haque , Mohammad Ali , Ahmed Shatil Alam , Naeemul Hassan

A quarter of US adults regularly get their news from YouTube. Yet, despite the massive political content available on the platform, to date no classifier has been proposed to identify the political leaning of YouTube videos. To fill this…

Computation and Language · Computer Science 2024-04-09 Nouar AlDahoul , Talal Rahwan , Yasir Zaki

Warning users about misinformation on social media is not a simple usability task. Soft moderation has to balance between debunking falsehoods and avoiding moderation bias while preserving the social media consumption flow. Platforms thus…

Computers and Society · Computer Science 2022-05-04 Filipo Sharevski , Amy Devine , Emma Pieroni , Peter Jacnim

Online toxic language causes real harm, especially in regions with limited moderation tools. In this study, we evaluate how large language models handle toxic comments in Serbian, Croatian, and Bosnian, languages with limited labeled data.…

Computation and Language · Computer Science 2025-06-16 Amel Muminovic , Amela Kadric Muminovic

Online social media platforms such as YouTube have a wide, global reach. However, little is known about the experience of low-resourced language speakers on such platforms; especially in how they experience and navigate harmful content. To…

Human-Computer Interaction · Computer Science 2024-05-28 Hellina Hailu Nigatu , Inioluwa Deborah Raji

The increase in active users on social networking sites (SNSs) has also observed an increase in harmful content on social media sites. Harmful content is described as an inappropriate activity to harm or deceive an individual or a group of…

Social and Information Networks · Computer Science 2024-03-05 Gautam Kishore Shahi