English
Related papers

Related papers: HateBuffer: Safeguarding Content Moderators' Menta…

200 papers

Social media systems allow Internet users a congenial platform to freely express their thoughts and opinions. Although this property represents incredible and unique communication opportunities, it also brings along important challenges.…

Social and Information Networks · Computer Science 2016-03-25 Leandro Silva , Mainack Mondal , Denzil Correa , Fabricio Benevenuto , Ingmar Weber

The rapid proliferation of harmful and emotionally damaging content on social media platforms has intensified concerns regarding societal harm. While content moderation efforts primarily focus on detecting and removing harmful posts, less…

Social and Information Networks · Computer Science 2026-05-01 Syed Mhamudul Hasan , Mohd. Farhan Israk Soumik , Abdur R. Shahid

Commercial content moderation APIs are marketed as scalable solutions to combat online hate speech. However, the reliance on these APIs risks both silencing legitimate speech, called over-moderation, and failing to protect online platforms…

Human-Computer Interaction · Computer Science 2025-03-04 David Hartmann , Amin Oueslati , Dimitri Staufer , Lena Pohlmann , Simon Munzert , Hendrik Heuer

The internet has become a central medium through which `networked publics' express their opinions and engage in debate. Offensive comments and personal attacks can inhibit participation in these spaces. Automated content moderation aims to…

Computers and Society · Computer Science 2017-09-06 Reuben Binns , Michael Veale , Max Van Kleek , Nigel Shadbolt

Harmful or abusive online content has been increasing over time, raising concerns for social media platforms, government agencies, and policymakers. Such harmful or abusive content can have major negative impact on society, e.g.,…

Computation and Language · Computer Science 2022-05-10 Rabindra Nath Nandi , Firoj Alam , Preslav Nakov

Hateful Meme Challenge proposed by Facebook AI has attracted contestants around the world. The challenge focuses on detecting hateful speech in multimodal memes. Various state-of-the-art deep learning models have been applied to this…

Computer Vision and Pattern Recognition · Computer Science 2021-12-22 Aijing Gao , Bingjun Wang , Jiaqi Yin , Yating Tian

The automatic identification of harmful content online is of major concern for social media platforms, policymakers, and society. Researchers have studied textual, visual, and audio content, but typically in isolation. Yet, harmful content…

A significant challenge in automating hate speech detection on social media is distinguishing hate speech from regular and offensive language. These identify an essential category of content that web filters seek to remove. Only automated…

Computation and Language · Computer Science 2024-11-12 Faria Naznin , Md Touhidur Rahman , Shahran Rahman Alve

Citizen-generated counter speech is a promising way to fight hate speech and promote peaceful, non-polarized discourse. However, there is a lack of large-scale longitudinal studies of its effectiveness for reducing hate speech. To this end,…

Social and Information Networks · Computer Science 2021-09-07 Joshua Garland , Keyan Ghazi-Zahedi , Jean-Gabriel Young , Laurent Hébert-Dufresne , Mirta Galesic

This study investigates how online counterspeech, defined as direct responses to harmful online content with the intention of dissuading the perpetrator from further engaging in such behavior, is influenced by the match between a target of…

Human-Computer Interaction · Computer Science 2024-11-05 Kaike Ping , James Hawdon , Eugenia Rho

This paper investigates how content moderation affects content creation in an ideologically diverse online environments. We develop a model in which users act as both creators and consumers, differing in their ideological affiliation and…

General Economics · Economics 2025-11-26 Ying Bao , Jessie Liu

The goal of hate speech detection is to filter negative online content aiming at certain groups of people. Due to the easy accessibility of social media platforms it is crucial to protect everyone which requires building hate speech…

Computation and Language · Computer Science 2022-01-19 Irina Bigoulaeva , Viktor Hangya , Iryna Gurevych , Alexander Fraser

Hate speech in social media is a growing phenomenon, and detecting such toxic content has recently gained significant traction in the research community. Existing studies have explored fine-tuning language models (LMs) to perform hate…

Computation and Language · Computer Science 2023-03-07 Md Rabiul Awal , Roy Ka-Wei Lee , Eshaan Tanwar , Tanmay Garg , Tanmoy Chakraborty

In the past few years, there has been a surge of interest in multi-modal problems, from image captioning to visual question answering and beyond. In this paper, we focus on hate speech detection in multi-modal memes wherein memes pose an…

Computer Vision and Pattern Recognition · Computer Science 2021-01-01 Abhishek Das , Japsimar Singh Wahi , Siyao Li

Hate speech detection is a crucial area of research in natural language processing, essential for ensuring online community safety. However, detecting implicit hate speech, where harmful intent is conveyed in subtle or indirect ways,…

Computation and Language · Computer Science 2025-04-17 Yumin Kim , Hwanhee Lee

Social media platforms like Facebook and Reddit host thousands of user-governed online communities. These platforms sanction communities that frequently violate platform policies; however, public perceptions of such sanctions remain…

Human-Computer Interaction · Computer Science 2024-09-10 Shagun Jhaver

The explosive growth of social media has not only revolutionized communication but also brought challenges such as political polarization, misinformation, hate speech, and echo chambers. This dissertation employs computational social…

Social and Information Networks · Computer Science 2025-09-16 Julie Jiang

Online hate speech can harmfully impact individuals and groups, specifically on non-moderated platforms such as 4chan where users can post anonymous content. This work focuses on analysing and measuring the prevalence of online hate on…

Computation and Language · Computer Science 2025-04-02 Adrian Bermudez-Villalva , Maryam Mehrnezhad , Ehsan Toreini

Recent research at the intersection of AI explainability and fairness has focused on how explanations can improve human-plus-AI task performance as assessed by fairness measures. We propose to characterize what constitutes an explanation…

Computation and Language · Computer Science 2023-10-24 Tin Nguyen , Jiannan Xu , Aayushi Roy , Hal Daumé , Marine Carpuat

The widespread use of social media necessitates reliable and efficient detection of offensive content to mitigate harmful effects. Although sophisticated models perform well on individual datasets, they often fail to generalize due to…

Computation and Language · Computer Science 2024-10-08 Huy Nghiem , Hal Daumé