English
Related papers

Related papers: Demarked: A Strategy for Enhanced Abusive Speech M…

200 papers

While social media offer great communication opportunities, they also increase the vulnerability of young people to threatening situations online. Recent studies report that cyberbullying constitutes a growing problem among youngsters.…

Computation and Language · Computer Science 2020-03-03 Cynthia Van Hee , Gilles Jacobs , Chris Emmery , Bart Desmet , Els Lefever , Ben Verhoeven , Guy De Pauw , Walter Daelemans , Véronique Hoste

We present the first English corpus study on abusive language towards three conversational AI systems gathered "in the wild": an open-domain social bot, a rule-based chatbot, and a task-based system. To account for the complexity of the…

Computation and Language · Computer Science 2021-09-21 Amanda Cercas Curry , Gavin Abercrombie , Verena Rieser

The evolution of digital communication systems and the designs of online platforms have inadvertently facilitated the subconscious propagation of toxic behavior. Giving rise to reactive responses to toxic behavior. Toxicity in online…

Computers and Society · Computer Science 2025-10-01 Smita Khapre , Melkamu Abay Mersha , Hassan Shakil , Jonali Baruah , Jugal Kalita

With the widespread use of toxic language online, platforms are increasingly using automated systems that leverage advances in natural language processing to automatically flag and remove toxic comments. However, most automated systems --…

Human-Computer Interaction · Computer Science 2021-02-11 Austin P Wright , Omar Shaikh , Haekyu Park , Will Epperson , Muhammed Ahmed , Stephane Pinel , Duen Horng Chau , Diyi Yang

Hate speech detection is a crucial task, especially on social media, where harmful content can spread quickly. Implementing machine learning models to automatically identify and address hate speech is essential for mitigating its impact and…

Computation and Language · Computer Science 2025-08-19 Somaiyeh Dehghan , Mehmet Umut Sen , Berrin Yanikoglu

The damaging effects of hate speech on social media are evident during the last few years, and several organizations, researchers and social media platforms tried to harness them in various ways. Despite these efforts, social media users…

Information Retrieval · Computer Science 2020-05-04 Polychronis Charitidis , Stavros Doropoulos , Stavros Vologiannidis , Ioannis Papastergiou , Sophia Karakeva

In response to the escalating threat of misinformation, social media platforms have introduced a wide range of interventions aimed at reducing the spread and influence of false information. However, there is a lack of a coherent macrolevel…

Computers and Society · Computer Science 2025-10-21 Amir Karami

Cyberbullying is of extreme prevalence today. Online-hate comments, toxicity, cyberbullying amongst children and other vulnerable groups are only growing over online classes, and increased access to social platforms, especially post…

Computation and Language · Computer Science 2021-07-20 Bhumika Bhatia , Anuj Verma , Anjum , Rahul Katarya

Citizen-generated counter speech is a promising way to fight hate speech and promote peaceful, non-polarized discourse. However, there is a lack of large-scale longitudinal studies of its effectiveness for reducing hate speech. To this end,…

Social and Information Networks · Computer Science 2021-09-07 Joshua Garland , Keyan Ghazi-Zahedi , Jean-Gabriel Young , Laurent Hébert-Dufresne , Mirta Galesic

AI-generated counterspeech offers a promising and scalable strategy to curb online toxicity through direct replies that promote civil discourse. However, current counterspeech is one-size-fits-all, lacking adaptation to the moderation…

Human-Computer Interaction · Computer Science 2025-02-10 Lorenzo Cima , Alessio Miaschi , Amaury Trujillo , Marco Avvenuti , Felice Dell'Orletta , Stefano Cresci

How should conversational agents respond to verbal abuse through the user? To answer this question, we conduct a large-scale crowd-sourced evaluation of abuse response strategies employed by current state-of-the-art systems. Our results…

Human-Computer Interaction · Computer Science 2019-09-11 Amanda Cercas Curry , Verena Rieser

Automated hate speech detection is an important tool in combating the spread of hate speech, particularly in social media. Numerous methods have been developed for the task, including a recent proliferation of deep-learning based…

Computation and Language · Computer Science 2023-12-08 Jitendra Singh Malik , Hezhe Qiao , Guansong Pang , Anton van den Hengel

A rapidly increasing amount of human conversation occurs online. But divisiveness and conflict can fester in text-based interactions on social media platforms, in messaging apps, and on other digital forums. Such toxicity increases…

Human-Computer Interaction · Computer Science 2023-10-24 Lisa P. Argyle , Ethan Busby , Joshua Gubler , Chris Bail , Thomas Howe , Christopher Rytting , David Wingate

Communication on social media platforms is not only culturally and politically relevant, it is also increasingly widespread across societies. Users not only communicate via social media platforms, but also search specifically for…

Social and Information Networks · Computer Science 2020-11-30 Dennis Klinkhammer

This article proposes a framework for content moderation based on a decentralized market where no one party, neither governments nor firms, controls the flow of information.

Cryptography and Security · Computer Science 2023-09-19 Marshall Van Alstyne , Michael David Smith , Herb Lin

Toxic Online Content (TOC) includes messages on digital platforms that are harmful, hostile, or damaging to constructive public discourse. Individuals, organizations, and LLMs respond to TOC through counterspeech or counternarrative…

Computers and Society · Computer Science 2025-10-15 Lisa Schirch , Kristina Radivojevic , Cathy Buerger

The mental health of social media users has started more and more to be put at risk by harmful, hateful, and offensive content. In this paper, we propose \textsc{StopHC}, a harmful content detection and mitigation architecture for social…

Social and Information Networks · Computer Science 2024-11-12 Ciprian-Octavian Truică , Ana-Teodora Constantinescu , Elena-Simona Apostol

Language models, while capable of generating remarkably coherent and seemingly accurate text, can occasionally produce undesirable content, including harmful or toxic outputs. In this paper, we present a new two-stage approach to detect and…

Machine Learning · Computer Science 2025-11-11 Bao Nguyen , Binh Nguyen , Duy Nguyen , Viet Anh Nguyen

Text detoxification is a conditional text generation task aiming to remove offensive content from toxic text. It is highly useful for online forums and social media, where offensive content is frequently encountered. Intuitively, there are…

Computation and Language · Computer Science 2023-06-16 Griffin Floto , Mohammad Mahdi Abdollah Pour , Parsa Farinneya , Zhenwei Tang , Ali Pesaranghader , Manasa Bharadwaj , Scott Sanner

Social media platforms have become central to modern communication, yet they also harbor offensive content that challenges platform safety and inclusivity. While prior research has primarily focused on textual indicators of offense, the…

Computation and Language · Computer Science 2025-06-03 Yuhang Zhou , Yimin Xiao , Wei Ai , Ge Gao