English
Related papers

Related papers: Generating Counter Narratives against Online Hate …

200 papers

The damaging effects of hate speech on social media are evident during the last few years, and several organizations, researchers and social media platforms tried to harness them in various ways. Despite these efforts, social media users…

Information Retrieval · Computer Science 2020-05-04 Polychronis Charitidis , Stavros Doropoulos , Stavros Vologiannidis , Ioannis Papastergiou , Sophia Karakeva

Abusive speech on social media poses a persistent and evolving challenge, driven by the continuous emergence of novel slang and obfuscated terms designed to circumvent detection systems. In this work, we present a data efficient strategy…

Computation and Language · Computer Science 2025-12-03 Pritish N. Desai , Tanay Kewalramani , Srimanta Mandal

The phenomenal growth on the internet has helped in empowering individual's expressions, but the misuse of freedom of expression has also led to the increase of various cyber crimes and anti-social activities. Hate speech is one such issue…

Computation and Language · Computer Science 2020-06-01 Prashant Kapil , Asif Ekbal , Dipankar Das

Recently, many studies have tried to create generation models to assist counter speakers by providing counterspeech suggestions for combating the explosive proliferation of online hate. However, since these suggestions are from a vanilla…

Computation and Language · Computer Science 2022-05-10 Punyajoy Saha , Kanishk Singh , Adarsh Kumar , Binny Mathew , Animesh Mukherjee

This work presents a thorough review concerning recent studies and text generation advancements using Generative Adversarial Networks. The usage of adversarial learning for text generation is promising as it provides alternatives to…

Computation and Language · Computer Science 2022-12-22 Gustavo Henrique de Rosa , João Paulo Papa

Machine learning (ML)-based content moderation tools are essential to keep online spaces free from hateful communication. Yet, ML tools can only be as capable as the quality of the data they are trained on allows them. While there is…

Computation and Language · Computer Science 2024-06-14 Zehui Yu , Indira Sen , Dennis Assenmacher , Mattia Samory , Leon Fröhling , Christina Dahn , Debora Nozza , Claudia Wagner

Automated content moderation has long been used to help identify and filter undesired user-generated content online. But such systems have a history of incorrectly flagging content by and about marginalized identities for removal.…

Computation and Language · Computer Science 2025-07-25 Grace Proebsting , Oghenefejiro Isaacs Anigboro , Charlie M. Crawford , Danaé Metaxa , Sorelle A. Friedler

Hate speech and misinformation frequently co-occur online, amplifying prejudice and polarization. Given their scale, using Large Language Models (LLMs) to assist expert counterspeech (CS) writing has gained interest, yet prior work has…

Computation and Language · Computer Science 2026-05-22 Genoveffa Martone , Helena Bonaldi , Marco Guerini

The online hate speech is proliferating with several organization and countries implementing laws to ban such harmful speech. While these restrictions might reduce the amount of such hateful content, it does so by restricting freedom of…

Social and Information Networks · Computer Science 2018-12-07 Binny Mathew , Navish Kumar , Ravina , Pawan Goyal , Animesh Mukherjee

The dissemination of online hate speech can have serious negative consequences for individuals, online communities, and entire societies. This and the large volume of hateful online content prompted both practitioners', i.e., in content…

Computation and Language · Computer Science 2025-04-14 Julian Bäumler , Louis Blöcher , Lars-Joel Frey , Xian Chen , Markus Bayer , Christian Reuter

Social media has seen a worrying rise in hate speech in recent times. Branching to several distinct categories of cyberbullying, gender discrimination, or racism, the combined label for such derogatory content can be classified as toxic…

Computation and Language · Computer Science 2022-01-11 Sourav Das , Prasanta Mandal , Sanjay Chatterji

Counterspeech, i.e. the practice of responding to online hate speech, has gained traction in NLP as a promising intervention. While early work emphasised collaboration with non-governmental organisation stakeholders, recent research trends…

Computation and Language · Computer Science 2025-08-07 Tanvi Dinkar , Aiqi Jiang , Simona Frenda , Poppy Gerrard-Abbott , Nancie Gunson , Gavin Abercrombie , Ioannis Konstas

Hate Speech takes many forms to target communities with derogatory comments, and takes humanity a step back in societal progress. HateXplain is a recently published and first dataset to use annotated spans in the form of rationales, along…

Computation and Language · Computer Science 2022-08-10 Arvind Subramaniam , Aryan Mehra , Sayani Kundu

User-generated replies to hate speech are promising means to combat hatred, but questions about whether they can stop incivility in follow-up conversations linger. We argue that effective replies stop incivility from emerging in follow-up…

Computers and Society · Computer Science 2023-12-11 Xinchen Yu , Eduardo Blanco , Lingzi Hong

With the proliferation of social media, accurate detection of hate speech has become critical to ensure safety online. To combat nuanced forms of hate speech, it is important to identify and thoroughly explain hate speech to help users…

Computation and Language · Computer Science 2023-11-23 Yongjin Yang , Joonkee Kim , Yujin Kim , Namgyu Ho , James Thorne , Se-young Yun

Counterspeech is a key strategy against harmful online content, but scaling expert-driven efforts is challenging. Large Language Models (LLMs) present a potential solution, though their use in countering conspiracy theories is…

Computation and Language · Computer Science 2025-08-04 Mareike Lisker , Christina Gottschalk , Helena Mihaljević

Despite regulations imposed by nations and social media platforms, e.g. (Government of India, 2021; European Parliament and Council of the European Union, 2022), inter alia, hateful content persists as a significant challenge. Existing…

Automatic detection of online hate speech serves as a crucial step in the detoxification of the online discourse. Moreover, accurate classification can promote a better understanding of the proliferation of hate as a social phenomenon.…

Computation and Language · Computer Science 2024-09-24 Tom Marzea , Abraham Israeli , Oren Tsur

Sophisticated language models such as OpenAI's GPT-3 can generate hateful text that targets marginalized groups. Given this capacity, we are interested in whether large language models can be used to identify hate speech and classify text…

Computation and Language · Computer Science 2022-03-25 Ke-Li Chiu , Annie Collins , Rohan Alexander

To protect users from massive hateful content, existing works studied automated hate speech detection. Despite the existing efforts, one question remains: do automated hate speech detectors conform to social media content policies? A…

Software Engineering · Computer Science 2024-03-20 Jiangrui Zheng , Xueqing Liu , Guanqun Yang , Mirazul Haque , Xing Qian , Ravishka Rathnasuriya , Wei Yang , Girish Budhrani
‹ Prev 1 4 5 6 7 8 10 Next ›