English
Related papers

Related papers: Explainability and Hate Speech: Structured Explana…

200 papers

The damaging effects of hate speech on social media are evident during the last few years, and several organizations, researchers and social media platforms tried to harness them in various ways. Despite these efforts, social media users…

Information Retrieval · Computer Science 2020-05-04 Polychronis Charitidis , Stavros Doropoulos , Stavros Vologiannidis , Ioannis Papastergiou , Sophia Karakeva

Mental health forums are online communities where people express their issues and seek help from moderators and other users. In such forums, there are often posts with severe content indicating that the user is in acute distress and there…

Computation and Language · Computer Science 2017-02-23 Arman Cohan , Sydney Young , Andrew Yates , Nazli Goharian

In many online communities, community leaders (i.e., moderators and administrators) can proactively filter undesired content by requiring posts to be approved before publication. But although many communities adopt post approvals, there has…

Social and Information Networks · Computer Science 2022-05-09 Manoel Horta Ribeiro , Justin Cheng , Robert West

Despite the extensive communication benefits offered by social media platforms, numerous challenges must be addressed to ensure user safety. One of the most significant risks faced by users on these platforms is targeted hate speech. Social…

Computation and Language · Computer Science 2024-07-18 Sadar Jaf , Basel Barakat

With the recent surge and exponential growth of social media usage, scrutinizing social media content for the presence of any hateful content is of utmost importance. Researchers have been diligently working since the past decade on…

Computation and Language · Computer Science 2024-01-22 Atanu Mandal , Gargi Roy , Amit Barman , Indranil Dutta , Sudip Kumar Naskar

Social media platforms have traditionally relied on internal moderation teams and partnerships with independent fact-checking organizations to identify and flag misleading content. Recently, however, platforms including X (formerly Twitter)…

The spread of media bias is a significant concern as political discourse shapes beliefs and opinions. Addressing this challenge computationally requires improved methods for interpreting news. While large language models (LLMs) can scale…

Human-Computer Interaction · Computer Science 2026-02-24 Qile Wang , Prerana Khatiwada , Avinash Chouhan , Ashrey Mahesh , Joy Mwaria , Duy Duc Tran , Kenneth E. Barner , Matthew Louis Mauriello

Internet users highly rely on and trust web search engines, such as Google, to find relevant information online. However, scholars have documented numerous biases and inaccuracies in search outputs. To improve the quality of search results,…

Computers and Society · Computer Science 2023-10-06 Aleksandra Urman , Aniko Hannak , Mykola Makhortykh

Incivility on platforms such as Twitter (now X) and Reddit complicates the development of AI systems that can support productive, rhetorically sound political argumentation. We present experiments with \textit{GPT-3.5 Turbo} fine-tuned on…

Computation and Language · Computer Science 2025-11-04 Svetlana Churina , Kokil Jaidka

Politeness is a core dimension of human communication, yet its role in human-AI information seeking remains underexplored. We investigate how user politeness behaviour shapes conversational outcomes in a cooking-assistance setting. First,…

Human-Computer Interaction · Computer Science 2026-01-16 David Elsweiler , Christine Elsweiler , Anna Ziegner

Flagging mechanisms on social media platforms allow users to report inappropriate posts/accounts for review by content moderators. These reports are pivotal to platforms' efforts toward regulating norm violations. This paper examines how…

Human-Computer Interaction · Computer Science 2024-12-20 Yunhee Shim , Shagun Jhaver

Large Language Models (LLMs) are increasingly deployed as gateways to information, yet their content moderation practices remain underexplored. This work investigates the extent to which LLMs refuse to answer or omit information when…

Computation and Language · Computer Science 2025-04-08 Sander Noels , Guillaume Bied , Maarten Buyl , Alexander Rogiers , Yousra Fettach , Jefrey Lijffijt , Tijl De Bie

Content moderators review disturbing content to protect social media users, often at significant cost to their mental health. Recent reports document the mental health conditions of African moderators as notably problematic. Beyond the…

Content moderation (removing or limiting the distribution of posts based on their contents) is one tool social networks use to fight problems such as harassment and disinformation. Manually screening all content is usually impractical given…

Information Retrieval · Computer Science 2021-08-31 Eugene Yang , David D. Lewis , Ophir Frieder

The rapid growth of social media in recent years has fed into some highly undesirable phenomena such as proliferation of abusive and offensive language on the Internet. Previous research suggests that such hateful content tends to come from…

Computation and Language · Computer Science 2019-02-19 Pushkar Mishra , Marco Del Tredici , Helen Yannakoudakis , Ekaterina Shutova

Detecting and classifying instances of hate in social media text has been a problem of interest in Natural Language Processing in the recent years. Our work leverages state of the art Transformer language models to identify hate speech in a…

Computation and Language · Computer Science 2021-01-12 Sayar Ghosh Roy , Ujwal Narayan , Tathagata Raha , Zubair Abid , Vasudeva Varma

The content of today's social media is becoming more and more rich, increasingly mixing text, images, videos, and audio. It is an intriguing research question to model the interplay between these different modes in attracting user attention…

Social and Information Networks · Computer Science 2017-03-07 Jack Hessel , Lillian Lee , David Mimno

Decisions in organizations are about evaluating alternatives and choosing the one that would best serve organizational goals. To the extent that the evaluation of alternatives could be formulated as a predictive task with appropriate…

Human-Computer Interaction · Computer Science 2022-06-30 Charles Wan , Rodrigo Belo , Leid Zejnilović

Content moderation practices and technologies need to change over time as requirements and community expectations shift. However, attempts to restructure existing moderation practices can be difficult, especially for platforms that rely on…

Human-Computer Interaction · Computer Science 2026-03-10 Chau Tran , Kejsi Take , Kaylea Champion , Benjamin Mako Hill , Rachel Greenstadt

While social media empowers freedom of expression and individual voices, it also enables anti-social behavior, online harassment, cyberbullying, and hate speech. In this paper, we deepen our understanding of online hate speech by focusing…

Computation and Language · Computer Science 2018-04-13 Mai ElSherief , Vivek Kulkarni , Dana Nguyen , William Yang Wang , Elizabeth Belding