English
Related papers

Related papers: Who, Why, and How: Disentangling the Effects of Mo…

200 papers

Current content moderation follows a reactive, trial-and-error approach, where interventions are applied and their effects are only measured post-hoc. In contrast, we introduce a proactive, predictive approach that enables moderators to…

Computers and Society · Computer Science 2026-02-09 Benedetta Tessa , Lorenzo Cima , Amaury Trujillo , Marco Avvenuti , Stefano Cresci

Accurately estimating how users respond to moderation interventions is paramount for developing effective and user-centred moderation strategies. However, this requires a clear understanding of which user characteristics are associated with…

Computers and Society · Computer Science 2025-10-24 Benedetta Tessa , Alejandro Moreo , Stefano Cresci , Tiziano Fagni , Fabrizio Sebastiani

This work explores the impact of moderation on users' enjoyment of conversational AI systems. While recent advancements in Large Language Models (LLMs) have led to highly capable conversational AIs that are increasingly deployed in…

Human-Computer Interaction · Computer Science 2023-04-21 Xiaoding Lu , Aleksey Korshuk , Zongyi Liu , William Beauchamp , Chai Research

Content moderation is the process of flagging content based on pre-defined platform rules. There has been a growing need for AI moderators to safeguard users as well as protect the mental health of human moderators from traumatic content.…

Computation and Language · Computer Science 2023-02-21 Meng Ye , Karan Sikka , Katherine Atwell , Sabit Hassan , Ajay Divakaran , Malihe Alikhani

We study the impact of content moderation policies in online communities. In our theoretical model, a platform chooses a content moderation policy and individuals choose whether or not to participate in the community according to the…

Data Structures and Algorithms · Computer Science 2023-10-17 Cynthia Dwork , Chris Hays , Jon Kleinberg , Manish Raghavan

While recent research has focused on developing safeguards for generative AI (GAI) model-level content safety, little is known about how content moderation to prevent malicious content performs for end-users in real-world GAI products. To…

Human-Computer Interaction · Computer Science 2025-06-18 Lan Gao , Oscar Chen , Rachel Lee , Nick Feamster , Chenhao Tan , Marshini Chetty

Online social media platforms use automated moderation systems to remove or reduce the visibility of rule-breaking content. While previous work has documented the importance of manual content moderation, the effects of automated content…

Computers and Society · Computer Science 2023-02-17 Manoel Horta Ribeiro , Justin Cheng , Robert West

Social media platforms like Facebook and Reddit host thousands of user-governed online communities. These platforms sanction communities that frequently violate platform policies; however, public perceptions of such sanctions remain…

Human-Computer Interaction · Computer Science 2024-09-10 Shagun Jhaver

Automated content moderation has long been used to help identify and filter undesired user-generated content online. But such systems have a history of incorrectly flagging content by and about marginalized identities for removal.…

Computation and Language · Computer Science 2025-07-25 Grace Proebsting , Oghenefejiro Isaacs Anigboro , Charlie M. Crawford , Danaé Metaxa , Sorelle A. Friedler

This study examines social media users' preferences for the use of platform-wide moderation in comparison to user-controlled, personalized moderation tools to regulate three categories of norm-violating content - hate speech, sexually…

Human-Computer Interaction · Computer Science 2023-02-08 Shagun Jhaver , Amy Zhang

Growing evidence shows that proactive content moderation supported by AI can help improve online discourse. However, we know little about designing these systems, how design impacts efficacy and user experience, and how people perceive…

Human-Computer Interaction · Computer Science 2024-01-22 Mark Warner , Angelika Strohmayer , Matthew Higgs , Husnain Rafiq , Liying Yang , Lynne Coventry

AI-mediated communication is increasingly being utilized to help facilitate interactions; however, in privacy sensitive domains, an AI mediator has the additional challenge of considering how to preserve privacy. In these contexts, a…

Human-Computer Interaction · Computer Science 2026-03-27 Roshni Kaushik , Maarten Sap , Koichi Onoue

Social media platforms increasingly employ proactive moderation techniques, such as detecting and curbing toxic and uncivil comments, to prevent the spread of harmful content. Despite these efforts, such approaches are often criticized for…

Human-Computer Interaction · Computer Science 2025-07-30 Xiaotian Su , Naim Zierau , Soomin Kim , April Yi Wang , Thiemo Wambsganss

Moderators of online communities often employ comment deletion as a tool. We ask here whether, beyond the positive effects of shielding a community from undesirable content, does comment removal actually cause the behavior of the comment's…

Computers and Society · Computer Science 2019-10-23 Kumar Bhargav Srinivasan , Cristian Danescu-Niculescu-Mizil , Lillian Lee , Chenhao Tan

Online communities serve as essential support channels for People Who Use Drugs (PWUD), providing access to peer support and harm reduction information. The moderation of these communities involves consequential decisions affecting member…

Human-Computer Interaction · Computer Science 2026-01-06 Kaixuan Wang , Loraine Clarke , Carl-Cyril J Dreue , Guancheng Zhou , Jason T. Jacques

Online content moderation is essential for maintaining a healthy digital environment, and reliance on AI for this task continues to grow. Consider a user comment using national stereotypes to insult a politician. This example illustrates…

Artificial Intelligence · Computer Science 2026-03-03 Houde Dong , Yifei She , Kai Ye , Liangcai Su , Chenxiong Qian , Jie Hao

We study human AI-detection behaviour at scale using a year of activity from r/RealOrAI, a Reddit community where users collaboratively assess whether visual media is real or AI-generated. The community is moderated by a bot that solicits…

Social and Information Networks · Computer Science 2026-05-26 Tuğrulcan Elmas

As more people flock to social media to connect with others and form virtual communities, it is important to research how members of these groups interact to understand human behavior on the Web. In response to an increase in hate speech,…

Social and Information Networks · Computer Science 2021-01-07 Pamela Bilo Thomas , Daniel Riehm , Maria Glenski , Tim Weninger

Online spaces involve diverse communities engaging in various forms of collaboration, which naturally give rise to discussions, some of which inevitably escalate into conflict or disputes. To address such situations, AI has primarily been…

Human-Computer Interaction · Computer Science 2025-09-16 Soobin Cho , Mark Zachry , David W. McDonald

There is an ongoing debate about how to moderate toxic speech on social media and the impact of content moderation on online discourse. This paper proposes and validates a methodology for measuring the content-moderation-induced distortions…

Social and Information Networks · Computer Science 2026-03-04 Mahyar Habibi , Dirk Hovy , Carlo Schwarz
‹ Prev 1 2 3 10 Next ›