English
Related papers

Related papers: Analyzing Norm Violations in Live-Stream Chat

200 papers

With the in-depth integration of mobile Internet and widespread adoption of social platforms, user-generated content in the Chinese cyberspace has witnessed explosive growth. Among this content, the proliferation of toxic comments poses…

Audio and Speech Processing · Electrical Eng. & Systems 2026-01-22 Ruixing Ren , Junhui Zhao , Xiaoke Sun , Qiuping Li

The rapid growth of social media in recent years has fed into some highly undesirable phenomena such as proliferation of abusive and offensive language on the Internet. Previous research suggests that such hateful content tends to come from…

Computation and Language · Computer Science 2019-02-19 Pushkar Mishra , Marco Del Tredici , Helen Yannakoudakis , Ekaterina Shutova

An increasing amount of attention has been devoted to the problem of "toxic" or antisocial behavior on social media. In this paper we analyze such behavior at very large scales: we analyze toxicity over a 14-year time span on nearly 500…

Social and Information Networks · Computer Science 2025-09-23 Katy Blumer , Jon Kleinberg

This paper investigates the effectiveness of TikTok's enforcement mechanisms for limiting the exposure of harmful content to youth accounts. We collect over 7000 videos, classify them as harmful vs not-harmful, and then simulate…

Computers and Society · Computer Science 2025-09-09 Linda Xue , Francesco Corso , Nicolo' Fontana , Geng Liu , Stefano Ceri , Francesco Pierri

Millions of mobile apps have been available through various app markets. Although most app markets have enforced a number of automated or even manual mechanisms to vet each app before it is released to the market, thousands of low-quality…

Software Engineering · Computer Science 2021-03-02 Yangyu Hu , Haoyu Wang , Tiantong Ji , Xusheng Xiao , Xiapu Luo , Peng Gao , Yao Guo

We present a survey of methods for assessing and enhancing the quality of online discussions, focusing on the potential of LLMs. While online discourses aim, at least in theory, to foster mutual understanding, they often devolve into…

The detection of sensitive content in large datasets is crucial for ensuring that shared and analysed data is free from harmful material. However, current moderation tools, such as external APIs, suffer from limitations in customisation,…

Computation and Language · Computer Science 2025-06-25 Dimosthenis Antypas , Indira Sen , Carla Perez-Almendros , Jose Camacho-Collados , Francesco Barbieri

Detecting norm violations in online communities is critical to maintaining healthy and safe spaces for online discussions. Existing machine learning approaches often struggle to adapt to the diverse rules and interpretations across…

Computation and Language · Computer Science 2024-04-18 Zihao He , Jonathan May , Kristina Lerman

Hate content in social media is ever-increasing. While Facebook, Twitter, Google have attempted to take several steps to tackle the hateful content, they have mostly been unsuccessful. Counterspeech is seen as an effective way of tackling…

Social and Information Networks · Computer Science 2019-04-08 Binny Mathew , Punyajoy Saha , Hardik Tharad , Subham Rajgaria , Prajwal Singhania , Suman Kalyan Maity , Pawan Goyal , Animesh Mukherje

Toxic sentiment analysis on Twitter (X) often focuses on specific topics and events such as politics and elections. Datasets of toxic users in such research are typically gathered through lexicon-based techniques, providing only a…

Social and Information Networks · Computer Science 2024-06-06 Hina Qayyum , Muhammad Ikram , Benjamin Zhao , Ian Wood , Mohamad Ali Kaafar , Nicolas Kourtellis

The pervasiveness of abusive content on the internet can lead to severe psychological and physical harm. Significant effort in Natural Language Processing (NLP) research has been devoted to addressing this problem through abusive content…

Computation and Language · Computer Science 2021-07-23 Svetlana Kiritchenko , Isar Nejadgholi , Kathleen C. Fraser

We present a holistic approach to building a robust and useful natural language classification system for real-world content moderation. The success of such a system relies on a chain of carefully designed and executed steps, including the…

Computation and Language · Computer Science 2023-02-16 Todor Markov , Chong Zhang , Sandhini Agarwal , Tyna Eloundou , Teddy Lee , Steven Adler , Angela Jiang , Lilian Weng

The internet has become a central medium through which `networked publics' express their opinions and engage in debate. Offensive comments and personal attacks can inhibit participation in these spaces. Automated content moderation aims to…

Computers and Society · Computer Science 2017-09-06 Reuben Binns , Michael Veale , Max Van Kleek , Nigel Shadbolt

The context-dependent nature of online aggression makes annotating large collections of data extremely difficult. Previously studied datasets in abusive language detection have been insufficient in size to efficiently train deep learning…

Computation and Language · Computer Science 2018-08-31 Younghun Lee , Seunghyun Yoon , Kyomin Jung

Violence-provoking speech -- speech that implicitly or explicitly promotes violence against the members of the targeted community, contributed to a massive surge in anti-Asian crimes during the pandemic. While previous works have…

Computation and Language · Computer Science 2024-07-23 Gaurav Verma , Rynaa Grover , Jiawei Zhou , Binny Mathew , Jordan Kraemer , Munmun De Choudhury , Srijan Kumar

Understanding toxicity in user conversations is undoubtedly an important problem. Addressing "covert" or implicit cases of toxicity is particularly hard and requires context. Very few previous studies have analysed the influence of…

Computation and Language · Computer Science 2022-10-19 Atijit Anuchitanukul , Julia Ive , Lucia Specia

Social media and online forums are increasingly becoming popular. Unfortunately, these platforms are being used for spreading hate speech. In this paper, we design black-box techniques to protect users from hate-speech on online platforms…

Cryptography and Security · Computer Science 2025-05-23 Sampanna Yashwant Kahu , Naman Ahuja

Reddit administrators have generally struggled to prevent or contain such discourse for several reasons including: (1) the inability for a handful of human administrators to track and react to millions of posts and comments per day and (2)…

Social and Information Networks · Computer Science 2019-07-01 Hussam Habib , Maaz Bin Musa , Fareed Zaffar , Rishab Nithyanand

Current content moderation follows a reactive, trial-and-error approach, where interventions are applied and their effects are only measured post-hoc. In contrast, we introduce a proactive, predictive approach that enables moderators to…

Computers and Society · Computer Science 2026-02-09 Benedetta Tessa , Lorenzo Cima , Amaury Trujillo , Marco Avvenuti , Stefano Cresci

The detection of hate speech or toxic content online is a complex and sensitive issue. While the identification itself is highly dependent on the context of the situation, sensitive personal attributes such as age, language, and nationality…

Multiagent Systems · Computer Science 2024-10-11 Jan Fillies , Theodoros Mitsikas , Ralph Schäfermeier , Adrian Paschke