中文
相关论文

相关论文: Impact Of Content Features For Automatic Online Ab…

200 篇论文

Effective content moderation systems require explicit classification criteria, yet online communities like subreddits often operate with diverse, implicit standards. This work introduces a novel approach to identify and extract these…

计算与语言 · 计算机科学 2025-09-04 Youngwoo Kim , Himanshu Beniwal , Steven L. Johnson , Thomas Hartvigsen

The detection of hate speech or toxic content online is a complex and sensitive issue. While the identification itself is highly dependent on the context of the situation, sensitive personal attributes such as age, language, and nationality…

多智能体系统 · 计算机科学 2024-10-11 Jan Fillies , Theodoros Mitsikas , Ralph Schäfermeier , Adrian Paschke

Recent advances in text mining and natural language processing technology have enabled researchers to detect an authors identity or demographic characteristics, such as age and gender, in several text genres by automatically analysing the…

密码学与安全 · 计算机科学 2022-11-30 Claudia Peersman , Matthew Edwards , Emma Williams , Awais Rashid

Offensive speech detection is a key component of content moderation. However, what is offensive can be highly subjective. This paper investigates how machine and human moderators disagree on what is offensive when it comes to real-world…

Incivility remains a major challenge for online discussion platforms, to such an extent that even conversations between well-intentioned users can often derail into uncivil behavior. Traditionally, platforms have relied on moderators to --…

人机交互 · 计算机科学 2022-12-06 Jonathan P. Chang , Charlotte Schluger , Cristian Danescu-Niculescu-Mizil

Hate speech, offensive language, sexism, racism and other types of abusive behavior have become a common phenomenon in many online social media platforms. In recent years, such diverse abusive behaviors have been manifesting with increased…

Understanding interpersonal communication requires, in part, understanding the social context and norms in which a message is said. However, current methods for identifying offensive content in such communication largely operate independent…

计算与语言 · 计算机科学 2023-07-07 David Jurgens , Agrima Seth , Jackson Sargent , Athena Aghighi , Michael Geraci

Online communities serve as essential support channels for People Who Use Drugs (PWUD), providing access to peer support and harm reduction information. The moderation of these communities involves consequential decisions affecting member…

人机交互 · 计算机科学 2026-01-06 Kaixuan Wang , Loraine Clarke , Carl-Cyril J Dreue , Guancheng Zhou , Jason T. Jacques

Social media companies as well as authorities make extensive use of artificial intelligence (AI) tools to monitor postings of hate speech, celebrations of violence or profanity. Since AI software requires massive volumes of data to train…

计算与语言 · 计算机科学 2021-09-30 Hadeel Saadany , Constantin Orasan

Despite the common use of rule-based tools for online content moderation, human moderators still spend a lot of time monitoring them to ensure that they work as intended. Based on surveys and interviews with Reddit moderators who use…

人机交互 · 计算机科学 2025-02-21 Jean Y. Song , Sangwook Lee , Jisoo Lee , Mina Kim , Juho Kim

This study examines social media users' preferences for the use of platform-wide moderation in comparison to user-controlled, personalized moderation tools to regulate three categories of norm-violating content - hate speech, sexually…

人机交互 · 计算机科学 2023-02-08 Shagun Jhaver , Amy Zhang

With rising concern around abusive and hateful behavior on social media platforms, we present an ensemble learning method to identify and analyze the linguistic properties of such content. Our stacked ensemble comprises of three machine…

计算与语言 · 计算机科学 2020-06-08 Gaurav Verma , Niyati Chhaya , Vishwa Vinay

Although automated harmful content detection systems are frequently used to monitor online platforms, moderators and end users frequently cannot understand the logic underlying their predictions. While recent studies have focused on…

计算与语言 · 计算机科学 2026-03-20 Trishita Dhara , Siddhesh Sheth

Social media systems rely on user feedback and rating mechanisms for personalization, ranking, and content filtering. However, when users evaluate content contributed by fellow users (e.g., by liking a post or voting on a comment), these…

社会与信息网络 · 计算机科学 2014-05-08 Justin Cheng , Cristian Danescu-Niculescu-Mizil , Jure Leskovec

Now-a-days, derogatory comments are often made by one another, not only in offline environment but also immensely in online environments like social networking websites and online communities. So, an Identification combined with Prevention…

计算与语言 · 计算机科学 2019-03-19 Navoneel Chakrabarty

To support safety and inclusion in online communications, significant efforts in NLP research have been put towards addressing the problem of abusive content detection, commonly defined as a supervised classification task. The research…

计算与语言 · 计算机科学 2020-10-29 Svetlana Kiritchenko , Isar Nejadgholi

Social bots play a significant role in many online social networks (OSN) as they imitate human behavior. This fact raises difficult questions about their capabilities and potential risks. Given the recent advances in Generative AI (GenAI),…

社会与信息网络 · 计算机科学 2024-05-06 Shaghayegh Najari , Davood Rafiee , Mostafa Salehi , Reza Farahbakhsh

Abusive language is a massive problem in online social platforms. Existing abusive language detection techniques are particularly ill-suited to comments containing heterogeneous abusive language patterns, i.e., both abusive and non-abusive…

计算与语言 · 计算机科学 2021-05-25 Hongyu Gong , Alberto Valido , Katherine M. Ingram , Giulia Fanti , Suma Bhat , Dorothy L. Espelage

Flagging mechanisms on social media platforms allow users to report inappropriate posts/accounts for review by content moderators. These reports are pivotal to platforms' efforts toward regulating norm violations. This paper examines how…

人机交互 · 计算机科学 2024-12-20 Yunhee Shim , Shagun Jhaver

A major task for moderators of online spaces is norm-setting, essentially creating shared norms for user behavior in their communities. Platform design principles emphasize the importance of highlighting norm-adhering examples and…

人机交互 · 计算机科学 2026-04-07 Agam Goyal , Charlotte Lambert , Yoshee Jain , Eshwar Chandrasekharan