中文
相关论文

相关论文: Venire: A Machine Learning-Guided Panel Review Sys…

200 篇论文

Over the past decade, social media platforms have been key in spreading rumors, leading to significant negative impacts. To counter this, the community has developed various Rumor Detection (RD) algorithms to automatically identify them…

计算与语言 · 计算机科学 2025-04-08 Bing Wang , Bingrui Zhao , Ximing Li , Changchun Li , Wanfu Gao , Shengsheng Wang

Current content moderation follows a reactive, trial-and-error approach, where interventions are applied and their effects are only measured post-hoc. In contrast, we introduce a proactive, predictive approach that enables moderators to…

计算机与社会 · 计算机科学 2026-02-09 Benedetta Tessa , Lorenzo Cima , Amaury Trujillo , Marco Avvenuti , Stefano Cresci

We deal with the problem of localized in-video taxonomic human annotation in the video content moderation domain, where the goal is to identify video segments that violate granular policies, e.g., community guidelines on an online video…

机器学习 · 计算机科学 2022-10-19 Meghana Deodhar , Xiao Ma , Yixin Cai , Alex Koes , Alex Beutel , Jilin Chen

Whose labels should a machine learning (ML) algorithm learn to emulate? For ML tasks ranging from online comment toxicity to misinformation detection to medical diagnosis, different groups in society may have irreconcilable disagreements…

Moderating content in social media platforms is a formidable challenge due to the unprecedented scale of such systems, which typically handle billions of posts per day. Some of the largest platforms such as Facebook blend machine learning…

Social media platforms face increasing scrutiny over the rapid spread of misinformation. In response, many have adopted community-based content moderation systems, including Community Notes (formerly Birdwatch) on X (formerly Twitter),…

社会与信息网络 · 计算机科学 2026-02-02 Gabriela Juncosa , Saeedeh Mohammadi , Margaret Samahita , Taha Yasseri

We study the impact of content moderation policies in online communities. In our theoretical model, a platform chooses a content moderation policy and individuals choose whether or not to participate in the community according to the…

数据结构与算法 · 计算机科学 2023-10-17 Cynthia Dwork , Chris Hays , Jon Kleinberg , Manish Raghavan

Effective content moderation systems require explicit classification criteria, yet online communities like subreddits often operate with diverse, implicit standards. This work introduces a novel approach to identify and extract these…

计算与语言 · 计算机科学 2025-09-04 Youngwoo Kim , Himanshu Beniwal , Steven L. Johnson , Thomas Hartvigsen

One trending application of LLM (large language model) is to use it for content moderation in online platforms. Most current studies on this application have focused on the metric of accuracy -- the extent to which LLMs make correct…

计算机与社会 · 计算机科学 2025-06-03 Tao Huang

Online spaces involve diverse communities engaging in various forms of collaboration, which naturally give rise to discussions, some of which inevitably escalate into conflict or disputes. To address such situations, AI has primarily been…

人机交互 · 计算机科学 2025-09-16 Soobin Cho , Mark Zachry , David W. McDonald

Hate speech moderation remains a challenging task for social media platforms. Human-AI collaborative systems offer the potential to combine the strengths of humans' reliability and the scalability of machine learning to tackle this issue…

人机交互 · 计算机科学 2023-07-25 Philippe Lammerts , Philip Lippmann , Yen-Chia Hsu , Fabio Casati , Jie Yang

Despite the common use of rule-based tools for online content moderation, human moderators still spend a lot of time monitoring them to ensure that they work as intended. Based on surveys and interviews with Reddit moderators who use…

人机交互 · 计算机科学 2025-02-21 Jean Y. Song , Sangwook Lee , Jisoo Lee , Mina Kim , Juho Kim

Incivility on platforms such as Twitter (now X) and Reddit complicates the development of AI systems that can support productive, rhetorically sound political argumentation. We present experiments with \textit{GPT-3.5 Turbo} fine-tuned on…

计算与语言 · 计算机科学 2025-11-04 Svetlana Churina , Kokil Jaidka

Human feedback is essential for building human-centered AI systems across domains where disagreement is prevalent, such as AI safety, content moderation, or sentiment analysis. Many disagreements, particularly in politically charged…

Content moderation is often performed by a collaboration between humans and machine learning models. However, it is not well understood how to design the collaborative process so as to maximize the combined moderator-model system…

机器学习 · 计算机科学 2021-07-12 Ian D. Kivlichan , Zi Lin , Jeremiah Liu , Lucy Vasserman

The prevalence and impact of toxic discussions online have made content moderation crucial.Automated systems can play a vital role in identifying toxicity, and reducing the reliance on human moderation.Nevertheless, identifying toxic…

Reddit relies on volunteer moderators to enforce community rules, configure tools, and review flagged content. This labor is substantial, worth millions in unpaid effort, and increasingly hard to sustain as communities grow. While recent…

人机交互 · 计算机科学 2025-10-07 Tanvi Bajpai , Eshwar Chandrasekharan

Developing a strong community requires empowered leadership capable of overcoming governance challenges. New online platforms have given users opportunities to practice governance through content moderation roles. The over 2.8 million…

计算机与社会 · 计算机科学 2022-06-14 Hannah M Wang , Beril Bulat , Stephen Fujimoto , Seth Frey

Detecting what content communities value is a foundational challenge for social computing systems -- from feed curation and content ranking to moderation tools and personalized recommendation systems. Yet existing approaches remain…

人机交互 · 计算机科学 2026-01-21 Agam Goyal , Xianyang Zhan , Charlotte Lambert , Koustuv Saha , Eshwar Chandrasekharan

Reddit administrators have generally struggled to prevent or contain such discourse for several reasons including: (1) the inability for a handful of human administrators to track and react to millions of posts and comments per day and (2)…

社会与信息网络 · 计算机科学 2019-07-01 Hussam Habib , Maaz Bin Musa , Fareed Zaffar , Rishab Nithyanand