中文
相关论文

相关论文: Like trainer, like bot? Inheritance of bias in alg…

200 篇论文

Crowdsourced annotation is vital to both collecting labelled data to train and test automated content moderation systems and to support human-in-the-loop review of system decisions. However, annotation tasks such as judging hate speech are…

Developing machine learning models to characterize political polarization on online social media presents significant challenges. These challenges mainly stem from various factors such as the lack of annotated data, presence of noise in…

社会与信息网络 · 计算机科学 2023-11-22 Sadia Kamal , Brenner Little , Jade Gullic , Trevor Harms , Kristin Olofsson , Arunkumar Bagavathi

As robots and other intelligent agents move from simple environments and problems to more complex, unstructured settings, manually programming their behavior has become increasingly challenging and expensive. Often, it is easier for a…

机器人学 · 计算机科学 2018-11-19 Takayuki Osa , Joni Pajarinen , Gerhard Neumann , J. Andrew Bagnell , Pieter Abbeel , Jan Peters

There is strong agreement that generative AI should be regulated, but strong disagreement on how to approach regulation. While some argue that AI regulation should mostly rely on extensions of existing laws, others argue that entirely new…

计算机与社会 · 计算机科学 2024-12-17 Ruth Elisabeth Appel

In recent years, social media companies have grappled with defining and enforcing content moderation policies surrounding political content on their platforms, due in part to concerns about political bias, disinformation, and polarization.…

人机交互 · 计算机科学 2023-05-25 Jacob Thebault-Spieker , Sukrit Venkatagiri , Naomi Mine , Kurt Luther

Hateful rhetoric is plaguing online discourse, fostering extreme societal movements and possibly giving rise to real-world violence. A potential solution to this growing global problem is citizen-generated counter speech where citizens…

计算机与社会 · 计算机科学 2020-06-09 Joshua Garland , Keyan Ghazi-Zahedi , Jean-Gabriel Young , Laurent Hébert-Dufresne , Mirta Galesic

Social media websites, electronic newspapers and Internet forums allow visitors to leave comments for others to read and interact. This exchange is not free from participants with malicious intentions, who troll others by positing messages…

计算与语言 · 计算机科学 2016-12-30 Luis Gerardo Mojica

Value alignment is the task of creating autonomous systems whose values align with those of humans. Past work has shown that stories are a potentially rich source of information on human values; however, past work has been limited to…

计算与语言 · 计算机科学 2022-12-13 Md Sultan Al Nahian , Spencer Frazier , Brent Harrison , Mark Riedl

The ability of Natural Language Processing (NLP) methods to categorize text into multiple classes has motivated their use in online content moderation tasks, such as hate speech and fake news detection. However, there is limited…

计算与语言 · 计算机科学 2025-03-11 Neemesh Yadav , Jiarui Liu , Francesco Ortu , Roya Ensafi , Zhijing Jin , Rada Mihalcea

Toxic comment classification models are often found biased toward identity terms which are terms characterizing a specific group of people such as "Muslim" and "black". Such bias is commonly reflected in false-positive predictions, i.e.…

计算与语言 · 计算机科学 2022-10-18 Zhixue Zhao , Ziqi Zhang , Frank Hopfgartner

Machine learning has significantly enhanced the abilities of robots, enabling them to perform a wide range of tasks in human environments and adapt to our uncertain real world. Recent works in various machine learning domains have…

机器人学 · 计算机科学 2023-10-31 Laura Londoño , Juana Valeria Hurtado , Nora Hertz , Philipp Kellmeyer , Silja Voeneky , Abhinav Valada

Though detection systems have been developed to identify obscene content such as pornography and violence, artificial intelligence is simply not good enough to fully automate this task yet. Due to the need for manual verification, social…

人机交互 · 计算机科学 2020-01-07 Brandon Dang , Martin J. Riedl , Matthew Lease

To meet the demands of content moderation, online platforms have resorted to automated systems. Newer forms of real-time engagement($\textit{e.g.}$, users commenting on live streams) on platforms like Twitch exert additional pressures on…

计算与语言 · 计算机科学 2025-06-11 Prarabdh Shukla , Wei Yin Chong , Yash Patel , Brennan Schaffner , Danish Pruthi , Arjun Bhagoji

Online conversations can be toxic and subjected to threats, abuse, or harassment. To identify toxic text comments, several deep learning and machine learning models have been proposed throughout the years. However, recent studies…

机器学习 · 计算机科学 2023-11-09 Md Azim Khan

We deal with the problem of localized in-video taxonomic human annotation in the video content moderation domain, where the goal is to identify video segments that violate granular policies, e.g., community guidelines on an online video…

机器学习 · 计算机科学 2022-10-19 Meghana Deodhar , Xiao Ma , Yixin Cai , Alex Koes , Alex Beutel , Jilin Chen

This work focuses on the nature of visibility in societies where the behaviours of humans and algorithms influence each other - termed algorithmically infused societies. We propose a quantitative measure of visibility, with implications and…

社会与信息网络 · 计算机科学 2024-12-09 Shaojing Sun , Zhiyuan Liu , David Waxman

Content moderation is a central mechanism through which platforms attempt to balance user engagement with community governance. Yet existing research has largely treated moderation as a uniform intervention, overlooking how moderator…

计算机与社会 · 计算机科学 2026-05-18 Siyi Zhou , Lindsay Young , Marlon Twyman , Emilio Ferrara

Much of our modern digital infrastructure relies critically upon open sourced software. The communities responsible for building this cyberinfrastructure require maintenance and moderation, which is often supported by volunteer efforts.…

人机交互 · 计算机科学 2023-08-16 Jane Hsieh , Joselyn Kim , Laura Dabbish , Haiyi Zhu

We describe the current content moderation strategy employed by Meta to remove policy-violating content from its platforms. Meta relies on both handcrafted and learned risk models to flag potentially violating content for human review. Our…

Public participation is indispensable for an insightful understanding of the ethics issues raised by AI technologies. Twitter is selected in this paper to serve as an online public sphere for exploring discourse on AI ethics, facilitating…

计算机与社会 · 计算机科学 2024-06-21 Mengyi Wei , Puzhen Zhang , Chuan Chen , Dongsheng Chen , Chenyu Zuo , Liqiu Meng