English
Related papers

Related papers: Toxic Bias: Perspective API Misreads German as Mor…

200 papers

If large language models like GPT-3 preferably produce a particular point of view, they may influence people's opinions on an unknown scale. This study investigates whether a language-model-powered writing assistant that generates some…

Human-Computer Interaction · Computer Science 2023-02-02 Maurice Jakesch , Advait Bhat , Daniel Buschek , Lior Zalmanson , Mor Naaman

Positive, supportive online communication in social media (candy speech) has the potential to foster civility, yet automated detection of such language remains underexplored, limiting systematic analysis of its impact. We investigate how…

Computation and Language · Computer Science 2025-09-17 Christian Rene Thelen , Patrick Gustav Blaneck , Tobias Bornheim , Niklas Grieger , Stephan Bialonski

Large Language Models (LLMs) are increasingly embedded in autonomous agents that engage, converse, and co-evolve in online social platforms. While prior work has documented the generation of toxic content by LLMs, far less is known about…

Multiagent Systems · Computer Science 2026-01-22 Erica Coppolillo , Luca Luceri , Emilio Ferrara

We use structural topic modeling to examine racial bias in data collected to train models to detect hate speech and abusive language in social media posts. We augment the abusive language dataset by adding an additional feature indicating…

Computation and Language · Computer Science 2020-05-28 Thomas Davidson , Debasmita Bhattacharya

Google's Vision API analyses images and provides a variety of output predictions, one such type is context-based labelling. In this paper, it is shown that adversarial examples that cause incorrect label prediction and spoofing can be…

Computer Vision and Pattern Recognition · Computer Science 2019-11-19 Aman Apte , Aritra Bandyopadhyay , K Akhilesh Shenoy , Jason Peter Andrews , Aditya Rathod , Manish Agnihotri , Aditya Jajodia

Marking biased texts is a practical approach to increase media bias awareness among news consumers. However, little is known about the generalizability of such awareness to new topics or unmarked news articles, and the role of…

Human-Computer Interaction · Computer Science 2024-12-31 Timo Spinde , Fei Wu , Wolfgang Gaissmaier , Gianluca Demartini , Helge Giese

Online hate speech poses a serious threat to individual well-being and societal cohesion. A promising solution to curb online hate speech is counterspeech. Counterspeech is aimed at encouraging users to reconsider hateful posts by direct…

Social and Information Networks · Computer Science 2024-11-26 Dominik Bär , Abdurahman Maarouf , Stefan Feuerriegel

The rapid progress of generative AI has enabled remarkable creative capabilities, yet it also raises urgent concerns regarding the safety of AI-generated visual content in real-world applications such as content moderation, platform…

Computer Vision and Pattern Recognition · Computer Science 2025-06-05 Qiang Fu , Zonglei Jing , Zonghao Ying , Xiaoqian Li

The presence of toxic content has become a major problem for many online communities. Moderators try to limit this problem by implementing more and more refined comment filters, but toxic users are constantly finding new ways to circumvent…

Computation and Language · Computer Science 2018-12-06 Éloi Brassard-Gourdeau , Richard Khoury

Recent advances in generative Artificial Intelligence have raised public awareness, shaping expectations and concerns about their societal implications. Central to these debates is the question of AI alignment -- how well AI systems meet…

Computers and Society · Computer Science 2025-04-18 Andreas Jungherr , Adrian Rauchfleisch

In this era of digitization, knowing the user's sociolect aspects have become essential features to build the user specific recommendation systems. These sociolect aspects could be found by mining the user's language sharing in the form of…

Computation and Language · Computer Science 2018-04-13 Barathi Ganesh HB , Anand Kumar M , Soman KP

The intensification of affective polarization worldwide has raised new questions about how social media platforms might be further fracturing an already-divided public sphere. As opposed to ideological polarization, affective polarization…

Social and Information Networks · Computer Science 2021-10-13 Martin Saveski , Nabeel Gillani , Ann Yuan , Prashanth Vijayaraghavan , Deb Roy

Social media platforms increasingly employ proactive moderation techniques, such as detecting and curbing toxic and uncivil comments, to prevent the spread of harmful content. Despite these efforts, such approaches are often criticized for…

Human-Computer Interaction · Computer Science 2025-07-30 Xiaotian Su , Naim Zierau , Soomin Kim , April Yi Wang , Thiemo Wambsganss

This paper presents TextComplexityDE, a dataset consisting of 1000 sentences in German language taken from 23 Wikipedia articles in 3 different article-genres to be used for developing text-complexity predictor models and automatic text…

Computation and Language · Computer Science 2019-04-17 Babak Naderi , Salar Mohtaj , Kaspar Ensikat , Sebastian Möller

This paper explores the design of a propaganda detection tool using Large Language Models (LLMs). Acknowledging the inherent biases in AI models, especially in political contexts, we investigate how these biases might be leveraged to…

Human-Computer Interaction · Computer Science 2025-12-01 Liudmila Zavolokina , Kilian Sprenkamp , Zoya Katashinskaya , Daniel Gordon Jones

The proliferation of harmful and offensive content is a problem that many online platforms face today. One of the most common approaches for moderating offensive content online is via the identification and removal after it has been posted,…

Social and Information Networks · Computer Science 2021-12-03 Matthew Katsaros , Kathy Yang , Lauren Fratamico

Interpretability is a topic that has been in the spotlight for the past few years. Most existing interpretability techniques produce interpretations in the form of rules or feature importance. These interpretations, while informative, may…

Computation and Language · Computer Science 2024-10-15 Nikolaos Mylonas , Nikolaos Stylianou , Theodora Tsikrika , Stefanos Vrochidis , Ioannis Kompatsiaris

Internet memes, channels for humor, social commentary, and cultural expression, are increasingly used to spread toxic messages. Studies on the computational analyses of toxic memes have significantly grown over the past five years, and the…

Computation and Language · Computer Science 2026-04-20 Delfina Sol Martinez Pandiani , Erik Tjong Kim Sang , Davide Ceolin

Twitter, a popular social media outlet, has evolved into a vast source of linguistic data, rich with opinion, sentiment, and discussion. Due to the increasing popularity of Twitter, its perceived potential for exerting social influence has…

The moderation of content on online platforms is usually non-transparent. On Wikipedia, however, this discussion is carried out publicly and the editors are encouraged to use the content moderation policies as explanations for making…

Machine Learning · Computer Science 2024-11-01 Lucie-Aimée Kaffee , Arnav Arora , Isabelle Augenstein