English
Related papers

Related papers: Impact of Stricter Content Moderation on Parler's …

200 papers

Research across different disciplines has documented the expanding polarization in social media. However, much of it focused on the US political system or its culturally controversial topics. In this work, we explore polarization on Twitter…

Social and Information Networks · Computer Science 2021-04-13 Ramon Villa-Cox , Helen , Zeng , Ashiqur R. KhudaBukhsh , Kathleen M. Carley

Users online tend to consume information adhering to their system of beliefs and to ignore dissenting information. During the COVID-19 pandemic, users get exposed to a massive amount of information about a new topic having a high level of…

Social and Information Networks · Computer Science 2021-06-09 Gabriele Etta , Matteo Cinelli , Alessandro Galeazzi , Carlo Michele Valensise , Mauro Conti , Walter Quattrociocchi

The proliferation of harmful content shared online poses a threat to online information integrity and the integrity of discussion across platforms. Despite various moderation interventions adopted by social media platforms, researchers and…

Social and Information Networks · Computer Science 2023-04-07 Valerio La Gatta , Luca Luceri , Francesco Fabbri , Emilio Ferrara

Current content moderation follows a reactive, trial-and-error approach, where interventions are applied and their effects are only measured post-hoc. In contrast, we introduce a proactive, predictive approach that enables moderators to…

Computers and Society · Computer Science 2026-02-09 Benedetta Tessa , Lorenzo Cima , Amaury Trujillo , Marco Avvenuti , Stefano Cresci

To protect users from massive hateful content, existing works studied automated hate speech detection. Despite the existing efforts, one question remains: do automated hate speech detectors conform to social media content policies? A…

Software Engineering · Computer Science 2024-03-20 Jiangrui Zheng , Xueqing Liu , Guanqun Yang , Mirazul Haque , Xing Qian , Ravishka Rathnasuriya , Wei Yang , Girish Budhrani

Toxic language, such as hate speech, can deter users from participating in online communities and enjoying popular platforms. Previous approaches to detecting toxic language and norm violations have been primarily concerned with…

Computation and Language · Computer Science 2023-10-10 Jihyung Moon , Dong-Ho Lee , Hyundong Cho , Woojeong Jin , Chan Young Park , Minwoo Kim , Jonathan May , Jay Pujara , Sungjoon Park

Content moderation is the process of flagging content based on pre-defined platform rules. There has been a growing need for AI moderators to safeguard users as well as protect the mental health of human moderators from traumatic content.…

Computation and Language · Computer Science 2023-02-21 Meng Ye , Karan Sikka , Katherine Atwell , Sabit Hassan , Ajay Divakaran , Malihe Alikhani

The exponential growth of social media platforms such as Twitter and Facebook has revolutionized textual communication and textual content publication in human society. However, they have been increasingly exploited to propagate toxic…

Computation and Language · Computer Science 2023-02-14 Wenxuan Wang , Jen-tse Huang , Weibin Wu , Jianping Zhang , Yizhan Huang , Shuqing Li , Pinjia He , Michael Lyu

While recent research has focused on developing safeguards for generative AI (GAI) model-level content safety, little is known about how content moderation to prevent malicious content performs for end-users in real-world GAI products. To…

Human-Computer Interaction · Computer Science 2025-06-18 Lan Gao , Oscar Chen , Rachel Lee , Nick Feamster , Chenhao Tan , Marshini Chetty

Despite the valuable social interactions that online media promote, these systems provide space for speech that would be potentially detrimental to different groups of people. The moderation of content imposed by many social media has…

Social and Information Networks · Computer Science 2021-08-30 Lucas Henrique Costa de Lima , Julio Reis , Philipe Melo , Fabricio Murai , Fabricio Benevenuto

The Internet and online forums such as Reddit have become an increasingly popular medium for citizens to engage in political conversations. However, the online disinhibition effect resulting from the ability to use pseudonymous identities…

Computation and Language · Computer Science 2017-07-20 Rishab Nithyanand , Brian Schaffner , Phillipa Gill

Social media platforms provide users the freedom of expression and a medium to exchange information and express diverse opinions. Unfortunately, this has also resulted in the growth of abusive content with the purpose of discriminating…

Computation and Language · Computer Science 2021-07-01 Sohail Akhtar , Valerio Basile , Viviana Patti

Calls are escalating for social media platforms to do more to mitigate extreme online communities whose views can lead to real-world harms, e.g., mis/disinformation and distrust that increased Covid-19 fatalities, and now extend to…

Social and Information Networks · Computer Science 2022-07-22 Elvira Maria Restrepo , Martin Moreno , Lucia Illari , Neil F. Johnson

Formally announced to the public following former President Donald Trump's bans and suspensions from mainstream social networks in early 2022 after his role in the January 6 Capitol Riots, Truth Social was launched as an "alternative"…

Social and Information Networks · Computer Science 2023-04-04 Patrick Gerard , Nicholas Botzer , Tim Weninger

Evaluating the effects of moderation interventions is a task of paramount importance, as it allows assessing the success of content moderation processes. So far, intervention effects have been almost solely evaluated at the aggregated…

Social and Information Networks · Computer Science 2023-03-02 Amaury Trujillo , Stefano Cresci

The detection of sensitive content in large datasets is crucial for ensuring that shared and analysed data is free from harmful material. However, current moderation tools, such as external APIs, suffer from limitations in customisation,…

Computation and Language · Computer Science 2025-06-25 Dimosthenis Antypas , Indira Sen , Carla Perez-Almendros , Jose Camacho-Collados , Francesco Barbieri

The detection of hate speech or toxic content online is a complex and sensitive issue. While the identification itself is highly dependent on the context of the situation, sensitive personal attributes such as age, language, and nationality…

Multiagent Systems · Computer Science 2024-10-11 Jan Fillies , Theodoros Mitsikas , Ralph Schäfermeier , Adrian Paschke

Harmful content detection models tend to have higher false positive rates for content from marginalized groups. In the context of marginal abuse modeling on Twitter, such disproportionate penalization poses the risk of reduced visibility,…

Computation and Language · Computer Science 2022-10-13 Kyra Yee , Alice Schoenauer Sebag , Olivia Redfield , Emily Sheng , Matthias Eck , Luca Belli

WARNING: This paper contains examples of offensive materials. To address the proliferation of toxic content on social media, we introduce SMARTER, we introduce SMARTER, a data-efficient two-stage framework for explainable content moderation…

Computation and Language · Computer Science 2026-04-23 Huy Nghiem , Advik Sachdeva , Hal Daumé

Social Network sites are fertile ground for several polluting phenomena affecting online and offline spaces. Among these phenomena are included echo chambers, closed systems in which the opinions expressed by the people inside are…

Social and Information Networks · Computer Science 2024-01-22 Erica Cau , Virginia Morini , Giulio Rossetti