English
Related papers

Related papers: TAR on Social Media: A Framework for Online Conten…

200 papers

The sheer scale of data required to train modern large language models (LLMs) poses significant risks, as models are likely to gain knowledge of sensitive topics such as bio-security, as well the ability to replicate copyrighted works.…

Computation and Language · Computer Science 2024-12-17 Harry J. Davies , Giorgos Iacovides , Danilo P. Mandic

Anonymity in social media platforms keeps users hidden behind a keyboard. This absolves users of responsibility, allowing them to engage in online rage, hate speech, and other text-based toxicity that harms online well-being. Recent…

Human-Computer Interaction · Computer Science 2023-03-03 Akriti Verma , Shama Islam , Valeh Moghaddam , Adnan Anwar

Centralized content moderation paradigm both falls short and over-reaches: 1) it fails to account for the subjective nature of harm, and 2) it acts with blunt suppression in response to content deemed harmful, even when such content can be…

Human-Computer Interaction · Computer Science 2026-02-11 Rayhan Rashed , Farnaz Jahanbakhsh

Recently, social media platforms are heavily moderated to prevent the spread of online hate speech, which is usually fertile in toxic words and is directed toward an individual or a community. Owing to such heavy moderation, newer and more…

Social and Information Networks · Computer Science 2023-03-21 Punyajoy Saha , Kiran Garimella , Narla Komal Kalyan , Saurabh Kumar Pandey , Pauras Mangesh Meher , Binny Mathew , Animesh Mukherjee

Hate speech remains a pressing challenge on social media, where platform moderation often fails to protect targeted users. Personal moderation tools that let users decide how content is filtered can address some of these shortcomings.…

Human-Computer Interaction · Computer Science 2026-03-03 Anna Ricarda Luther , Hendrik Heuer , Stephanie Geise , Sebastian Haunss , Andreas Breiter

Content moderation remains a critical yet challenging task for large-scale user-generated video platforms, especially in livestreaming environments where moderation must be timely, multimodal, and robust to evolving forms of unwanted…

Computer Vision and Pattern Recognition · Computer Science 2026-01-29 Wei Chee Yew , Hailun Xu , Sanjay Saha , Xiaotian Fan , Hiok Hian Ong , David Yuchen Wang , Kanchan Sarkar , Zhenheng Yang , Danhui Guan

Much of our modern digital infrastructure relies critically upon open sourced software. The communities responsible for building this cyberinfrastructure require maintenance and moderation, which is often supported by volunteer efforts.…

Human-Computer Interaction · Computer Science 2023-08-16 Jane Hsieh , Joselyn Kim , Laura Dabbish , Haiyi Zhu

The detection of sensitive content in large datasets is crucial for ensuring that shared and analysed data is free from harmful material. However, current moderation tools, such as external APIs, suffer from limitations in customisation,…

Computation and Language · Computer Science 2025-06-25 Dimosthenis Antypas , Indira Sen , Carla Perez-Almendros , Jose Camacho-Collados , Francesco Barbieri

Content moderation is a global challenge, yet major tech platforms prioritize high-resource languages, leaving low-resource languages with scarce native moderators. Since effective moderation depends on understanding contextual cues, this…

Computation and Language · Computer Science 2025-11-11 Junyeong Park , Seogyeong Jeong , Seyoung Song , Yohan Lee , Alice Oh

Social media platforms have traditionally relied on internal moderation teams and partnerships with independent fact-checking organizations to identify and flag misleading content. Recently, however, platforms including X (formerly Twitter)…

Detecting harmful content on social media, such as Twitter, is made difficult by the fact that the seemingly simple yes/no classification conceals a significant amount of complexity. Unfortunately, while several datasets have been collected…

Computation and Language · Computer Science 2023-11-14 Saad Almohaimeed , Saleh Almohaimeed , Ashfaq Ali Shafin , Bogdan Carbunar , Ladislau Bölöni

In the pursuit of bolstering user safety, social media platforms deploy active moderation strategies, including content removal and user suspension. These measures target users engaged in discussions marked by hate speech or toxicity, often…

Social and Information Networks · Computer Science 2024-01-26 Hina Qayyum , Muhammad Ikram , Benjamin Zi Hao Zhao , Ian D. Wood , Nicolas Kourtellis , Mohamed Ali Kaafar

With the widespread use of toxic language online, platforms are increasingly using automated systems that leverage advances in natural language processing to automatically flag and remove toxic comments. However, most automated systems --…

Human-Computer Interaction · Computer Science 2021-02-11 Austin P Wright , Omar Shaikh , Haekyu Park , Will Epperson , Muhammed Ahmed , Stephane Pinel , Duen Horng Chau , Diyi Yang

Online harassment is a significant social problem. Prevention of online harassment requires rapid detection of harassing, offensive, and negative social media posts. In this paper, we propose the use of word embedding models to identify…

Machine Learning · Computer Science 2019-11-19 Anqi Liu , Maya Srikanth , Nicholas Adams-Cohen , R. Michael Alvarez , Anima Anandkumar

Among the topics discussed in Social Media, some lead to controversy. A number of recent studies have focused on the problem of identifying controversy in social media mostly based on the analysis of textual content or rely on global…

Social and Information Networks · Computer Science 2017-03-16 Mauro Coletto , Kiran Garimella , Aristides Gionis , Claudio Lucchese

Mental health forums are online communities where people express their issues and seek help from moderators and other users. In such forums, there are often posts with severe content indicating that the user is in acute distress and there…

Computation and Language · Computer Science 2017-02-23 Arman Cohan , Sydney Young , Andrew Yates , Nazli Goharian

Online social platforms centered around content creators often allow comments on content, where creators moderate the comments they receive. As creators can face overwhelming numbers of comments, with some of them harassing or hateful,…

Human-Computer Interaction · Computer Science 2022-02-18 Shagun Jhaver , Quan Ze Chen , Detlef Knauss , Amy Zhang

Hateful comments are prevalent on social media platforms. Although tools for automatically detecting, flagging, and blocking such false, offensive, and harmful content online have lately matured, such reactive and brute force methods alone…

Computation and Language · Computer Science 2024-01-17 Sougata Saha , Rohini Srihari

We propose an agent-based framework for personalized filtering of categorized harassing communication in online social networks. Unlike global moderation systems that apply uniform filtering rules, our approach models user-specific…

Artificial Intelligence · Computer Science 2026-03-17 Zenefa Rahaman , Sandip Sen

Online hate is an escalating problem that negatively impacts the lives of Internet users, and is also subject to rapid changes due to evolving events, resulting in new waves of online hate that pose a critical threat. Detecting and…

Computation and Language · Computer Science 2024-07-03 Nishant Vishwamitra , Keyan Guo , Farhan Tajwar Romit , Isabelle Ondracek , Long Cheng , Ziming Zhao , Hongxin Hu
‹ Prev 1 8 9 10 Next ›