中文
相关论文

相关论文: TAR on Social Media: A Framework for Online Conten…

200 篇论文

The sheer scale of data required to train modern large language models (LLMs) poses significant risks, as models are likely to gain knowledge of sensitive topics such as bio-security, as well the ability to replicate copyrighted works.…

计算与语言 · 计算机科学 2024-12-17 Harry J. Davies , Giorgos Iacovides , Danilo P. Mandic

Anonymity in social media platforms keeps users hidden behind a keyboard. This absolves users of responsibility, allowing them to engage in online rage, hate speech, and other text-based toxicity that harms online well-being. Recent…

人机交互 · 计算机科学 2023-03-03 Akriti Verma , Shama Islam , Valeh Moghaddam , Adnan Anwar

Centralized content moderation paradigm both falls short and over-reaches: 1) it fails to account for the subjective nature of harm, and 2) it acts with blunt suppression in response to content deemed harmful, even when such content can be…

人机交互 · 计算机科学 2026-02-11 Rayhan Rashed , Farnaz Jahanbakhsh

Recently, social media platforms are heavily moderated to prevent the spread of online hate speech, which is usually fertile in toxic words and is directed toward an individual or a community. Owing to such heavy moderation, newer and more…

Hate speech remains a pressing challenge on social media, where platform moderation often fails to protect targeted users. Personal moderation tools that let users decide how content is filtered can address some of these shortcomings.…

人机交互 · 计算机科学 2026-03-03 Anna Ricarda Luther , Hendrik Heuer , Stephanie Geise , Sebastian Haunss , Andreas Breiter

Content moderation remains a critical yet challenging task for large-scale user-generated video platforms, especially in livestreaming environments where moderation must be timely, multimodal, and robust to evolving forms of unwanted…

计算机视觉与模式识别 · 计算机科学 2026-01-29 Wei Chee Yew , Hailun Xu , Sanjay Saha , Xiaotian Fan , Hiok Hian Ong , David Yuchen Wang , Kanchan Sarkar , Zhenheng Yang , Danhui Guan

Much of our modern digital infrastructure relies critically upon open sourced software. The communities responsible for building this cyberinfrastructure require maintenance and moderation, which is often supported by volunteer efforts.…

人机交互 · 计算机科学 2023-08-16 Jane Hsieh , Joselyn Kim , Laura Dabbish , Haiyi Zhu

The detection of sensitive content in large datasets is crucial for ensuring that shared and analysed data is free from harmful material. However, current moderation tools, such as external APIs, suffer from limitations in customisation,…

计算与语言 · 计算机科学 2025-06-25 Dimosthenis Antypas , Indira Sen , Carla Perez-Almendros , Jose Camacho-Collados , Francesco Barbieri

Content moderation is a global challenge, yet major tech platforms prioritize high-resource languages, leaving low-resource languages with scarce native moderators. Since effective moderation depends on understanding contextual cues, this…

计算与语言 · 计算机科学 2025-11-11 Junyeong Park , Seogyeong Jeong , Seyoung Song , Yohan Lee , Alice Oh

Social media platforms have traditionally relied on internal moderation teams and partnerships with independent fact-checking organizations to identify and flag misleading content. Recently, however, platforms including X (formerly Twitter)…

Detecting harmful content on social media, such as Twitter, is made difficult by the fact that the seemingly simple yes/no classification conceals a significant amount of complexity. Unfortunately, while several datasets have been collected…

计算与语言 · 计算机科学 2023-11-14 Saad Almohaimeed , Saleh Almohaimeed , Ashfaq Ali Shafin , Bogdan Carbunar , Ladislau Bölöni

In the pursuit of bolstering user safety, social media platforms deploy active moderation strategies, including content removal and user suspension. These measures target users engaged in discussions marked by hate speech or toxicity, often…

社会与信息网络 · 计算机科学 2024-01-26 Hina Qayyum , Muhammad Ikram , Benjamin Zi Hao Zhao , Ian D. Wood , Nicolas Kourtellis , Mohamed Ali Kaafar

With the widespread use of toxic language online, platforms are increasingly using automated systems that leverage advances in natural language processing to automatically flag and remove toxic comments. However, most automated systems --…

Online harassment is a significant social problem. Prevention of online harassment requires rapid detection of harassing, offensive, and negative social media posts. In this paper, we propose the use of word embedding models to identify…

机器学习 · 计算机科学 2019-11-19 Anqi Liu , Maya Srikanth , Nicholas Adams-Cohen , R. Michael Alvarez , Anima Anandkumar

Among the topics discussed in Social Media, some lead to controversy. A number of recent studies have focused on the problem of identifying controversy in social media mostly based on the analysis of textual content or rely on global…

社会与信息网络 · 计算机科学 2017-03-16 Mauro Coletto , Kiran Garimella , Aristides Gionis , Claudio Lucchese

Mental health forums are online communities where people express their issues and seek help from moderators and other users. In such forums, there are often posts with severe content indicating that the user is in acute distress and there…

计算与语言 · 计算机科学 2017-02-23 Arman Cohan , Sydney Young , Andrew Yates , Nazli Goharian

Online social platforms centered around content creators often allow comments on content, where creators moderate the comments they receive. As creators can face overwhelming numbers of comments, with some of them harassing or hateful,…

人机交互 · 计算机科学 2022-02-18 Shagun Jhaver , Quan Ze Chen , Detlef Knauss , Amy Zhang

Hateful comments are prevalent on social media platforms. Although tools for automatically detecting, flagging, and blocking such false, offensive, and harmful content online have lately matured, such reactive and brute force methods alone…

计算与语言 · 计算机科学 2024-01-17 Sougata Saha , Rohini Srihari

We propose an agent-based framework for personalized filtering of categorized harassing communication in online social networks. Unlike global moderation systems that apply uniform filtering rules, our approach models user-specific…

人工智能 · 计算机科学 2026-03-17 Zenefa Rahaman , Sandip Sen

Online hate is an escalating problem that negatively impacts the lives of Internet users, and is also subject to rapid changes due to evolving events, resulting in new waves of online hate that pose a critical threat. Detecting and…

计算与语言 · 计算机科学 2024-07-03 Nishant Vishwamitra , Keyan Guo , Farhan Tajwar Romit , Isabelle Ondracek , Long Cheng , Ziming Zhao , Hongxin Hu
‹ 上一页 1 8 9 10 下一页 ›