中文
相关论文

相关论文: HateModerate: Testing Hate Speech Detectors agains…

200 篇论文

Despite the valuable social interactions that online media promote, these systems provide space for speech that would be potentially detrimental to different groups of people. The moderation of content imposed by many social media has…

社会与信息网络 · 计算机科学 2021-08-30 Lucas Henrique Costa de Lima , Julio Reis , Philipe Melo , Fabricio Murai , Fabricio Benevenuto

Exploiting social media to spread hate has tremendously increased over the years. Lately, multi-modal hateful content such as memes has drawn relatively more traction than uni-modal content. Moreover, the availability of implicit content…

计算与语言 · 计算机科学 2023-02-14 Piush Aggarwal , Pranit Chawla , Mithun Das , Punyajoy Saha , Binny Mathew , Torsten Zesch , Animesh Mukherjee

Content moderation is a central mechanism through which platforms attempt to balance user engagement with community governance. Yet existing research has largely treated moderation as a uniform intervention, overlooking how moderator…

计算机与社会 · 计算机科学 2026-05-18 Siyi Zhou , Lindsay Young , Marlon Twyman , Emilio Ferrara

The spread of hatred that was formerly limited to verbal communications has rapidly moved over the Internet. Social media and community forums that allow people to discuss and express their opinions are becoming platforms for the spreading…

计算与语言 · 计算机科学 2022-04-15 D. C Asogwa , C. I Chukwuneke , C. C Ngene , G. N Anigbogu

Hate Speech has become a major content moderation issue for online social media platforms. Given the volume and velocity of online content production, it is impossible to manually moderate hate speech related content on any platform. In…

计算与语言 · 计算机科学 2021-01-28 Sudhanshu Mishra , Shivangi Prasad , Shubhanshu Mishra

Hate speech detection has become a hot topic in recent years due to the exponential growth of offensive language in social media. It has proven that, state-of-the-art hate speech classifiers are efficient only when tested on the data with…

计算与语言 · 计算机科学 2021-07-06 Hadi Mansourifar , Dana Alsagheer , Weidong Shi , Lan Ni , Yan Huang

In the recent past, social media platforms have helped people in connecting and communicating to a wider audience. But this has also led to a drastic increase in cyberbullying. It is essential to detect and curb hate speech to keep the…

计算与语言 · 计算机科学 2021-11-02 Ravindra Nayak , Raviraj Joshi

Large language models (LLMs) excel in many diverse applications beyond language generation, e.g., translation, summarization, and sentiment analysis. One intriguing application is in text classification. This becomes pertinent in the realm…

计算与语言 · 计算机科学 2024-03-14 Tharindu Kumarage , Amrita Bhattacharjee , Joshua Garland

Online social media platforms use automated moderation systems to remove or reduce the visibility of rule-breaking content. While previous work has documented the importance of manual content moderation, the effects of automated content…

计算机与社会 · 计算机科学 2023-02-17 Manoel Horta Ribeiro , Justin Cheng , Robert West

Hate speech is a specific type of controversial content that is widely legislated as a crime that must be identified and blocked. However, due to the sheer volume and velocity of the Twitter data stream, hate speech detection cannot be…

计算与语言 · 计算机科学 2021-08-09 Moin Khan , Khurram Shahzad , Kamran Malik

Social media platforms serve as accessible outlets for individuals to express their thoughts and experiences, resulting in an influx of user-generated data spanning all age groups. While these platforms enable free expression, they also…

计算与语言 · 计算机科学 2023-12-12 Nikhil Narayan , Mrutyunjay Biswal , Pramod Goyal , Abhranta Panigrahi

Amidst the rise of Large Multimodal Models (LMMs) and their widespread application in generating and interpreting complex content, the risk of propagating biased and harmful memes remains significant. Current safety measures often fail to…

人工智能 · 计算机科学 2025-05-01 Xuanyu Su , Yansong Li , Diana Inkpen , Nathalie Japkowicz

The rise in harmful online content not only distorts public discourse but also poses significant challenges to maintaining a healthy digital environment. In response to this, we introduce a multimodal dataset uniquely crafted for…

Detecting harmful content on social media, such as Twitter, is made difficult by the fact that the seemingly simple yes/no classification conceals a significant amount of complexity. Unfortunately, while several datasets have been collected…

计算与语言 · 计算机科学 2023-11-14 Saad Almohaimeed , Saleh Almohaimeed , Ashfaq Ali Shafin , Bogdan Carbunar , Ladislau Bölöni

The curation of hate speech datasets involves complex design decisions that balance competing priorities. This paper critically examines these methodological choices in a diverse range of datasets, highlighting common themes and practices,…

计算与语言 · 计算机科学 2025-06-23 Luna Wang , Andrew Caines , Alice Hutchings

The widespread presence of hate speech on the internet, including formats such as text-based tweets and vision-language memes, poses a significant challenge to digital platform safety. Recent research has developed detection models tailored…

计算与语言 · 计算机科学 2024-10-10 Ming Shan Hee , Aditi Kumaresan , Roy Ka-Wei Lee

Most current approaches to characterize and detect hate speech focus on \textit{content} posted in Online Social Networks. They face shortcomings to collect and annotate hateful speech due to the incompleteness and noisiness of OSN text and…

计算机与社会 · 计算机科学 2018-03-28 Manoel Horta Ribeiro , Pedro H. Calais , Yuri A. Santos , Virgílio A. F. Almeida , Wagner Meira

Subtle and indirect hate speech remains an underexplored challenge in online safety research, particularly when harmful intent is embedded within misleading or manipulative narratives. Existing hate speech datasets primarily capture overt…

计算与语言 · 计算机科学 2026-03-04 Sai Kartheek Reddy Kasu , Shankar Biradar , Sunil Saumya , Md. Shad Akhtar

Online hatred is a growing concern on many social media platforms. To address this issue, different social media platforms have introduced moderation policies for such content. They also employ moderators who can check the posts violating…

计算与语言 · 计算机科学 2021-12-01 Mithun Das , Somnath Banerjee , Punyajoy Saha