中文
相关论文

相关论文: Into the LAIONs Den: Investigating Hate in Multimo…

200 篇论文

The proliferation of hate speech and offensive comments on social media has become increasingly prevalent due to user activities. Such comments can have detrimental effects on individuals' psychological well-being and social behavior. While…

Online hate is a growing concern on many social media platforms and other sites. To combat it, technology companies are increasingly identifying and sanctioning `hateful users' rather than simply moderating hateful content. Yet, most…

社会与信息网络 · 计算机科学 2021-03-23 Zo Ahmed , Bertie Vidgen , Scott A. Hale

Hateful speech in Online Social Networks (OSNs) is a key challenge for companies and governments, as it impacts users and advertisers, and as several countries have strict legislation against the practice. This has motivated work on…

社会与信息网络 · 计算机科学 2018-01-16 Manoel Horta Ribeiro , Pedro H. Calais , Yuri A. Santos , Virgílio A. F. Almeida , Wagner Meira

Detecting harmful content on social media, such as Twitter, is made difficult by the fact that the seemingly simple yes/no classification conceals a significant amount of complexity. Unfortunately, while several datasets have been collected…

计算与语言 · 计算机科学 2023-11-14 Saad Almohaimeed , Saleh Almohaimeed , Ashfaq Ali Shafin , Bogdan Carbunar , Ladislau Bölöni

We suggest a multilabel Korean online hate speech dataset that covers seven categories of hate speech: (1) Race and Nationality, (2) Religion, (3) Regionalism, (4) Ageism, (5) Misogyny, (6) Sexual Minorities, and (7) Male. Our 35K dataset…

计算与语言 · 计算机科学 2022-04-11 TaeYoung Kang , Eunrang Kwon , Junbum Lee , Youngeun Nam , Junmo Song , JeongKyu Suh

The purpose of this paper is to ascertain the influence of sociocultural factors (i.e., social, cultural, and political) in the development of hate speech detection systems. We set out to investigate the suitability of using open-source…

计算与语言 · 计算机科学 2024-07-02 Sidney G. -J. Wong

Perceptions of hate can vary greatly across cultural contexts. Hate speech (HS) datasets, however, have traditionally been developed by language. This hides potential cultural biases, as one language may be spoken in different countries…

计算与语言 · 计算机科学 2025-05-20 Manuel Tonneau , Diyi Liu , Samuel Fraiberger , Ralph Schroeder , Scott A. Hale , Paul Röttger

Online platforms struggle to curb hate speech without over-censoring legitimate discourse. Early bidirectional transformer encoders made big strides, but the arrival of ultra-large autoregressive LLMs promises deeper context-awareness.…

计算与语言 · 计算机科学 2025-07-15 Ariadna Mon , Saúl Fenollosa , Jon Lecumberri

The prevalence of digital media and evolving sociopolitical dynamics have significantly amplified the dissemination of hateful content. Existing studies mainly focus on classifying texts into binary categories, often overlooking the…

计算与语言 · 计算机科学 2024-04-19 Abinew Ali Ayele , Esubalew Alemneh Jalew , Adem Chanie Ali , Seid Muhie Yimam , Chris Biemann

The surge of interest in data augmentation within the realm of NLP has been driven by the need to address challenges posed by hate speech domains, the dynamic nature of social media vocabulary, and the demands for large-scale neural…

计算与语言 · 计算机科学 2024-04-02 Md Saroar Jahan , Mourad Oussalah , Djamila Romaissa Beddia , Jhuma kabir Mim , Nabil Arhab

Large Language Models (LLMs) inherit explicit and implicit biases from their training datasets. Identifying and mitigating biases in LLMs is crucial to ensure fair outputs, as they can perpetuate harmful stereotypes and misinformation. This…

机器学习 · 计算机科学 2025-11-19 Fatima Kazi , Alex Young , Yash Inani , Setareh Rafatirad

Hate speech detection refers to the task of detecting hateful content that aims at denigrating an individual or a group based on their religion, gender, sexual orientation, or other characteristics. Due to the different policies of the…

计算与语言 · 计算机科学 2023-10-10 Paras Sheth , Tharindu Kumarage , Raha Moraffah , Aman Chadha , Huan Liu

Large datasets underlying much of current machine learning raise serious issues concerning inappropriate content such as offensive, insulting, threatening, or might otherwise cause anxiety. This calls for increased dataset documentation,…

人工智能 · 计算机科学 2022-07-15 Patrick Schramowski , Christopher Tauchmann , Kristian Kersting

Knowledge of slang is a desirable feature of LLMs in the context of user interaction, as slang often reflects an individual's social identity. Several works on informal language processing have defined and curated benchmarks for tasks such…

计算与语言 · 计算机科学 2025-09-23 Leonor Veloso , Lea Hirlimann , Philipp Wicke , Hinrich Schütze

Hate speech remains prevalent in human society and continues to evolve in its forms and expressions. Modern advancements in internet and online anonymity accelerate its rapid spread and complicate its detection. However, hate speech…

计算与语言 · 计算机科学 2026-04-24 Yejin Lee , Hyeseon Ahn , Yo-Sub Han

Warning: this paper contains content that may be offensive or upsetting. Most hate speech datasets neglect the cultural diversity within a single language, resulting in a critical shortcoming in hate speech detection. To address this, we…

计算与语言 · 计算机科学 2024-04-04 Nayeon Lee , Chani Jung , Junho Myung , Jiho Jin , Jose Camacho-Collados , Juho Kim , Alice Oh

Large Language models (LLMs), such as ChatGPT, have gained popularity in recent years with the advancement of Natural Language Processing (NLP), with use cases spanning many disciplines and daily lives as well. LLMs inherit explicit and…

计算与语言 · 计算机科学 2025-12-01 Fatima Kazi

Although social media platforms are a prominent arena for users to engage in interpersonal discussions and express opinions, the facade and anonymity offered by social media may allow users to spew hate speech and offensive content. Given…

计算与语言 · 计算机科学 2024-05-09 Ayushi Nirmal , Amrita Bhattacharjee , Paras Sheth , Huan Liu

Online hate speech is a recent problem in our society that is rising at a steady pace by leveraging the vulnerabilities of the corresponding regimes that characterise most social media platforms. This phenomenon is primarily fostered by…

计算与语言 · 计算机科学 2022-01-05 Ioannis Mollas , Zoe Chrysopoulou , Stamatis Karlos , Grigorios Tsoumakas

Recent studies have proposed models that yielded promising performance for the hateful meme classification task. Nevertheless, these proposed models do not generate interpretable explanations that uncover the underlying meaning and support…

计算与语言 · 计算机科学 2023-06-21 Ming Shan Hee , Wen-Haw Chong , Roy Ka-Wei Lee