中文
相关论文

相关论文: ProvocationProbe: Instigating Hate Speech Dataset …

200 篇论文

We present a neural-network based approach to classifying online hate speech in general, as well as racist and sexist speech in particular. Using pre-trained word embeddings and max/mean pooling from simple, fully-connected transformations…

计算与语言 · 计算机科学 2018-09-28 Rohan Kshirsagar , Tyus Cukuvac , Kathleen McKeown , Susan McGregor

Algorithmic hate speech detection faces significant challenges due to the diverse definitions and datasets used in research and practice. Social media platforms, legal frameworks, and institutions each apply distinct yet overlapping…

计算与语言 · 计算机科学 2025-03-10 Jan Fillies , Adrian Paschke

The expanding influence of social media platforms over the past decade has impacted the way people communicate. The level of obscurity provided by social media and easy accessibility of the internet has facilitated the spread of hate…

计算与语言 · 计算机科学 2024-12-02 Susmita Das , Arpita Dutta , Kingshuk Roy , Abir Mondal , Arnab Mukhopadhyay

We introduce a novel multi-labeled scheme for joint annotation of hate and counter-hate speech in social media conversations, categorizing hate and counter-hate messages into thematic and rhetorical dimensions. The thematic categories…

计算与语言 · 计算机科学 2025-07-29 Effi Levi , Gal Ron , Odelia Oshri , Shaul R. Shenhav

The ability to accurately detect and filter offensive content automatically is important to ensure a rich and diverse digital discourse. Trolling is a type of hurtful or offensive content that is prevalent in social media, but is…

计算机与社会 · 计算机科学 2020-08-04 Hitkul , Karmanya Aggarwal , Pakhi Bamdev , Debanjan Mahata , Rajiv Ratn Shah , Ponnurangam Kumaraguru

Hate speech has grown into a pervasive phenomenon, intensifying during times of crisis, elections, and social unrest. Multiple approaches have been developed to detect hate speech using artificial intelligence, but a generalized model is…

计算与语言 · 计算机科学 2024-10-10 Gautam Kishore Shahi , Tim A. Majchrzak

To tackle the rising phenomenon of hate speech, efforts have been made towards data curation and analysis. When it comes to analysis of bias, previous work has focused predominantly on race. In our work, we further investigate bias in hate…

计算与语言 · 计算机科学 2022-05-19 Antonis Maronikolakis , Philip Baader , Hinrich Schütze

We address a challenging problem of identifying main sources of hate speech on Twitter. On one hand, we carefully annotate a large set of tweets for hate speech, and deploy advanced deep learning to produce high quality hate speech…

社会与信息网络 · 计算机科学 2022-03-21 Bojan Evkoski , Andraz Pelicon , Igor Mozetic , Nikola Ljubesic , Petra Kralj Novak

The context-dependent nature of online aggression makes annotating large collections of data extremely difficult. Previously studied datasets in abusive language detection have been insufficient in size to efficiently train deep learning…

计算与语言 · 计算机科学 2018-08-31 Younghun Lee , Seunghyun Yoon , Kyomin Jung

Our study addresses a significant gap in online hate speech detection research by focusing on homophobia, an area often neglected in sentiment analysis research. Utilising advanced sentiment analysis models, particularly BERT, and…

计算与语言 · 计算机科学 2024-05-16 Josh McGiff , Nikola S. Nikolov

Hateful speech in Online Social Networks (OSNs) is a key challenge for companies and governments, as it impacts users and advertisers, and as several countries have strict legislation against the practice. This has motivated work on…

社会与信息网络 · 计算机科学 2018-01-16 Manoel Horta Ribeiro , Pedro H. Calais , Yuri A. Santos , Virgílio A. F. Almeida , Wagner Meira

We introduce a generic, language-independent method to collect a large percentage of offensive and hate tweets regardless of their topics or genres. We harness the extralinguistic information embedded in the emojis to collect a large number…

计算与语言 · 计算机科学 2022-05-20 Hamdy Mubarak , Sabit Hassan , Shammur Absar Chowdhury

Perceptions of hate can vary greatly across cultural contexts. Hate speech (HS) datasets, however, have traditionally been developed by language. This hides potential cultural biases, as one language may be spoken in different countries…

计算与语言 · 计算机科学 2025-05-20 Manuel Tonneau , Diyi Liu , Samuel Fraiberger , Ralph Schroeder , Scott A. Hale , Paul Röttger

Hate speech detection has been extensively studied, yet existing methods often overlook a real-world complexity: training labels are biased, and interpretations of what is considered hate vary across individuals with different cultural…

计算与语言 · 计算机科学 2025-10-17 Weibin Cai , Reza Zafarani

Social media is awash with hateful content, much of which is often veiled with linguistic and topical diversity. The benchmark datasets used for hate speech detection do not account for such divagation as they are predominantly compiled…

计算与语言 · 计算机科学 2023-06-16 Atharva Kulkarni , Sarah Masud , Vikram Goyal , Tanmoy Chakraborty

User-generated content online is shaped by many factors, including endogenous elements such as platform affordances and norms, as well as exogenous elements, in particular significant events. These impact what users say, how they say it,…

社会与信息网络 · 计算机科学 2018-04-23 Alexandra Olteanu , Carlos Castillo , Jeremy Boy , Kush R. Varshney

Hate speech detection has become a hot topic in recent years due to the exponential growth of offensive language in social media. It has proven that, state-of-the-art hate speech classifiers are efficient only when tested on the data with…

计算与语言 · 计算机科学 2021-07-06 Hadi Mansourifar , Dana Alsagheer , Weidong Shi , Lan Ni , Yan Huang

Detecting offensive language on Twitter has many applications ranging from detecting/predicting bullying to measuring polarization. In this paper, we focus on building a large Arabic offensive tweet dataset. We introduce a method for…

计算与语言 · 计算机科学 2021-03-11 Hamdy Mubarak , Ammar Rashed , Kareem Darwish , Younes Samih , Ahmed Abdelali

The spread of COVID-19 has sparked racism and hate on social media targeted towards Asian communities. However, little is known about how racial hate spreads during a pandemic and the role of counterspeech in mitigating this spread. In this…

社会与信息网络 · 计算机科学 2021-11-12 Bing He , Caleb Ziems , Sandeep Soni , Naren Ramakrishnan , Diyi Yang , Srijan Kumar

Despite the growing body of research tackling offensive language in social media, this research is predominantly reactive, determining if content already posted in social media is abusive. There is a gap in predictive approaches, which we…

计算与语言 · 计算机科学 2025-03-07 Raneem Alharthi , Rajwa Alharthi , Ravi Shekhar , Aiqi Jiang , Arkaitz Zubiaga