中文
相关论文

相关论文: HateCheck: Functional Tests for Hate Speech Detect…

200 篇论文

Hate speech is a specific type of controversial content that is widely legislated as a crime that must be identified and blocked. However, due to the sheer volume and velocity of the Twitter data stream, hate speech detection cannot be…

计算与语言 · 计算机科学 2021-08-09 Moin Khan , Khurram Shahzad , Kamran Malik

Recent research has highlighted a key issue in speech deepfake detection: models trained on one set of deepfakes perform poorly on others. The question arises: is this due to the continuously improving quality of Text-to-Speech (TTS)…

声音 · 计算机科学 2024-06-13 Nicolas M. Müller , Nicholas Evans , Hemlata Tak , Philip Sperl , Konstantin Böttinger

Social media platforms are critical spaces for public discourse, shaping opinions and community dynamics, yet their widespread use has amplified harmful content, particularly hate speech, threatening online safety and inclusivity. While…

计算与语言 · 计算机科学 2025-06-11 Muhammad Usman , Muhammad Ahmad , M. Shahiki Tash , Irina Gelbukh , Rolando Quintero Tellez , Grigori Sidorov

Automatic identification of hateful and abusive content is vital in combating the spread of harmful online content and its damaging effects. Most existing works evaluate models by examining the generalization error on train-test splits on…

计算与语言 · 计算机科学 2025-04-07 Lanqin Yuan , Marian-Andrei Rizoiu

The growth of social networks makes toxic content spread rapidly. Hate speech detection is a task to help decrease the number of harmful comments. With the diversity in the hate speech created by users, it is necessary to interpret the hate…

计算与语言 · 计算机科学 2025-02-11 Cuong Nhat Vo , Khanh Bao Huynh , Son T. Luu , Trong-Hop Do

The proliferation of online hate speech poses a significant threat to the harmony of the web. While explicit hate is easily recognized through overt slurs, implicit hate speech is often conveyed through sarcasm, irony, stereotypes, or coded…

计算与语言 · 计算机科学 2026-02-04 Chengshuai Zhao , Shu Wan , Paras Sheth , Karan Patwa , K. Selçuk Candan , Huan Liu

Society needs to develop a system to detect hate and offense to build a healthy and safe environment. However, current research in this field still faces four major shortcomings, including deficient pre-processing techniques, indifference…

计算与语言 · 计算机科学 2022-06-02 Khanh Q. Tran , An T. Nguyen , Phu Gia Hoang , Canh Duc Luu , Trong-Hop Do , Kiet Van Nguyen

The curation of hate speech datasets involves complex design decisions that balance competing priorities. This paper critically examines these methodological choices in a diverse range of datasets, highlighting common themes and practices,…

计算与语言 · 计算机科学 2025-06-23 Luna Wang , Andrew Caines , Alice Hutchings

The ubiquity of social media has transformed online interactions among individuals. Despite positive effects, it has also allowed anti-social elements to unite in alternative social media environments (eg. Gab.com) like never before.…

社会与信息网络 · 计算机科学 2020-07-28 Michael Ridenhour , Arunkumar Bagavathi , Elaheh Raisi , Siddharth Krishnan

Hate speech detection is key to online content moderation, but current models struggle to generalise beyond their training data. This has been linked to dataset biases and the use of sentence-level labels, which fail to teach models the…

计算与语言 · 计算机科学 2025-06-05 Agostina Calabrese , Tom Sherborne , Björn Ross , Mirella Lapata

With proliferation of user generated contents in social media platforms, establishing mechanisms to automatically identify toxic and abusive content becomes a prime concern for regulators, researchers, and society. Keeping the balance…

计算与语言 · 计算机科学 2021-06-10 Djamila Romaissa Beddiar , Md Saroar Jahan , Mourad Oussalah

Social media platforms, while enabling global connectivity, have become hubs for the rapid spread of harmful content, including hate speech and fake narratives \cite{davidson2017automated, shu2017fake}. The Faux-Hate shared task focuses on…

计算与语言 · 计算机科学 2025-12-19 Yash Bhaskar , Sankalp Bahad , Parameswari Krishnamurthy

Recent studies have proposed models that yielded promising performance for the hateful meme classification task. Nevertheless, these proposed models do not generate interpretable explanations that uncover the underlying meaning and support…

计算与语言 · 计算机科学 2023-06-21 Ming Shan Hee , Wen-Haw Chong , Roy Ka-Wei Lee

Solutions for defending against deepfake speech fall into two categories: proactive watermarking models and passive conventional deepfake detectors. While both address common threats, their differences in training, optimization, and…

声音 · 计算机科学 2025-06-18 Chia-Hua Wu , Wanying Ge , Xin Wang , Junichi Yamagishi , Yu Tsao , Hsin-Min Wang

Hate speech is plaguing the cyberspace along with user-generated content. This paper investigates the role of conversational context in the annotation and detection of online hate and counter speech, where context is defined as the…

计算与语言 · 计算机科学 2022-06-15 Xinchen Yu , Eduardo Blanco , Lingzi Hong

The recognition of hate speech and offensive language (HOF) is commonly formulated as a classification task to decide if a text contains HOF. We investigate whether HOF detection can profit by taking into account the relationships between…

计算与语言 · 计算机科学 2022-07-12 Flor Miriam Plaza-del-Arco , Sercan Halat , Sebastian Padó , Roman Klinger

Standard approaches to hate speech detection rely on sufficient available hate speech annotations. Extending previous work that repurposes natural language inference (NLI) models for zero-shot text classification, we propose a simple…

计算与语言 · 计算机科学 2022-10-04 Janis Goldzycher , Gerold Schneider

Counterfactually Augmented Data (CAD) aims to improve out-of-domain generalizability, an indicator of model robustness. The improvement is credited with promoting core features of the construct over spurious artifacts that happen to…

计算与语言 · 计算机科学 2022-05-10 Indira Sen , Mattia Samory , Claudia Wagner , Isabelle Augenstein

The proliferation of online hate speech has necessitated the creation of algorithms which can detect toxicity. Most of the past research focuses on this detection as a classification task, but assigning an absolute toxicity label is often…

计算与语言 · 计算机科学 2022-06-28 Millon Madhur Das , Punyajoy Saha , Mithun Das

Our work advances an approach for predicting hate speech in social media, drawing out the critical need to consider the discussions that follow a post to successfully detect when hateful discourse may arise. Using graph transformer…

机器学习 · 计算机科学 2023-05-02 Liam Hebert , Hong Yi Chen , Robin Cohen , Lukasz Golab
‹ 上一页 1 8 9 10 下一页 ›