中文
相关论文

相关论文: Leveraging Argument Structure to Predict Content H…

200 篇论文

Hateful speech detection is a key component of content moderation, yet current evaluation frameworks rarely assess why a text is deemed hateful. We introduce \textsf{HateXScore}, a four-component metric suite designed to evaluate the…

计算与语言 · 计算机科学 2026-01-21 Yujia Hu , Roy Ka-Wei Lee

Social media platforms provide users the freedom of expression and a medium to exchange information and express diverse opinions. Unfortunately, this has also resulted in the growth of abusive content with the purpose of discriminating…

计算与语言 · 计算机科学 2021-07-01 Sohail Akhtar , Valerio Basile , Viviana Patti

Tackling online hatred using informed textual responses - called counter narratives - has been brought under the spotlight recently. Accordingly, a research line has emerged to automatically generate counter narratives in order to…

计算与语言 · 计算机科学 2021-06-23 Yi-Ling Chung , Serra Sinem Tekiroglu , Marco Guerini

In this paper, we discuss the development of an annotation schema to build datasets for evaluating the offline harm potential of social media texts. We define "harm potential" as the potential for an online public post to cause real-world…

计算与语言 · 计算机科学 2024-03-19 Ritesh Kumar , Ojaswee Bhalla , Madhu Vanthi , Shehlat Maknoon Wani , Siddharth Singh

Hate speech online targets individuals or groups based on identity attributes and spreads rapidly, posing serious social risks. Memes, which combine images and text, have emerged as a nuanced vehicle for disseminating hate speech, often…

多智能体系统 · 计算机科学 2026-03-26 Rui Xing , Qi Chai , Jie Ma , Jing Tao , Pinghui Wang , Shuming Zhang , Xinping Wang , Hao Wang

In this paper we propose a framework characterizing information disorders as disputes of narratives. Such narratives present claims to readers, who must decide whether to accept the statements in the claims as facts. We point out that this…

计算机与社会 · 计算机科学 2021-08-31 Daniel Schwabe

Text data can pose a risk of harm. However, the risks are not fully understood, and how to handle, present, and discuss harmful text in a safe way remains an unresolved issue in the NLP community. We provide an analytical framework…

计算与语言 · 计算机科学 2023-02-28 Hannah Rose Kirk , Abeba Birhane , Bertie Vidgen , Leon Derczynski

In recent years, we have witnessed the proliferation of large amounts of online content generated directly by users with virtually no form of external control, leading to the possible spread of misinformation. The search for effective…

信息检索 · 计算机科学 2024-07-12 Rishabh Upadhyay , Gabriella Pasi , Marco Viviani

With the recent surge and exponential growth of social media usage, scrutinizing social media content for the presence of any hateful content is of utmost importance. Researchers have been diligently working since the past decade on…

计算与语言 · 计算机科学 2024-01-22 Atanu Mandal , Gargi Roy , Amit Barman , Indranil Dutta , Sudip Kumar Naskar

News will be biased so long as people have opinions. As social media becomes the primary entry point for news and partisan differences increase, it is increasingly important for informed citizens to be able to recognize bias. If people are…

计算与语言 · 计算机科学 2025-05-22 Jessica Zhu , Iain Cruickshank , Michel Cukier

Abusive speech on social media poses a persistent and evolving challenge, driven by the continuous emergence of novel slang and obfuscated terms designed to circumvent detection systems. In this work, we present a data efficient strategy…

计算与语言 · 计算机科学 2025-12-03 Pritish N. Desai , Tanay Kewalramani , Srimanta Mandal

Humans are black boxes -- we cannot observe their neural processes, yet society functions by evaluating verifiable arguments. AI explainability should follow this principle: stakeholders need verifiable reasoning chains, not mechanistic…

机器学习 · 计算机科学 2025-10-07 Ege Cakar , Per Ola Kristensson

The explosive growth of social media has not only revolutionized communication but also brought challenges such as political polarization, misinformation, hate speech, and echo chambers. This dissertation employs computational social…

社会与信息网络 · 计算机科学 2025-09-16 Julie Jiang

Subtle and indirect hate speech remains an underexplored challenge in online safety research, particularly when harmful intent is embedded within misleading or manipulative narratives. Existing hate speech datasets primarily capture overt…

计算与语言 · 计算机科学 2026-03-04 Sai Kartheek Reddy Kasu , Shankar Biradar , Sunil Saumya , Md. Shad Akhtar

Preventing the spread of misinformation is challenging. The detection of misleading content presents a significant hurdle due to its extreme linguistic and domain variability. Content-based models have managed to identify deceptive language…

计算与语言 · 计算机科学 2024-01-30 Flavio Merenda , José Manuel Gómez-Pérez

Platforms have struggled to keep pace with the spread of disinformation. Current responses like user reports, manual analysis, and third-party fact checking are slow and difficult to scale, and as a result, disinformation can spread…

计算机与社会 · 计算机科学 2020-09-30 Austin Hounsel , Jordan Holland , Ben Kaiser , Kevin Borgolte , Nick Feamster , Jonathan Mayer

Hateful and Toxic content has become a significant concern in today's world due to an exponential rise in social media. The increase in hate speech and harmful content motivated researchers to dedicate substantial efforts to the challenging…

计算与语言 · 计算机科学 2021-01-25 Suman Dowlagar , Radhika Mamidi

Argumentation provides a representation of arguments and attacks between these arguments. Argumentation can be used to represent a reasoning process over evidence to reach conclusions. Within such a reasoning process, understanding the…

人工智能 · 计算机科学 2021-02-17 Todd Robinson

In this paper, we delve into the rapidly evolving challenge of misinformation detection, with a specific focus on the nuanced manipulation of narrative frames - an under-explored area within the AI community. The potential for Generative AI…

计算与语言 · 计算机科学 2024-02-27 Guan Wang , Rebecca Frederick , Jinglong Duan , William Wong , Verica Rupar , Weihua Li , Quan Bai

An increasingly common expression of online hate speech is multimodal in nature and comes in the form of memes. Designing systems to automatically detect hateful content is of paramount importance if we are to mitigate its undesirable…