中文
相关论文

相关论文: DeMod: A Holistic Tool with Explainable Detection …

200 篇论文

Previous work has examined how debiasing language models affect downstream tasks, specifically, how debiasing techniques influence task performance and whether debiased models also make impartial predictions in downstream tasks or not.…

计算与语言 · 计算机科学 2022-06-03 Sullam Jeoung , Jana Diesner

Mental health forums are online communities where people express their issues and seek help from moderators and other users. In such forums, there are often posts with severe content indicating that the user is in acute distress and there…

计算与语言 · 计算机科学 2017-02-23 Arman Cohan , Sydney Young , Andrew Yates , Nazli Goharian

Social media platforms increasingly employ proactive moderation techniques, such as detecting and curbing toxic and uncivil comments, to prevent the spread of harmful content. Despite these efforts, such approaches are often criticized for…

人机交互 · 计算机科学 2025-07-30 Xiaotian Su , Naim Zierau , Soomin Kim , April Yi Wang , Thiemo Wambsganss

This paper analyzes the community safety guidelines of five text-to-image (T2I) generation platforms and audits five T2I models, focusing on prompts related to the representation of humans in areas that might lead to societal stigma. While…

计算机与社会 · 计算机科学 2024-09-27 Piera Riccio , Georgina Curto , Nuria Oliver

Conversation agents, commonly referred to as chatbots, are increasingly deployed in many domains to allow people to have a natural interaction while trying to solve a specific problem. Given their widespread use, it is important to provide…

社会与信息网络 · 计算机科学 2020-10-13 Biplav Srivastava , Francesca Rossi , Sheema Usmani , and Mariana Bernagozzi

Though detection systems have been developed to identify obscene content such as pornography and violence, artificial intelligence is simply not good enough to fully automate this task yet. Due to the need for manual verification, social…

人机交互 · 计算机科学 2020-01-07 Brandon Dang , Martin J. Riedl , Matthew Lease

Harmful and offensive communication or content is detrimental to social bonding and the mental state of users on social media platforms. Text detoxification is a crucial task in natural language processing (NLP), where the goal is removing…

计算与语言 · 计算机科学 2024-04-05 Ali Pesaranghader , Nikhil Verma , Manasa Bharadwaj

NSFW (Not Safe for Work) content, in the context of a dialogue, can have severe side effects on users in open-domain dialogue systems. However, research on detecting NSFW language, especially sexually explicit content, within a dialogue…

计算与语言 · 计算机科学 2024-03-22 Huachuan Qiu , Shuai Zhang , Hongliang He , Anqi Li , Zhenzhong Lan

Online platforms rely on moderation interventions to curb harmful behavior such hate speech, toxicity, and the spread of mis- and disinformation. Yet research on the effects and possible biases of such interventions faces multiple…

社会与信息网络 · 计算机科学 2026-01-28 Aldo Cerulli , Lorenzo Cima , Benedetta Tessa , Serena Tardelli , Stefano Cresci

Generative AI chatbots have proven surprisingly effective at persuading people to change their beliefs and attitudes in lab settings. However, the practical implications of these findings are not yet clear. In this work, we explore the…

人机交互 · 计算机科学 2026-01-29 Jeremy Foote , Deepak Kumar , Bedadyuti Jha , Ryan Funkhouser , Loizos Bitsikokos , Hitesh Goel , Hsuen-Chi Chiu

Moderation and blocking behavior, both closely related to the mitigation of abuse and misinformation on social platforms, are fundamental mechanisms for maintaining healthy online communities. However, while centralized platforms typically…

社会与信息网络 · 计算机科学 2026-03-20 Carlo Bono , Nick Liu , Giuseppe Russo , Francesco Pierri

Twitter bot detection has become a crucial task in efforts to combat online misinformation, mitigate election interference, and curb malicious propaganda. However, advanced Twitter bots often attempt to mimic the characteristics of genuine…

社会与信息网络 · 计算机科学 2023-04-26 Yuhan Liu , Zhaoxuan Tan , Heng Wang , Shangbin Feng , Qinghua Zheng , Minnan Luo

Text-to-image diffusion models can generate diverse content with flexible prompts, which makes them well-suited for customization through fine-tuning with a small amount of user-provided data. However, controllable fine-tuning that prevents…

计算机视觉与模式识别 · 计算机科学 2025-11-19 Ziyao Zeng , Jingcheng Ni , Ruyi Liu , Alex Wong

Internet censorship limits the access of nodes residing within a specific network environment to the public Internet, and vice versa. During the last decade, techniques for conducting Internet censorship have been developed further.…

密码学与安全 · 计算机科学 2025-10-31 Steffen Wendzel , Simon Volpert , Sebastian Zillien , Julia Lenz , Philip Rünz , Luca Caviglione

Decoding from the output distributions of large language models to produce high-quality text is a complex challenge in language modeling. Various approaches, such as beam search, sampling with temperature, $k-$sampling, nucleus…

计算与语言 · 计算机科学 2024-10-22 Esteban Garces Arias , Julian Rodemann , Meimingwei Li , Christian Heumann , Matthias Aßenmacher

The recent development of decentralised and interoperable social networks (such as the "fediverse") creates new challenges for content moderators. This is because millions of posts generated on one server can easily "spread" to another,…

计算机与社会 · 计算机科学 2024-04-18 Vibhor Agarwal , Aravindh Raman , Nishanth Sastry , Ahmed M. Abdelmoniem , Gareth Tyson , Ignacio Castro

Toxic content is one of the most critical issues for social media platforms today. India alone had 518 million social media users in 2020. In order to provide a good experience to content creators and their audience, it is crucial to flag…

计算与语言 · 计算机科学 2022-01-04 Manan Jhaveri , Devanshu Ramaiya , Harveen Singh Chadha

Although not all bots are malicious, the vast majority of them are responsible for spreading misinformation and manipulating the public opinion about several issues, i.e., elections and many more. Therefore, the early detection of bots is…

计算与语言 · 计算机科学 2024-07-31 Loukas Ilias , Ioannis Michail Kazelidis , Dimitris Askounis

The Fediverse, a group of interconnected servers providing a variety of interoperable services (e.g. micro-blogging in Mastodon) has gained rapid popularity. This sudden growth, partly driven by Elon Musk's acquisition of Twitter, has…

社会与信息网络 · 计算机科学 2025-01-13 Haris Bin Zia , Aravindh Raman , Ignacio Castro , Gareth Tyson

Identifying changes in individuals' behaviour and mood, as observed via content shared on online platforms, is increasingly gaining importance. Most research to-date on this topic focuses on either: (a) identifying individuals at risk or…

计算与语言 · 计算机科学 2022-05-12 Adam Tsakalidis , Federico Nanni , Anthony Hills , Jenny Chim , Jiayu Song , Maria Liakata