English
Related papers

Related papers: Explainable Abuse Detection as Intent Classificati…

200 papers

One of the main tasks of cybersecurity is recognizing malicious interactions with an arbitrary system. Currently, the logging information from each interaction can be collected in almost unrestricted amounts, but identification of attacks…

Cryptography and Security · Computer Science 2019-07-02 Linara Adilova , Livin Natious , Siming Chen , Olivier Thonnard , Michael Kamp

The age of social media is flooded with Internet memes, necessitating a clear grasp and effective identification of harmful ones. This task presents a significant challenge due to the implicit meaning embedded in memes, which is not…

Computation and Language · Computer Science 2024-01-25 Hongzhan Lin , Ziyang Luo , Wei Gao , Jing Ma , Bo Wang , Ruichao Yang

Sensitive information detection is crucial in content moderation to maintain safe online communities. Assisting in this traditionally manual process could relieve human moderators from overwhelming and tedious tasks, allowing them to focus…

Online platforms and communities establish their own norms that govern what behavior is acceptable within the community. Substantial effort in NLP has focused on identifying unacceptable behaviors and, recently, on forecasting them before…

Computation and Language · Computer Science 2021-10-12 Chan Young Park , Julia Mendelsohn , Karthik Radhakrishnan , Kinjal Jain , Tushar Kanakagiri , David Jurgens , Yulia Tsvetkov

In recent years social media has become an increasingly popular tool for communication. People use it to share their ideas, exchange information, and discuss thoughts. Given its prevalence and widespread reach, social media must remain a…

Computation and Language · Computer Science 2026-05-22 Pranshu Rastogi , Madhav Mathur , Ramaneswaran S , Kshitij Mohan

Users of social platforms often perceive these sites as supportive spaces to post about their mental health issues. Those conversations contain important traces about individuals' health risks. Recently, researchers have exploited this…

Computation and Language · Computer Science 2024-08-21 Eliseo Bao , Anxo Pérez , Javier Parapar

Toxicity detection algorithms, originally designed with reactive content moderation in mind, are increasingly being deployed into proactive end-user interventions to moderate content. Through a socio-technical lens and focusing on contexts…

Human-Computer Interaction · Computer Science 2025-02-25 Mark Warner , Angelika Strohmayer , Matthew Higgs , Lynne Coventry

With machine learning models being increasingly used to aid decision making even in high-stakes domains, there has been a growing interest in developing interpretable models. Although many supposedly interpretable models have been proposed,…

Artificial Intelligence · Computer Science 2021-08-17 Forough Poursabzi-Sangdeh , Daniel G. Goldstein , Jake M. Hofman , Jennifer Wortman Vaughan , Hanna Wallach

Online Social Media represent a pervasive source of information able to reach a huge audience. Sadly, recent studies show how online social bots (automated, often malicious accounts, populating social networks and mimicking genuine users)…

Social and Information Networks · Computer Science 2019-09-10 Alessandro Balestrucci , Rocco De Nicola , Marinella Petrocchi , Catia Trubiani

Malicious actors exploit social media to inflate stock prices, sway elections, spread misinformation, and sow discord. To these ends, they employ tactics that include the use of inauthentic accounts and campaigns. Methods to detect these…

Social and Information Networks · Computer Science 2022-11-02 Alexander C. Nwala , Alessandro Flammini , Filippo Menczer

Reducing hateful and offensive content in online social media pose a dual problem for the moderators. On the one hand, rigid censorship on social media cannot be imposed. On the other, the free flow of such content cannot be allowed. Hence,…

Social and Information Networks · Computer Science 2019-09-30 Punyajoy Saha , Binny Mathew , Pawan Goyal , Animesh Mukherjee

For more than a decade now, academicians and online platform administrators have been studying solutions to the problem of bot detection. Bots are computer algorithms whose use is far from being benign: malicious bots are purposely created…

Cryptography and Security · Computer Science 2025-06-25 Rocco De Nicola , Marinella Petrocchi , Manuel Pratelli

The pervasive use of social media platforms, such as Facebook, Instagram, and X, has significantly amplified our electronic interconnectedness. Moreover, these platforms are now easily accessible from any location at any given time.…

Social and Information Networks · Computer Science 2024-02-21 Abulkarim Faraj Alqahtani , Mohammad Ilyas

The rise of social media platforms has led to an increase in cyber-aggressive behavior, encompassing a broad spectrum of hostile behavior, including cyberbullying, online harassment, and the dissemination of offensive and hate speech. These…

Computation and Language · Computer Science 2024-12-31 Swapnil Mane , Suman Kundu , Rajesh Sharma

This paper addresses the important problem of discerning hateful content in social media. We propose a detection scheme that is an ensemble of Recurrent Neural Network (RNN) classifiers, and it incorporates various features associated with…

Computation and Language · Computer Science 2019-07-05 Georgios K. Pitsilis , Heri Ramampiaro , Helge Langseth

In recent years, online social networks have allowed worldwide users to meet and discuss. As guarantors of these communities, the administrators of these platforms must prevent users from adopting inappropriate behaviors. This verification…

Information Retrieval · Computer Science 2019-06-17 Noé Cecillon , Vincent Labatut , Richard Dufour , Georges Linarès

While social media offer great communication opportunities, they also increase the vulnerability of young people to threatening situations online. Recent studies report that cyberbullying constitutes a growing problem among youngsters.…

Computation and Language · Computer Science 2020-03-03 Cynthia Van Hee , Gilles Jacobs , Chris Emmery , Bart Desmet , Els Lefever , Ben Verhoeven , Guy De Pauw , Walter Daelemans , Véronique Hoste

With the spread of online social networks, it is more and more difficult to monitor all the user-generated content. Automating the moderation process of the inappropriate exchange content on Internet has thus become a priority task. Methods…

Computation and Language · Computer Science 2021-01-19 Noé Cecillon , Vincent Labatut , Richard Dufour , Georges Linares

The datasets most widely used for abusive language detection contain lists of messages, usually tweets, that have been manually judged as abusive or not by one or more annotators, with the annotation performed at message level. In this…

Computation and Language · Computer Science 2021-03-30 Stefano Menini , Alessio Palmero Aprosio , Sara Tonelli

Social media platforms increasingly employ proactive moderation techniques, such as detecting and curbing toxic and uncivil comments, to prevent the spread of harmful content. Despite these efforts, such approaches are often criticized for…

Human-Computer Interaction · Computer Science 2025-07-30 Xiaotian Su , Naim Zierau , Soomin Kim , April Yi Wang , Thiemo Wambsganss