English
Related papers

Related papers: Harmful Suicide Content Detection

200 papers

Existing methods for evaluating the harmfulness of content generated by large language models (LLMs) have been well studied. However, approaches tailored to multimodal large language models (MLLMs) remain underdeveloped and lack depth. This…

Artificial Intelligence · Computer Science 2025-09-30 Qi Xue , Minrui Jiang , Runjia Zhang , Xiurui Xie , Pei Ke , Guisong Liu

With the widespread availability of pretrained Large Language Models (LLMs) and their training datasets, concerns about the security risks associated with their usage has increased significantly. One of these security risks is the threat of…

Cryptography and Security · Computer Science 2025-06-10 Neil Fendley , Edward W. Staley , Joshua Carney , William Redman , Marie Chau , Nathan Drenkow

Text data can pose a risk of harm. However, the risks are not fully understood, and how to handle, present, and discuss harmful text in a safe way remains an unresolved issue in the NLP community. We provide an analytical framework…

Computation and Language · Computer Science 2023-02-28 Hannah Rose Kirk , Abeba Birhane , Bertie Vidgen , Leon Derczynski

Online communities serve as essential support channels for People Who Use Drugs (PWUD), providing access to peer support and harm reduction information. The moderation of these communities involves consequential decisions affecting member…

Human-Computer Interaction · Computer Science 2026-01-06 Kaixuan Wang , Loraine Clarke , Carl-Cyril J Dreue , Guancheng Zhou , Jason T. Jacques

Detecting online toxicity has always been a challenge due to its inherent subjectivity. Factors such as the context, geography, socio-political climate, and background of the producers and consumers of the posts play a crucial role in…

Social and Information Networks · Computer Science 2023-01-18 Tanmay Garg , Sarah Masud , Tharun Suresh , Tanmoy Chakraborty

The potential of large language models (LLMs) to generate harmful content poses a significant safety risk for data management, as LLMs are increasingly being used as engines for data generation. To assess this risk, numerous harmfulness…

Computation and Language · Computer Science 2026-03-19 Langqi Yang , Tianhang Zheng , Yixuan Chen , Kedong Xiu , Hao Zhou , Wangze Ni , Lei Chen , Zhan Qin , Kui Ren

Social media platforms are plagued by harmful content such as hate speech, misinformation, and extremist rhetoric. Machine learning (ML) models are widely adopted to detect such content; however, they remain highly vulnerable to adversarial…

Machine Learning · Computer Science 2025-12-30 Yidong Chai , Yi Liu , Mohammadreza Ebrahimi , Weifeng Li , Balaji Padmanabhan

As large language models (LLMs) are increasingly deployed in high-stakes settings, the risk of generating harmful or toxic content remains a central challenge. Post-hoc alignment methods are brittle: once unsafe patterns are learned during…

The proliferation of harmful content shared online poses a threat to online information integrity and the integrity of discussion across platforms. Despite various moderation interventions adopted by social media platforms, researchers and…

Social and Information Networks · Computer Science 2023-04-07 Valerio La Gatta , Luca Luceri , Francesco Fabbri , Emilio Ferrara

Peer review is crucial for advancing and improving science through constructive criticism. However, toxic feedback can discourage authors and hinder scientific progress. This work explores an important but underexplored area: detecting…

Computation and Language · Computer Science 2025-02-05 Man Luo , Bradley Peterson , Rafael Gan , Hari Ramalingame , Navya Gangrade , Ariadne Dimarogona , Imon Banerjee , Phillip Howard

Social media has become a valuable resource for the study of suicidal ideation and the assessment of suicide risk. Among social media platforms, Reddit has emerged as the most promising one due to its anonymity and its focus on topic-based…

Computation and Language · Computer Science 2021-06-22 Chenghao Yang , Yudong Zhang , Smaranda Muresan

Toxic comments in online platforms are an unavoidable social issue under the cloak of anonymity. Hate speech detection has been actively done for languages such as English, German, or Italian, where manually labeled corpus has been…

Computation and Language · Computer Science 2020-05-27 Jihyung Moon , Won Ik Cho , Junbum Lee

Online marketplaces, while revolutionizing global commerce, have inadvertently facilitated the proliferation of illicit activities, including drug trafficking, counterfeit sales, and cybercrimes. Traditional content moderation methods such…

Computation and Language · Computer Science 2026-03-06 Quoc Khoa Tran , Thanh Thi Nguyen , Campbell Wilson

The burgeoning capabilities of advanced large language models (LLMs) such as ChatGPT have led to an increase in synthetic content generation with implications across a variety of sectors, including media, cybersecurity, public discourse,…

Computation and Language · Computer Science 2023-10-25 Xianjun Yang , Liangming Pan , Xuandong Zhao , Haifeng Chen , Linda Petzold , William Yang Wang , Wei Cheng

The information ecosystem today is overwhelmed by an unprecedented quantity of data on versatile topics are with varied quality. However, the quality of information disseminated in the field of medicine has been questioned as the negative…

Computers and Society · Computer Science 2020-04-13 Fariha Afsana , Muhammad Ashad Kabir , Naeemul Hassan , Manoranjan Paul

With the widespread availability of LLMs since the release of ChatGPT and increased public scrutiny, commercial model development appears to have focused their efforts on 'safety' training concerning legal liabilities at the expense of…

Computation and Language · Computer Science 2024-08-02 Alina Leidinger , Richard Rogers

Suicide remains a critical global public health issue. While previous studies have provided valuable insights into detecting suicidal expressions in individual social media posts, limited attention has been paid to the analysis of…

Computation and Language · Computer Science 2025-10-17 Jun Li , Qun Zhao

Identifying the targets of hate speech is a crucial step in grasping the nature of such speech and, ultimately, in improving the detection of offensive posts on online forums. Much harmful content on online platforms uses implicit language…

Computation and Language · Computer Science 2024-07-01 Nazanin Jafari , James Allan , Sheikh Muhammad Sarwar

In the light of increasing clues on social media impact on self-harm and suicide risks, there is still no evidence on who are and how factually engaged in suicide-related online behaviors. This study reports new findings of high-performance…

Computers and Society · Computer Science 2024-01-17 Anastasia Peshkovskaya , Yu-Tao Xiang

The abstract outlines the problem of toxic comments on social media platforms, where individuals use disrespectful, abusive, and unreasonable language that can drive users away from discussions. This behavior is referred to as anti-social…

Machine Learning · Computer Science 2023-04-17 K. Poojitha , A. Sai Charish , M. Arun Kuamr Reddy , S. Ayyasamy
‹ Prev 1 8 9 10 Next ›