English
Related papers

Related papers: Sensitive Content Classification in Social Media: …

200 papers

Mental health challenges and cyberbullying are increasingly prevalent in digital spaces, necessitating scalable and interpretable detection systems. This paper introduces a unified multiclass classification framework for detecting ten…

Computation and Language · Computer Science 2026-03-26 Edward Ajayi , Martha Kachweka , Mawuli Deku , Emily Aiken

Social media text shows promise for monitoring trends in the opioid overdose crisis; however, the overwhelming majority of social media text is unrelated to opioids. When leveraging social media text to monitor trends in the ongoing opioid…

Computation and Language · Computer Science 2026-03-12 Kristy A. Carpenter , Issah A. Samori , Mathew V. Kiang , Keith Humphreys , Anna Lembke , Johannes C. Eichstaedt , Russ B. Altman

Commercial content moderation APIs are marketed as scalable solutions to combat online hate speech. However, the reliance on these APIs risks both silencing legitimate speech, called over-moderation, and failing to protect online platforms…

Human-Computer Interaction · Computer Science 2025-03-04 David Hartmann , Amin Oueslati , Dimitri Staufer , Lena Pohlmann , Simon Munzert , Hendrik Heuer

The volume of machine-generated content online has grown dramatically due to the widespread use of Large Language Models (LLMs), leading to new challenges for content moderation systems. Conventional content moderation classifiers, which…

Computation and Language · Computer Science 2026-05-26 Shaz Furniturewala , Arkaitz Zubiaga

The rapid growth in user generated content on social media has resulted in a significant rise in demand for automated content moderation. Various methods and frameworks have been proposed for the tasks of hate speech detection and toxic…

Computation and Language · Computer Science 2024-09-27 Elizaveta Korotkova , Isaac Chung

With the development of web technology, social media texts are becoming a rich source for automatic mental health analysis. As traditional discriminative methods bear the problem of low interpretability, the recent large language models…

Computation and Language · Computer Science 2024-02-14 Kailai Yang , Tianlin Zhang , Ziyan Kuang , Qianqian Xie , Jimin Huang , Sophia Ananiadou

Research shows that exposure to suicide-related news media content is associated with suicide rates, with some content characteristics likely having harmful and others potentially protective effects. Although good evidence exists for a few…

Computation and Language · Computer Science 2022-06-29 Hannah Metzler , Hubert Baginski , Thomas Niederkrotenthaler , David Garcia

Content moderation plays a critical role in shaping safe and inclusive online environments, balancing platform standards, user expectations, and regulatory frameworks. Traditionally, this process involves operationalising policies into…

Large-scale web-scraped text corpora used to train general-purpose AI models often contain harmful demographic-targeted social biases, creating a regulatory need for data auditing and developing scalable bias-detection methods. Although…

Computation and Language · Computer Science 2026-04-10 Ayan Majumdar , Feihao Chen , Jinghui Li , Xiaozhen Wang

Detecting hateful content is a challenging and important problem. Automated tools, like machine-learning models, can help, but they require continuous training to adapt to the ever-changing landscape of social media. In this work, we…

Computation and Language · Computer Science 2025-11-06 Jay Patel , Hrudayangam Mehta , Jeremy Blackburn

The widespread use of social media necessitates reliable and efficient detection of offensive content to mitigate harmful effects. Although sophisticated models perform well on individual datasets, they often fail to generalize due to…

Computation and Language · Computer Science 2024-10-08 Huy Nghiem , Hal Daumé

Effective toxic content detection relies heavily on high-quality and diverse data, which serve as the foundation for robust content moderation models. Synthetic data has become a common approach for training models across various NLP tasks.…

Computation and Language · Computer Science 2025-02-25 Zheng Hui , Zhaoxiao Guo , Hang Zhao , Juanyong Duan , Lin Ai , Yinheng Li , Julia Hirschberg , Congrui Huang

The proliferation of harmful content on online social media platforms has necessitated empirical understandings of experiences of harm online and the development of practices for harm mitigation. Both understandings of harm and approaches…

Human-Computer Interaction · Computer Science 2021-09-20 Morgan Klaus Scheuerman , Jialun Aaron Jiang , Casey Fiesler , Jed R. Brubaker

Violent threats remain a significant problem across social media platforms. Useful, high-quality data facilitates research into the understanding and detection of malicious content, including violence. In this paper, we introduce a…

Social media influence campaigns pose significant challenges to public discourse and democracy. Traditional detection methods fall short due to the complexity and dynamic nature of social media. Addressing this, we propose a novel detection…

Social and Information Networks · Computer Science 2023-11-15 Luca Luceri , Eric Boniardi , Emilio Ferrara

Most social media users come from the Global South, where harmful content usually appears in local languages. Yet, AI-driven moderation systems struggle with low-resource languages spoken in these regions. Through semi-structured interviews…

Computation and Language · Computer Science 2025-08-06 Farhana Shahid , Mona Elswah , Aditya Vashistha

The growing demand for accessible mental health support, compounded by workforce shortages and logistical barriers, has led to increased interest in utilizing Large Language Models (LLMs) for scalable and real-time assistance. However,…

Human-Computer Interaction · Computer Science 2025-05-27 Ugur Kursuncu , Trilok Padhi , Gaurav Sinha , Abdulkadir Erol , Jaya Krishna Mandivarapu , Christopher R. Larrison

Large Language Models (LLMs) have transformed natural language processing (NLP) by enabling robust text generation and understanding. However, their deployment in sensitive domains like healthcare, finance, and legal services raises…

Artificial Intelligence · Computer Science 2024-12-09 Georgios Feretzakis , Vassilios S. Verykios

Mental health forums are online communities where people express their issues and seek help from moderators and other users. In such forums, there are often posts with severe content indicating that the user is in acute distress and there…

Computation and Language · Computer Science 2017-02-23 Arman Cohan , Sydney Young , Andrew Yates , Nazli Goharian

Growing evidence shows that proactive content moderation supported by AI can help improve online discourse. However, we know little about designing these systems, how design impacts efficacy and user experience, and how people perceive…

Human-Computer Interaction · Computer Science 2024-01-22 Mark Warner , Angelika Strohmayer , Matthew Higgs , Husnain Rafiq , Liying Yang , Lynne Coventry