English
Related papers

Related papers: Using Grok to Avoid Personal Attacks While Correct…

200 papers

Many recent advances in natural language generation have been fueled by training large language models on internet-scale data. However, this paradigm can lead to models that generate toxic, inaccurate, and unhelpful content, and automatic…

This paper highlights the developing need for quantitative modes for capturing and monitoring malicious communication in social media. There has been a deliberate "weaponization" of messaging through the use of social networks including by…

Computers and Society · Computer Science 2024-05-28 Andy Skumanich , Han Kyul Kim

Most adversarial threats in artificial intelligence (AI) target the computational behavior of models rather than the humans who rely on them. Yet modern AI systems increasingly operate within human decision loops, where users interpret and…

Artificial Intelligence · Computer Science 2026-05-18 Shutong Fan , Lan Zhang , Xiaoyong Yuan

Large Language Models (LLMs) are becoming increasingly persuasive, demonstrating the ability to personalize arguments in conversation with humans by leveraging their personal data. This may have serious impacts on the scale and…

Computation and Language · Computer Science 2025-01-30 Jasper Timm , Chetan Talele , Jacob Haimes

Misconceptions in psychology and education persist despite clear contradictory evidence, resisting traditional correction methods. This study investigated whether personalised AI dialogue could effectively correct these stubborn beliefs. In…

Human-Computer Interaction · Computer Science 2025-06-12 Brooklyn J. Corbett , Jason M. Tangen

Polarization, declining trust, and wavering support for democratic norms are pressing threats to U.S. democracy. Exposure to verified and quality news may lower individual susceptibility to these threats and make citizens more resilient to…

Social and Information Networks · Computer Science 2024-04-02 Hadi Askari , Anshuman Chhabra , Bernhard Clemm von Hohenberg , Michael Heseltine , Magdalena Wojcieszak

Large language models have many beneficial applications, but can they also be used to attack content-filtering algorithms in social media platforms? We investigate the challenge of generating adversarial examples to test the robustness of…

Computation and Language · Computer Science 2025-09-04 Piotr Przybyła , Euan McGill , Horacio Saggion

Social media users drive the spread of misinformation online by sharing posts that include erroneous information or commenting on controversial topics with unsubstantiated arguments often in earnest. Work on echo chambers has suggested that…

Social and Information Networks · Computer Science 2024-04-25 Xinyu Wang , Jiayi Li , Sarah Rajtmajer

Since 2016, the amount of academic research with the keyword "misinformation" has more than doubled [2]. This research often focuses on article headlines shown in artificial testing environments, yet misinformation largely spreads through…

Human-Computer Interaction · Computer Science 2020-12-16 Emily Saltz , Claire Leibowicz , Claire Wardle

Advances in generative AI (GenAI) have raised concerns about detecting and discerning AI-generated content from human-generated content. Most existing literature assumes a paradigm where 'expert' organized disinformation creators and flawed…

Human-Computer Interaction · Computer Science 2024-06-19 Amelia Hassoun , Ariel Abonizio , Katy Osborn , Cameron Wu , Beth Goldberg

Generative AI and misinformation research has evolved since our 2024 survey. This paper presents an updated perspective, transitioning from literature review to practical countermeasures. We report on changes in the threat landscape,…

Computers and Society · Computer Science 2026-02-12 Alexander Loth , Martin Kappes , Marc-Oliver Pahl

This study evaluates the effectiveness of ChatGPT, an advanced AI model for natural language processing, in identifying targeting and inappropriate language in online comments. With the increasing challenge of moderating vast volumes of…

Computation and Language · Computer Science 2025-05-29 Barbarestani Baran , Maks Isa , Vossen Piek

Across various applications, humans increasingly use black-box artificial intelligence (AI) systems without insight into these systems' reasoning. To counter this opacity, explainable AI (XAI) methods promise enhanced transparency and…

Human-Computer Interaction · Computer Science 2025-01-09 Philipp Spitzer , Joshua Holstein , Katelyn Morrison , Kenneth Holstein , Gerhard Satzger , Niklas Kühl

Existing challenges in misinformation exposure and susceptibility vary across demographic groups, as some populations are more vulnerable to misinformation than others. Large language models (LLMs) introduce new dimensions to these…

Computation and Language · Computer Science 2025-10-15 Angana Borah , Rada Mihalcea , Verónica Pérez-Rosas

Social media platforms often assume that users can self-correct against misinformation. However, social media users are not equally susceptible to all misinformation as their biases influence what types of misinformation might thrive and…

The proliferation of misinformation poses a significant threat to society, exacerbated by the capabilities of generative AI. This demo paper introduces Veracity, an open-source AI system designed to empower individuals to combat…

Social media manipulation poses a significant threat to cognitive autonomy and unbiased opinion formation. Prior literature explored the relationship between online activity and emotional state, cognitive resources, sunlight and weather.…

Social and Information Networks · Computer Science 2023-12-21 Elisabeth Stockinger , Riccardo Gallotti , Carina I. Hausladen

Question generation has recently gained a lot of research interest, especially with the advent of large language models. In and of itself, question generation can be considered 'AI-hard', as there is a lack of unanimously agreed sense of…

Information Retrieval · Computer Science 2022-10-19 Sreehari Sankar , Zhihang Dong

Disinformation is among the top risks of generative artificial intelligence (AI) misuse. Global adoption of generative AI necessitates red-teaming evaluations (i.e., systematic adversarial probing) that are robust across diverse languages…

Computation and Language · Computer Science 2025-09-24 Alejandro Cuevas , Saloni Dash , Bharat Kumar Nayak , Dan Vann , Madeleine I. G. Daepp

In this paper, we explore the feasibility of leveraging large language models (LLMs) to automate or otherwise assist human raters with identifying harmful content including hate speech, harassment, violent extremism, and election…