English
Related papers

Related papers: Evaluating the Effectiveness of OpenAI's Parental …

200 papers

Smart voice assistants (SVAs) have become embedded in the daily lives of youth, introducing complex privacy challenges due to always-on listening, shared device usage, and opaque data practices. This study applies the Privacy-Ethics…

Computers and Society · Computer Science 2026-03-03 Molly Campbell , Yulia Bobkova , Ajay Kumar Shrestha

Background Large language models (LLMs) are increasingly deployed in medical consultations, yet their safety under realistic user pressures remains understudied. Prior assessments focused on neutral conditions, overlooking vulnerabilities…

Computation and Language · Computer Science 2026-01-16 Vahideh Zolfaghari

Large Language Models (LLMs) are increasingly used by teenagers and young adults in everyday life, ranging from emotional support and creative expression to educational assistance. However, their unique vulnerabilities and risk profiles…

Human-Computer Interaction · Computer Science 2025-09-12 Yaman Yu , Yiren Liu , Jacky Zhang , Yun Huang , Yang Wang

This literature review evaluates privacy-by-design frameworks, tools, and policies intended to protect youth in AI-enabled smart devices using a PRISMA-guided workflow. Sources from major academic and grey-literature repositories from the…

Computers and Society · Computer Science 2026-03-03 Molly Campbell , Mohamad Sheikho Al Jasem , Ajay Kumar Shrestha

There are growing concerns about the risks posed by AI companion applications designed for emotional engagement. Existing safety evaluations often rely on self-reported user data or interviews, offering limited insights into real-time…

Computation and Language · Computer Science 2026-05-04 Prerna Juneja , Lika Lomidze

The integration of Artificial Intelligence (AI) systems into technologies used by young digital citizens raises significant privacy concerns. This study investigates these concerns through a comparative analysis of stakeholder perspectives.…

Computers and Society · Computer Science 2025-03-18 Molly Campbell , Ankur Barthwal , Sandhya Joshi , Austin Shouli , Ajay Kumar Shrestha

To understand and identify the unprecedented risks posed by rapidly advancing artificial intelligence (AI) models, Frontier AI Risk Management Framework in Practice presents a comprehensive assessment of their frontier risks. As Large…

Recent advances in AI agents capable of solving complex, everyday tasks, from scheduling to customer service, have enabled deployment in real-world settings, but their possibilities for unsafe behavior demands rigorous evaluation. While…

Artificial Intelligence · Computer Science 2026-02-18 Sanidhya Vijayvargiya , Aditya Bharat Soni , Xuhui Zhou , Zora Zhiruo Wang , Nouha Dziri , Graham Neubig , Maarten Sap

This paper studies how parents want to moderate children's interactions with Generative AI chatbots, with the goal of informing the design of future GenAI parental control tools. We first used an LLM to generate synthetic child-GenAI…

Human-Computer Interaction · Computer Science 2026-03-13 John Driscoll , Yulin Chen , Viki Shi , Izak Vucharatavintara , Yaxing Yao , Haojian Jin

AI chatbots have quietly become the world's most popular therapists, coaches, and confidants. Users of cloud-based LLM services are increasingly shifting from simple queries like idea generation and poem writing, to deeply personal…

Human-Computer Interaction · Computer Science 2026-04-08 Max Holschneider , Saetbyeol LeeYouk

The prevalence of short form video platforms, combined with the ineffectiveness of age verification mechanisms, raises concerns about the potential harms facing children and teenagers in an algorithm-moderated online environment. We…

Computers and Society · Computer Science 2026-05-27 Haoning Xue , Brian Nishimine , Martin Hilbert , Drew Cingel , Samantha Vigil , Jane Shawcroft , Arti Thakur , Zubair Shafiq , Jingwen Zhang

We introduce a multi-turn benchmark for evaluating personalised alignment in LLM-based AI assistants, focusing on their ability to handle user-provided safety-critical contexts. Our assessment of ten leading models across five scenarios…

Human-Computer Interaction · Computer Science 2025-01-31 Lize Alberts , Benjamin Ellis , Andrei Lupu , Jakob Foerster

AI character platforms, which allow users to engage in conversations with AI personas, are a rapidly growing application domain. However, their immersive and personalized nature, combined with technical vulnerabilities, raises significant…

Cryptography and Security · Computer Science 2025-12-02 Yiluo Wei , Peixian Zhang , Gareth Tyson

Human safety awareness gaps often prevent the timely recognition of everyday risks. In solving this problem, a proactive safety artificial intelligence (AI) system would work better than a reactive one. Instead of just reacting to users'…

Computation and Language · Computer Science 2025-10-21 Youliang Yuan , Wenxiang Jiao , Yuejin Xie , Chihao Shen , Menghan Tian , Wenxuan Wang , Jen-tse Huang , Pinjia He

Parental control apps, which are mobile apps that allow parents to monitor and restrict their children's activities online, are becoming increasingly adopted by parents as a means of safeguarding their children's online safety. However, it…

Human-Computer Interaction · Computer Science 2021-09-14 Ge Wang , Jun Zhao , Max Van Kleek , Nigel Shadbolt

Generative AI (GenAI) is a powerful technology poised to reshape Trust & Safety. While misuse by attackers is a growing concern, its defensive capacity remains underexplored. This paper examines these effects through a qualitative study…

Human-Computer Interaction · Computer Science 2026-04-24 Patrick Gage Kelley , Steven Rousso-Schindler , Renee Shelby , Kurt Thomas , Allison Woodruff

AI safety is a rapidly growing area of research that seeks to prevent the harm and misuse of frontier AI technology, particularly with respect to generative AI (GenAI) tools that are capable of creating realistic and high-quality content…

Artificial Intelligence · Computer Science 2025-02-19 Pin-Yu Chen

While recent research has focused on developing safeguards for generative AI (GAI) model-level content safety, little is known about how content moderation to prevent malicious content performs for end-users in real-world GAI products. To…

Human-Computer Interaction · Computer Science 2025-06-18 Lan Gao , Oscar Chen , Rachel Lee , Nick Feamster , Chenhao Tan , Marshini Chetty

The current "notice and consent" paradigm is broken: consent dialogues are often manipulative, and users cannot realistically read or understand every privacy policy. While recent LLM-based tools empower users seeking active control, many…

Human-Computer Interaction · Computer Science 2026-04-24 Vincent Freiberger

This research offers a unique evaluation of how AI systems interpret the digital language of Generation Alpha (Gen Alpha, born 2010-2024). As the first cohort raised alongside AI, Gen Alpha faces new forms of online risk due to immersive…

Computers and Society · Computer Science 2025-05-19 Manisha Mehta , Fausto Giunchiglia