中文
相关论文

相关论文: Intersectionality in Conversational AI Safety: How…

200 篇论文

Bicycle safety is important for bikeability and transportation efficiency. However, conventional surveys often fall short in capturing how people actually perceive cycling environments because they rely heavily on respondents' recall rather…

计算机与社会 · 计算机科学 2026-04-10 Feiyang Ren , Zhaoxi Zhang , Tamir Mendel , Takahiro Yabe

As language models evolve into autonomous agents that act and communicate on behalf of users, ensuring safety in multi-agent ecosystems becomes a central challenge. Interactions between personal assistants and external service providers…

密码学与安全 · 计算机科学 2025-11-10 Amr Gomaa , Ahmed Salem , Sahar Abdelnabi

Dialogue safety remains a pervasive challenge in open-domain human-machine interaction. Existing approaches propose distinctive dialogue safety taxonomies and datasets for detecting explicitly harmful responses. However, these taxonomies…

计算与语言 · 计算机科学 2023-08-01 Huachuan Qiu , Tong Zhao , Anqi Li , Shuai Zhang , Hongliang He , Zhenzhong Lan

Large language models offer significant potential for increasing labour productivity, such as streamlining personnel selection, but raise concerns about perpetuating systemic biases embedded into their pre-training data. This study explores…

综合经济学 · 经济学 2024-02-13 Louis Lippens

This article makes two key contributions to methodological debates in automation research. First, we argue for and demonstrate how methods in this field must account for intersections of social difference, such as race, class, ethnicity,…

计算机与社会 · 计算机科学 2023-08-31 Shanthi Robertson , Liam Magee , Karen Soldatić

Safeguard models help large language models (LLMs) detect and block harmful content, but most evaluations remain English-centric and overlook linguistic and cultural diversity. Existing multilingual safety benchmarks often rely on…

计算与语言 · 计算机科学 2025-12-08 Panuthep Tasawong , Jian Gang Ngui , Alham Fikri Aji , Trevor Cohn , Peerat Limkonchotiwat

Gender stereotypes are pervasive beliefs about individuals based on their gender that play a significant role in shaping societal attitudes, behaviours, and even opportunities. Recognizing the negative implications of gender stereotypes,…

计算与语言 · 计算机科学 2024-04-19 Isar Nejadgholi , Kathleen C. Fraser , Anna Kerkhof , Svetlana Kiritchenko

AI chatbots, built using large language models, are increasingly integrated into society and mimic the patterns of human text exchanges. While previous research has raised concerns that humans may form romantic attachment to chatbots, the…

人机交互 · 计算机科学 2026-03-03 Julia Kieserman , Cat Mai , Sara Lignell , Lucy Qin , Athanasios Andreou , Damon McCoy , Rosanna Bellini

The issue of fairness in AI has received an increasing amount of attention in recent years. The problem can be approached by looking at different protected attributes (e.g., ethnicity, gender, etc) independently, but fairness for individual…

机器学习 · 计算机科学 2023-02-27 Giulio Filippi , Sara Zannone , Adriano Koshiyama

AI chatbots are an emerging security attack vector, vulnerable to threats such as prompt injection, and rogue chatbot creation. When deployed in domains such as corporate security policy, they could be weaponized to deliver guidance that…

人机交互 · 计算机科学 2025-10-13 Brandon Lit , Edward Crowder , Daniel Vogel , Hassan Khan

In recent months, the social impact of Artificial Intelligence (AI) has gained considerable public interest, driven by the emergence of Generative AI models, ChatGPT in particular. The rapid development of these models has sparked heated…

As AI systems become increasingly integrated into daily life, their potential to exacerbate or trigger severe psychological harms remains poorly understood and inadequately tested. This paper presents a proactive methodology for…

Artificial Intelligence (AI) is reshaping many societal domains, raising critical questions about its risks, benefits, and the potential misalignment between public and academic perspectives. This study examines how the general public…

计算机与社会 · 计算机科学 2026-05-05 Philipp Brauner , Felix Glawe , Gian Luca Liehner , Luisa Vervier , Martina Ziefle

Despite increasing AI chatbot deployment in public discourse, empirical evidence on their capacity to foster intercultural empathy remains limited. Through a randomized experiment, we assessed how different AI deliberation…

人机交互 · 计算机科学 2025-07-29 Isabel Villanueva , Tara Bobinac , Binwei Yao , Junjie Hu , Kaiping Chen

With the rapid uptake of generative AI, investigating human perceptions of generated responses has become crucial. A major challenge is their `aptitude' for hallucinating and generating harmful contents. Despite major efforts for…

As generative large model capabilities advance, safety concerns become more pronounced in their outputs. To ensure the sustainable growth of the AI ecosystem, it's imperative to undertake a holistic evaluation and refinement of associated…

人工智能 · 计算机科学 2023-12-01 Jiawen Deng , Jiale Cheng , Hao Sun , Zhexin Zhang , Minlie Huang

Xenophobia is one of the key drivers of marginalisation, discrimination, and conflict, yet many prominent machine learning (ML) fairness frameworks fail to comprehensively measure or mitigate the resulting xenophobic harms. Here we aim to…

计算机与社会 · 计算机科学 2023-10-09 Nenad Tomasev , Jonathan Leader Maynard , Iason Gabriel

Argumentative dialogues across political divides can reduce polarization, yet opportunities for citizens to engage with opposing views in accessible and structured ways remain limited. AI dialogue partners offer a scalable framework for…

计算机与社会 · 计算机科学 2026-05-25 Jianlong Zhu , Syed Muhammad Jhon Raza Naqvi , Carolin-Theresa Ziemer , Usman Naseem , Ingmar Weber

Selecting a college major is a difficult decision for many incoming freshmen. Traditional academic advising is often hindered by long wait times, intimidating environments, and limited personalization. AI Chatbots present an opportunity to…

计算机与社会 · 计算机科学 2025-10-14 Prarthana P. Kartholy , Thandi M. Labor , Neil N. Panchal , Sean H. Wang , Hillary N. Owusu

Conversational AI systems are becoming famous in day to day lives. In this paper, we are trying to address the following key question: To identify whether design, as well as development efforts for search oriented conversational AI are…

人工智能 · 计算机科学 2017-09-15 Mahipal Jadeja , Neelanshi Varia