中文
相关论文

相关论文: Red-Teaming for Generative AI: Silver Bullet or Se…

200 篇论文

Cyber warfare has become a central element of modern conflict, especially within multi-domain operations. As both a distinct and critical domain, cyber warfare requires integrating defensive and offensive technologies into coherent…

计算机科学与博弈论 · 计算机科学 2026-02-26 Ya-Ting Yang , Quanyan Zhu

Generative Artificial Intelligence (AI) stands as a transformative force that presents a paradox; it offers unprecedented opportunities for productivity growth while potentially posing significant threats to economic stability and societal…

Generative Artificial Intelligence (GenAI) is rapidly reshaping the global financial landscape, offering unprecedented opportunities to enhance customer engagement, automate complex workflows, and extract actionable insights from vast…

密码学与安全 · 计算机科学 2025-05-01 Bikash Saha , Nanda Rani , Sandeep Kumar Shukla

Generative AI has unleashed the power of content generation and it has also unwittingly opened the pandora box of realistic deepfake causing a number of social hazards and harm to businesses and personal reputation. The investigation &…

人工智能 · 计算机科学 2026-01-13 Prasanna Kumar

Red teaming assesses how large language models (LLMs) can produce content that violates norms, policies, and rules set during their safety training. However, most existing automated methods in the literature are not representative of the…

Behind the scenes of maintaining the safety of technology products from harmful and illegal digital content lies unrecognized human labor. The recent rise in the use of generative AI technologies and the accelerating demands to meet…

人机交互 · 计算机科学 2024-11-05 Alice Qian Zhang , Judith Amores , Mary L. Gray , Mary Czerwinski , Jina Suh

Generative Artificial Intelligence (AI) has seen mainstream adoption lately, especially in the form of consumer-facing, open-ended, text and image generating models. However, the use of such systems raises significant ethical and safety…

计算机与社会 · 计算机科学 2023-08-10 Avijit Ghosh , Dhanya Lakshmi

As generative Artificial Intelligence (genAI) technologies proliferate across sectors, they offer significant benefits but also risk exacerbating discrimination. This chapter explores how genAI intersects with non-discrimination laws,…

计算机与社会 · 计算机科学 2025-03-10 Philipp Hacker

As large language models (LLMs) are increasingly used for code generation, concerns over the security risks have grown substantially. Early research has primarily focused on red teaming, which aims to uncover and evaluate vulnerabilities…

软件工程 · 计算机科学 2025-10-22 Chengquan Guo , Yuzhou Nie , Chulin Xie , Zinan Lin , Wenbo Guo , Bo Li

There is an increasing need for young people to become critically AI literate, understanding not only how AI works but also its limitations and ethical nuances. Yet, designing learning experiences that make such complex, serious topics…

人机交互 · 计算机科学 2026-04-03 Jaemarie Solyst , Ruth Karen Nakigozi , Chloe Fong , R. Benjamin Shapiro

To understand and identify the unprecedented risks posed by rapidly advancing artificial intelligence (AI) models, this report presents a comprehensive assessment of their frontier risks. Drawing on the E-T-C analysis (deployment…

Studies of Generative AI (GenAI)-assisted creative workflows have focused on individuals overcoming challenges of prompting to produce what they envisioned. When designers work in teams, how do collaboration and prompting influence each…

人机交互 · 计算机科学 2025-09-29 Yuanning Han , Ziyi Qiu , Jiale Cheng , RAY LC

Generative AI tools are increasingly entering academic peer review workflows, raising questions about fairness, accountability, and the legitimacy of evaluative judgment. While these systems promise efficiency gains amid growing reviewer…

计算机与社会 · 计算机科学 2026-03-24 Tatiana Chakravorti , Pranav Narayanan Venkit , Sourojit Ghosh , Sarah Rajtmajer

Recent advancements in generative artificial intelligence (AI) have transformed collaborative work processes, yet the impact on team performance remains underexplored. Here we examine the role of generative AI in enhancing or replacing…

人机交互 · 计算机科学 2024-05-29 Ning Li , Huaikang Zhou , Kris Mikel-Hong

Agentic Artificial Intelligence (AI) builds upon Generative AI (GenAI). It constitutes the next major step in the evolution of AI with much stronger reasoning and interaction capabilities that enable more autonomous behavior to tackle…

人工智能 · 计算机科学 2025-04-29 Johannes Schneider

Generative artificial intelligence (GenAI) has the potential to improve healthcare through automation that enhances the quality and safety of patient care. Powered by foundation models that have been pretrained and can generate complex…

计算机与社会 · 计算机科学 2024-07-25 Laleh Jalilian , Daniel McDuff , Achuta Kadambi

The integration of Artificial Intelligence (AI) necessitates determining whether systems function as tools or collaborative teammates. In this study, by synthesizing Human-AI Interaction (HAI) literature, we analyze this distinction across…

Generative AI (GenAI) is a powerful technology poised to reshape Trust & Safety. While misuse by attackers is a growing concern, its defensive capacity remains underexplored. This paper examines these effects through a qualitative study…

人机交互 · 计算机科学 2026-04-24 Patrick Gage Kelley , Steven Rousso-Schindler , Renee Shelby , Kurt Thomas , Allison Woodruff

This research-in-progress paper presents a new project management framework that utilises GenAI technology. The framework is designed to address the common challenge of uniform team compositions in academic and research project teams,…

计算机与社会 · 计算机科学 2026-04-02 Johnny Chan , Yuming Li

Rapidly advancing artificial intelligence (AI) systems introduce novel, uncertain, and potentially catastrophic risks. Managing these risks requires a mature risk-management infrastructure whose cornerstone is rigorous risk modeling. We…