中文
相关论文

相关论文: X-Risk Analysis for AI Research

200 篇论文

The impact of Artificial Intelligence does not depend only on fundamental research and technological developments, but for a large part on how these systems are introduced into society and used in everyday situations. AI is changing the way…

计算机与社会 · 计算机科学 2022-05-24 Virginia Dignum

This paper examines the systemic risks posed by incremental advancements in artificial intelligence, developing the concept of `gradual disempowerment', in contrast to the abrupt takeover scenarios commonly discussed in AI safety. We…

计算机与社会 · 计算机科学 2025-01-30 Jan Kulveit , Raymond Douglas , Nora Ammann , Deger Turan , David Krueger , David Duvenaud

As Artificial Intelligence (AI) systems become increasingly integrated into various aspects of daily life, concerns about privacy and ethical accountability are gaining prominence. This study explores stakeholder perspectives on privacy in…

计算机与社会 · 计算机科学 2025-03-18 Ajay Kumar Shrestha , Sandhya Joshi

With the turmoil in cybersecurity and the mind-blowing advances in AI, it is only natural that cybersecurity practitioners consider further employing learning techniques to help secure their organizations and improve the efficiency of their…

密码学与安全 · 计算机科学 2019-12-17 Ricardo Morla

Artificial intelligence (AI) is an emerging technology that has the potential to transform many aspects of society, including the economy, healthcare, and transportation. This article synthesizes recent research literature on the global…

密码学与安全 · 计算机科学 2024-01-24 Chandregowda Pachegowda

Cybersecurity research involves publishing papers about malicious exploits as much as publishing information on how to design tools to protect cyber-infrastructure. It is this information exchange between ethical hackers and security…

人工智能 · 计算机科学 2017-06-06 Federico Pistono , Roman V. Yampolskiy

Our survey of 53 specialists across 105 AI reliability and security research areas identifies the most promising research prospects to guide strategic AI R&D investment. As companies are seeking to develop AI systems with broadly…

计算机与社会 · 计算机科学 2025-05-29 Joe O'Brien , Jeremy Dolan , Jay Kim , Jonah Dykhuizen , Jeba Sania , Sebastian Becker , Jam Kraprayoon , Cara Labrador

As Artificial Intelligence (AI) systems continue to grow in size and complexity, so does the difficulty of the quest for AI transparency. In a world of large models and complex AI systems, why do we explain AI and what should we explain?…

人工智能 · 计算机科学 2026-04-23 Karina Cortinas-Lorenzo , Gavin Doherty

Developing and certifying safe - or so-called trustworthy - AI has become an increasingly salient issue, especially in light of upcoming regulation such as the EU AI Act. In this context, the black-box nature of machine learning models…

As Artificial Intelligence (AI) systems proliferate, the need for systematic, transparent, and actionable processes for evaluating them is growing. While many resources exist to support AI evaluation, they have several limitations. Few…

计算机与社会 · 计算机科学 2026-02-02 Rachel M. Kim , Blaine Kuehnert , Alice Lai , Kenneth Holstein , Hoda Heidari , Rayid Ghani

AI-controlled robotic systems pose a risk to human workers and the environment. Classical risk assessment methods cannot adequately describe such black box systems. Therefore, new methods for a dynamic risk assessment of such AI-controlled…

机器人学 · 计算机科学 2024-01-26 Philipp Grimmeisen , Friedrich Sautter , Andrey Morozov

In recent years, agentic artificial intelligence (AI) systems are becoming increasingly widespread. These systems allow agents to use various tools, such as web browsers, compilers, and more. However, despite their popularity, agentic AI…

Artificial general intelligence (AGI) does not yet exist, but given the pace of technological development in artificial intelligence, it is projected to reach human-level intelligence within roughly the next two decades. After that, many…

计算机与社会 · 计算机科学 2023-11-16 David R. Mandel

Artificial Intelligence (AI) systems introduce unprecedented privacy challenges as they process increasingly sensitive data. Traditional privacy frameworks prove inadequate for AI technologies due to unique characteristics such as…

密码学与安全 · 计算机科学 2025-10-06 Grace Billiris , Asif Gill , Madhushi Bandara

AI advancements have been significantly driven by a combination of foundation models and curiosity-driven learning aimed at increasing capability and adaptability. Within this landscape, open-endedness, where AI agents autonomously and…

人工智能 · 计算机科学 2026-05-06 Ivaxi Sheth , Jan Wehner , Sahar Abdelnabi , Ruta Binkyte , Mario Fritz

Safety and responsibility evaluations of advanced AI models are a critical but developing field of research and practice. In the development of Google DeepMind's advanced AI models, we innovated on and applied a broad set of approaches to…

Scientific research organizations that are developing and deploying Artificial Intelligence (AI) systems are at the intersection of technological progress and ethical considerations. The push for Responsible AI (RAI) in such institutions…

人工智能 · 计算机科学 2023-12-18 Muneera Bano , Didar Zowghi , Pip Shea , Georgina Ibarra

As stories of human-AI interactions continue to be highlighted in the news and research platforms, the challenges are becoming more pronounced, including potential risks of overreliance, cognitive offloading, social and emotional…

人机交互 · 计算机科学 2025-10-22 Celeste Riley , Omar Al-Refai , Yadira Colunga Reyes , Eman Hammad

The rapid advancement of artificial intelligence (AI) has brought significant challenges to the education and workforce skills required to take advantage of AI for human-AI collaboration in the workplace. As AI continues to reshape…

人工智能 · 计算机科学 2024-05-31 Margarida Romero

This paper critically examines the evolving ethical and regulatory challenges posed by the integration of artificial intelligence (AI) in cybersecurity. We trace the historical development of AI regulation, highlighting major milestones…

密码学与安全 · 计算机科学 2025-01-22 Vikram Kulothungan