中文
相关论文

相关论文: Superintelligence Safety: A Requirements Engineeri…

200 篇论文

What does Artificial Intelligence (AI) have to contribute to health care? And what should we be looking out for if we are worried about its risks? In this paper we offer a survey, and initial evaluation, of hopes and fears about the…

计算机与社会 · 计算机科学 2025-05-13 Robert Sparrow , Joshua Hatherley

Every AI system is deployed by a human organization. In high risk applications, the combined human plus AI system must function as a high-reliability organization in order to avoid catastrophic errors. This short note reviews the properties…

人工智能 · 计算机科学 2018-11-28 Thomas G. Dietterich

In AI, the existential risk denotes the hypothetical threat posed by an artificial system that would possess both the capability and the objective, either directly or indirectly, to eradicate humanity. This issue is gaining prominence in…

人工智能 · 计算机科学 2026-05-18 Rufin VanRullen

As AI systems become more capable, widely deployed, and increasingly autonomous in critical areas such as cybersecurity, biological research, and healthcare, ensuring their safety and alignment with human values is paramount. Machine…

As AI systems become increasingly powerful, the need for safe AI has become more pressing. Humans are an attractive model for AI safety: as the only known agents capable of general intelligence, they perform robustly even under conditions…

Robotics, automation, and related Artificial Intelligence (AI) systems have become pervasive bringing in concerns related to security, safety, accuracy, and trust. With growing dependency on physical robots that work in close proximity to…

密码学与安全 · 计算机科学 2023-02-17 Sudip Mittal , Jingdao Chen

AI scientists powered by large language models have demonstrated substantial promise in autonomously conducting experiments and facilitating scientific discoveries across various disciplines. While their capabilities are promising, these…

The rapid development of AI systems poses unprecedented risks, including loss of control, misuse, geopolitical instability, and concentration of power. To navigate these risks and avoid worst-case outcomes, governments may proactively…

人工智能 · 计算机科学 2025-07-15 Peter Barnett , Aaron Scher , David Abecassis

Requirements Engineering (RE) is the discipline for identifying, analyzing, as well as ensuring the implementation and delivery of user, technical, and societal requirements. Recently reported issues concerning the acceptance of Artificial…

人工智能 · 计算机科学 2023-02-22 Walid Maalej , Yen Dieu Pham , Larissa Chazette

In this paper we discuss how systems with Artificial Intelligence (AI) can undergo safety assessment. This is relevant, if AI is used in safety related applications. Taking a deeper look into AI models, we show, that many models of…

人工智能 · 计算机科学 2021-05-17 Jens Braband , Hendrik Schäbe

While artificial intelligence (AI) is advancing rapidly and mastering increasingly complex problems with astonishing performance, the safety assurance of such systems is a major concern. Particularly in the context of safety-critical,…

人工智能 · 计算机科学 2025-07-01 Lars Ullrich , Walter Zimmer , Ross Greer , Knut Graichen , Alois C. Knoll , Mohan Trivedi

Superhuman artificial general intelligence could be created this century and would likely be a significant source of existential risk. Delaying the creation of superintelligent AI (ASI) could decrease total existential risk by increasing…

计算机与社会 · 计算机科学 2022-09-13 Stephen McAleese

The recent embrace of machine learning (ML) in the development of autonomous weapons systems (AWS) creates serious risks to geopolitical stability and the free exchange of ideas in AI research. This topic has received comparatively little…

计算机与社会 · 计算机科学 2024-06-04 Riley Simmons-Edler , Ryan Badman , Shayne Longpre , Kanaka Rajan

Framed in positive terms, this report examines how technical AI research might be steered in a manner that is more attentive to humanity's long-term prospects for survival as a species. In negative terms, we ask what existential risks…

计算机与社会 · 计算机科学 2020-06-11 Andrew Critch , David Krueger

Success in the quest for artificial intelligence has the potential to bring unprecedented benefits to humanity, and it is therefore worthwhile to investigate how to maximize these benefits while avoiding potential pitfalls. This article…

人工智能 · 计算机科学 2016-02-15 Stuart Russell , Daniel Dewey , Max Tegmark

With the progressive digitalisation of a majority of services to communities and individuals, humankind is facing new challenges. While energy sources are rapidly dwindling and rigorous choices have to be made to ensure the sustainability…

计算机与社会 · 计算机科学 2020-02-17 Birgitta Dresp-Langley

In 1950, Alan Turing proposed an imitation game as the ultimate test of whether a machine was intelligent: could a machine imitate a human so well that its answers to questions indistinguishable from a human. Ever since, creating…

综合经济学 · 经济学 2022-01-13 Erik Brynjolfsson

Rapid advances in AI are beginning to reshape national security. Destabilizing AI developments could rupture the balance of power and raise the odds of great-power conflict, while widespread proliferation of capable AI hackers and…

计算机与社会 · 计算机科学 2025-04-16 Dan Hendrycks , Eric Schmidt , Alexandr Wang

Superintelligence is a hypothetical agent that possesses intelligence far surpassing that of the brightest and most gifted human minds. In light of recent advances in machine intelligence, a number of scientists, philosophers and…

计算机与社会 · 计算机科学 2021-01-12 Manuel Alfonseca , Manuel Cebrian , Antonio Fernandez Anta , Lorenzo Coviello , Andres Abeliuk , Iyad Rahwan

Human oversight requirements are a core component of the European AI Act and in AI governance. In this paper, we highlight key challenges in testing for compliance with these requirements. A central difficulty lies in balancing simple, but…

人机交互 · 计算机科学 2025-07-25 Markus Langer , Veronika Lazar , Kevin Baum