中文
相关论文

相关论文: Shutdown Safety Valves for Advanced AI

200 篇论文

AI evaluations are an important component of the AI governance toolkit, underlying current approaches to safety cases for preventing catastrophic risks. Our paper examines what these evaluations can and cannot tell us. Evaluations can…

计算机与社会 · 计算机科学 2024-12-13 Peter Barnett , Lisa Thiergart

Artificial intelligence (AI) was initially developed as an implicit moral agent to solve simple and clearly defined tasks where all options are predictable. However, it is now part of our daily life powering cell phones, cameras, watches,…

计算机与社会 · 计算机科学 2020-02-11 Mohamed Akrout , Robert Steinbauer

Autonomous vehicles are a growing technology that aims to enhance safety, accessibility, efficiency, and convenience through autonomous maneuvers ranging from lane change to overtaking. Overtaking is one of the most challenging maneuvers…

机器人学 · 计算机科学 2024-10-28 Ehsan Malayjerdi , Gokhan Alcan , Eshagh Kargar , Hatem Darweesh , Raivo Sell , Ville Kyrki

AI-based systems have been used widely across various industries for different decisions ranging from operational decisions to tactical and strategic ones in low- and high-stakes contexts. Gradually the weaknesses and issues of these…

人机交互 · 计算机科学 2022-01-13 Morteza Saberi

Autonomous Vehicles (AV) are expected to bring considerable benefits to society, such as traffic optimization and accidents reduction. They rely heavily on advances in many Artificial Intelligence (AI) approaches and techniques. However,…

One of the current AI issues depicted in popular culture is the fear of conscious super AIs that try to take control over humanity. And as computational power goes upwards and that turns more and more into a reality, understanding…

神经元与认知 · 定量生物学 2023-05-22 Daniel Lopes

Artificial intelligence (AI), which enables machines to learn to perform a task by training on diverse datasets, is one of the most revolutionary developments in scientific history. Although AI and especially deep learning is relatively…

计算机与社会 · 计算机科学 2024-02-27 Nathan Wang , Paul Tonko , Nikil Ragav , Michael Chungyoun , Jonathan Plucker

Artificial intelligence has become a part of the provision of governmental services, from making decisions about benefits to issuing fines for parking violations. However, AI systems rarely live up to the promise of neutral optimisation,…

人工智能 · 计算机科学 2025-10-10 Dave Murray-Rust , Kars Alfrink , Cristina Zaga

Calls for transparency in AI systems are growing in number and urgency from diverse stakeholders ranging from regulators to researchers to users (with a comparative absence of companies developing AI). Notions of transparency for AI abound,…

密码学与安全 · 计算机科学 2025-02-03 Peter Hall , Olivia Mundahl , Sunoo Park

Mental health disorders create profound personal and societal burdens, yet conventional diagnostics are resource-intensive and limit accessibility. Advances in artificial intelligence, particularly natural language processing and multimodal…

计算与语言 · 计算机科学 2025-08-26 Aishik Mandal , Tanmoy Chakraborty , Iryna Gurevych

Ensuring safety for human-interactive robotics is important due to the potential for human injury. The key challenge is defining safety in a way that accounts for the complex range of human behaviors without modeling the human as an…

机器人学 · 计算机科学 2021-10-12 Jeevana Priya Inala , Yecheng Jason Ma , Osbert Bastani , Xin Zhang , Armando Solar-Lezama

The history of AI has included several "waves" of ideas. The first wave, from the mid-1950s to the 1980s, focused on logic and symbolic hand-encoded representations of knowledge, the foundations of so-called "expert systems". The second…

计算机与社会 · 计算机科学 2020-12-14 Odest Chadwicke Jenkins , Daniel Lopresti , Melanie Mitchell

General intelligence, the ability to solve arbitrary solvable problems, is supposed by many to be artificially constructible. Narrow intelligence, the ability to solve a given particularly difficult problem, has seen impressive recent…

人工智能 · 计算机科学 2020-07-22 Michael K Cohen , Badri Vellambi , Marcus Hutter

The rapid advancement of artificial intelligence (AI) systems suggests that artificial general intelligence (AGI) systems may soon arrive. Many researchers are concerned that AIs and AGIs will harm humans via intentional misuse (AI-misuse)…

人工智能 · 计算机科学 2023-05-31 Catalin Mitelut , Ben Smith , Peter Vamplew

In the current era, people and society have grown increasingly reliant on artificial intelligence (AI) technologies. AI has the potential to drive us towards a future in which all of humanity flourishes. It also comes with substantial risks…

计算机与社会 · 计算机科学 2021-08-24 Lu Cheng , Kush R. Varshney , Huan Liu

With the advent of the digital era, every day-to-day task is automated due to technological advances. However, technology has yet to provide people with enough tools and safeguards. As the internet connects more-and-more devices around the…

密码学与安全 · 计算机科学 2022-09-28 Abhilash Chakraborty , Anupam Biswas , Ajoy Kumar Khan

Is the output of generative AI entitled to First Amendment protection? We're inclined to say yes. Even though current AI programs are of course not people and do not themselves have constitutional rights, their speech may potentially be…

计算机与社会 · 计算机科学 2023-08-21 Eugene Volokh , Mark Lemley , Peter Henderson

Safety has become the central value around which dominant AI governance efforts are being shaped. Recently, this culminated in the publication of the International AI Safety Report, written by 96 experts of which 30 nominated by the…

计算机与社会 · 计算机科学 2025-03-10 Roel Dobbe

Of primary importance in formulating a response to the increasing prevalence and power of artificial intelligence (AI) applications in society are questions of ontology. Questions such as: What "are" these systems? How are they to be…

计算机与社会 · 计算机科学 2019-03-11 Scott H. Hawley

Superhuman artificial general intelligence could be created this century and would likely be a significant source of existential risk. Delaying the creation of superintelligent AI (ASI) could decrease total existential risk by increasing…

计算机与社会 · 计算机科学 2022-09-13 Stephen McAleese
‹ 上一页 1 8 9 10 下一页 ›