中文
相关论文

相关论文: Deliberative Technology for Alignment

200 篇论文

It's widely expected that humanity will someday create AI systems vastly more intelligent than us, leading to the unsolved alignment problem of "how to control superintelligence." However, this commonly expressed problem is not only…

人工智能 · 计算机科学 2024-12-02 James M. Mazzu

Large language models (LLMs) are increasingly deployed in collaborative tasks involving multiple agents, forming an "AI agent society: where agents interact and influence one another. Whether such groups can spontaneously coordinate on…

物理与社会 · 物理学 2025-04-16 Giordano De Marzo , Claudio Castellano , David Garcia

Recent progress in artificial intelligence (AI) has renewed interest in building systems that learn and think like people. Many advances have come from using deep neural networks trained end-to-end in tasks such as object recognition, video…

人工智能 · 计算机科学 2016-11-03 Brenden M. Lake , Tomer D. Ullman , Joshua B. Tenenbaum , Samuel J. Gershman

Artificial intelligence (AI), although not able to currently capture the many complexities of humans, are slowly adapting to have certain capabilities of humans, many of which can revolutionize our world. AI systems, such as ChatGPT and…

计算机与社会 · 计算机科学 2024-03-26 Jay Nemec

Artificial intelligence (AI) systems, such as machine learning algorithms, have allowed scientists, marketers and governments to shed light on correlations that remained invisible until now. Beforehand, the dots that we had to connect in…

计算机与社会 · 计算机科学 2022-02-08 Remy Demichelis

The abolitionist community faces challenges from both the carceral state and oppressive technologies which, by empowering the ruling class who have the resources to develop artificial intelligence (AI), serve to entrench societal inequities…

人机交互 · 计算机科学 2025-10-09 Carolyn Wang , Avriel Epps , Taylor Ferrari , Ra Ames

AI-enabled capabilities are reaching the requisite level of maturity to be deployed in the real world, yet do not always make correct or safe decisions. One way of addressing these concerns is to leverage AI control systems alongside and in…

机器学习 · 计算机科学 2024-10-10 Walt Woods , Alexander Grushin , Simon Khan , Alvaro Velasquez

Explainability and comprehensibility of AI are important requirements for intelligent systems deployed in real-world domains. Users want and frequently need to understand how decisions impacting them are made. Similarly it is important to…

计算机与社会 · 计算机科学 2019-07-10 Roman V. Yampolskiy

Artificial intelligence has become a part of the provision of governmental services, from making decisions about benefits to issuing fines for parking violations. However, AI systems rarely live up to the promise of neutral optimisation,…

人工智能 · 计算机科学 2025-10-10 Dave Murray-Rust , Kars Alfrink , Cristina Zaga

Emerging technologies, such as big data, Internet of things, cloud computing, mobile Internet, and robotics, breed and expedite new applications and fields. In the mean while, the long-term prosperity and happiness of human race demands…

计算机与社会 · 计算机科学 2015-12-22 Zhe Chen

Artificial intelligence (AI) governance is the body of standards and practices used to ensure that AI systems are deployed responsibly. Current AI governance approaches consist mainly of manual review and documentation processes. While such…

计算机与社会 · 计算机科学 2023-02-17 Sean McGregor , Jesse Hostetler

As AI is increasingly being adopted into application solutions, the challenge of supporting interaction with humans is becoming more apparent. Partly this is to support integrated working styles, in which humans and intelligent systems…

人工智能 · 计算机科学 2017-10-02 Maria Fox , Derek Long , Daniele Magazzeni

Technology is increasingly shaping our social structures and is becoming a driving force in altering human biology. Besides, human activities already proved to have a significant impact on the Earth system which in turn generates complex…

计算机与社会 · 计算机科学 2018-02-15 Eugenio Maria Battaglia , Jie Mei , Guillaume Dumas

The Aiming for AI Interoperability report investigates the ongoing challenge of achieving regulatory and technical AI interoperability as national and global AI governance efforts are proliferating. Here, technical interoperability is the…

计算机与社会 · 计算机科学 2026-03-24 Benjamin Faveri , Craig Shank , Richard Whitt , Phillip Dawson

AI systems have the potential to improve decision-making, but decision makers face the risk that the AI may be misaligned with their objectives. We study this problem in the context of a treatment decision, where a designer decides which…

理论经济学 · 经济学 2025-09-19 Drew Fudenberg , Annie Liang

Artificial intelligence promises to revolutionise medicine, yet its impact remains limited because of the pervasive translational gap. We posit that the prevailing technology-centric approaches underpin this challenge, rendering such…

人机交互 · 计算机科学 2025-06-06 Kacper Sokol , James Fackler , Julia E Vogt

Most governance frameworks assume that rules can be defined in advance, systems can be engineered to comply, and accountability can be applied after outcomes occur. This model worked when machines replaced physical labor or accelerated…

人工智能 · 计算机科学 2026-01-30 Wolfgang Rohde

Current AI systems lack several important human capabilities, such as adaptability, generalizability, self-control, consistency, common sense, and causal reasoning. We believe that existing cognitive theories of human decision making, such…

Artificial Intelligence (AI) systems are increasingly placed in positions where their decisions have real consequences, e.g., moderating online spaces, conducting research, and advising on policy. Ensuring they operate in a safe and…

The use of artificial intelligence (AI) in the public sector is best understood as a continuation and intensification of long standing rationalization and bureaucratization processes. Drawing on Weber, we take the core of these processes to…

人工智能 · 计算机科学 2024-07-09 Jakob Mokander , Ralph Schroeder