中文
相关论文

相关论文: Toward the Engineering of Virtuous Machines

200 篇论文

Artificial Intelligence systems are rapidly evolving, integrating extrinsic and intrinsic motivations. While these frameworks offer benefits, they risk misalignment at the algorithmic level while appearing superficially aligned with human…

人工智能 · 计算机科学 2024-11-08 Joshua T. S. Hewson

As we grant artificial intelligence increasing power and independence in contexts like healthcare, policing, and driving, AI faces moral dilemmas but lacks the tools to solve them. Warnings from regulators, philosophers, and computer…

人工智能 · 计算机科学 2022-07-22 Lavanya Singh

With the growing popularity of conversational agents based on large language models (LLMs), we need to ensure their behaviour is ethical and appropriate. Work in this area largely centres around the 'HHH' criteria: making outputs more…

计算与语言 · 计算机科学 2024-05-17 Lize Alberts , Geoff Keeling , Amanda McCroskery

Ethics and safety research in artificial intelligence is increasingly framed in terms of "alignment" with human values and interests. I argue that Turing's call for "fair play for machines" is an early and often overlooked contribution to…

计算机与社会 · 计算机科学 2018-12-07 Daniel Estrada

Generative AI enables automated, effective manipulation at scale. Despite the growing general ethical discussion around generative AI, the specific manipulation risks remain inadequately investigated. This article outlines essential…

计算机与社会 · 计算机科学 2025-03-10 Michael Klenk

This paper grounds ethics in evolutionary biology, viewing moral norms as adaptive mechanisms that render cooperation fitness-viable under selection pressure. Current alignment approaches add ethics post hoc, treating it as an external…

计算机与社会 · 计算机科学 2025-10-17 Dylan Waldner

What is agency, and why does it matter? In this work, we draw from the political science and philosophy literature and give two competing visions of what it means to be an (ethical) agent. The first view, which we term mechanistic, is…

计算机与社会 · 计算机科学 2024-04-23 Jessica Dai

In the age of algorithms, I focus on the question of how to ensure algorithms that will take over many of our familiar archival and library tasks, will behave according to human ethical norms that have evolved over many years. I start by…

人工智能 · 计算机科学 2018-01-08 Martijn van Otterlo

The concepts of blameworthiness and wrongness are of fundamental importance in human moral life. But to what extent are humans disposed to blame artificially intelligent agents, and to what extent will they judge their actions to be morally…

计算机与社会 · 计算机科学 2021-02-09 Michael T. Stuart , Markus Kneer

Several high-profile events, such as the mass testing of emotion recognition systems on vulnerable sub-populations and using question answering systems to make moral judgments, have highlighted how technology will often lead to more adverse…

人工智能 · 计算机科学 2022-03-22 Saif M. Mohammad

Although the problem of a critique of robotic behavior in near-unanimous agreement to human norms seems intractable, a starting point of such an ambition is a framework of the collection of knowledge a priori and experience a posteriori…

人工智能 · 计算机科学 2017-05-30 Christopher A. Tucker

Despite its successes, to date Artificial Intelligence (AI) is still characterized by a number of shortcomings with regards to different application domains and goals. These limitations are arguably both conceptual (e.g., related to…

We propose a conceptualization and implementation of AI ethics via the capability approach. We aim to show that conceptualizing AI ethics through the capability approach has two main advantages for AI ethics as a discipline. First, it helps…

计算机与社会 · 计算机科学 2025-02-07 Emanuele Ratti , Mark Graves

This paper explores the development of an ethical guardrail framework for AI systems, emphasizing the importance of customizable guardrails that align with diverse user values and underlying ethics. We address the challenges of AI ethics by…

计算机与社会 · 计算机科学 2024-11-25 Kristina Šekrst , Jeremy McHugh , Jonathan Rodriguez Cefalu

Across the technology industry, many companies have expressed their commitments to AI ethics and created dedicated roles responsible for translating high-level ethics principles into product. Yet it is unclear how effective this has been in…

人机交互 · 计算机科学 2024-09-12 Archana Ahlawat , Amy Winecoff , Jonathan Mayer

Widely considered a cornerstone of human morality, trust shapes many aspects of human social interactions. In this work, we present a theoretical analysis of the $\textit{trust game}$, the canonical task for studying trust in behavioral and…

人工智能 · 计算机科学 2023-12-21 Ardavan S. Nobandegani , Irina Rish , Thomas R. Shultz

Value learning is a crucial aspect of safe and ethical AI. This is primarily pursued by methods inferring human values from behaviour. However, humans care about much more than we are able to demonstrate through our actions. Consequently,…

人工智能 · 计算机科学 2025-05-28 Paul de Font-Reaulx

The widespread use of artificial intelligence (AI) in many domains has revealed numerous ethical issues from data and design to deployment. In response, countless broad principles and guidelines for ethical AI have been published, and…

人工智能 · 计算机科学 2021-09-21 Helen Bubinger , Jesse David Dinneen

Taking the stance that artificially conscious agents should be given human-like rights, in this paper we attempt to define consciousness, aggregate existing universal human rights, analyze robotic laws with roots in both reality and science…

计算机与社会 · 计算机科学 2020-11-16 Markian Hromiak

Research in Responsible AI has developed a range of principles and practices to ensure that machine learning systems are used in a manner that is ethical and aligned with human values. However, a critical yet often neglected aspect of…

计算机与社会 · 计算机科学 2024-08-21 Neha R. Gupta , Jessica Hullman , Hari Subramonyam