中文
相关论文

相关论文: A Computational Model of Commonsense Moral Decisio…

200 篇论文

Computational preference elicitation methods are tools used to learn people's preferences quantitatively in a given context. Recent works on preference elicitation advocate for active learning as an efficient method to iteratively construct…

人机交互 · 计算机科学 2024-07-29 Vijay Keswani , Vincent Conitzer , Hoda Heidari , Jana Schaich Borg , Walter Sinnott-Armstrong

Growing concerns about safety and alignment of AI systems highlight the importance of embedding moral capabilities in artificial agents: a promising solution is the use of learning from experience, i.e., Reinforcement Learning. In…

多智能体系统 · 计算机科学 2026-02-11 Elizaveta Tennant , Stephen Hailes , Mirco Musolesi

Decisions by humans depend on their estimations given some uncertain sensory data. These decisions can also be influenced by the behavior of others. Here we present a mathematical model to quantify this influence, inviting a further study…

物理与社会 · 物理学 2012-09-25 Gabriel Madirolas , Alfonso Perez-Escudero , Gonzalo G. de Polavieja

Traditionally, the way one evaluates the performance of an Artificial Intelligence (AI) system is via a comparison to human performance in specific tasks, treating humans as a reference for high-level cognition. However, these comparisons…

人工智能 · 计算机科学 2019-11-25 Camilo M. Signorelli , Xerxes D. Arsiwalla

Practical uses of Artificial Intelligence (AI) in the real world have demonstrated the importance of embedding moral choices into intelligent agents. They have also highlighted that defining top-down ethical constraints on AI according to…

多智能体系统 · 计算机科学 2023-08-31 Elizaveta Tennant , Stephen Hailes , Mirco Musolesi

Research in Responsible AI has developed a range of principles and practices to ensure that machine learning systems are used in a manner that is ethical and aligned with human values. However, a critical yet often neglected aspect of…

计算机与社会 · 计算机科学 2024-08-21 Neha R. Gupta , Jessica Hullman , Hari Subramonyam

Ethical thought experiments such as the trolley dilemma have been investigated extensively in the past, showing that humans act in a utilitarian way, trying to cause as little overall damage as possible. These trolley dilemmas have gained…

Artificial Moral Agents (AMA's) is a field in computer science with the purpose of creating autonomous machines that can make moral decisions akin to how humans do. Researchers have proposed theoretical means of creating such machines,…

计算机与社会 · 计算机科学 2020-03-03 George Rautenbach , C. Maria Keet

Although virtue ethics has repeatedly been proposed as a suitable framework for the development of artificial moral agents (AMAs), it has been proven difficult to approach from a computational perspective. In this work, we present the first…

人工智能 · 计算机科学 2022-10-07 Jakob Stenseke

Policy and guideline proposals for ethical artificial-intelligence research have proliferated in recent years. These are supposed to guide the socially-responsible development of AI for the common good. However, there typically exist…

计算机与社会 · 计算机科学 2021-01-20 Travis LaCroix , Aydin Mohseni

Machine Ethics decisions should consider the implications of uncertainty over decisions. Decisions should be made over sequences of actions to reach preferable outcomes long term. The evaluation of outcomes, however, may invoke one or more…

人工智能 · 计算机科学 2025-05-08 Simon Kolker , Louise A. Dennis , Ramon Fraga Pereira , Mengwei Xu

As AI systems become an increasing part of people's everyday lives, it becomes ever more important that they understand people's ethical norms. Motivated by descriptive ethics, a field of study that focuses on people's descriptive judgments…

计算与语言 · 计算机科学 2021-03-25 Nicholas Lourie , Ronan Le Bras , Yejin Choi

While various traditions under the 'virtue ethics' umbrella have been studied extensively and advocated by ethicists, it has not been clear that there exists a version of virtue ethics rigorous enough to be a target for machine ethics…

人工智能 · 计算机科学 2019-01-01 Naveen Sundar Govindarajulu , Selmer Bringsjord , Rikhiya Ghosh

The conceptual framework proposed in this paper centers on the development of a deliberative moral reasoning system - one designed to process complex moral situations by generating, filtering, and weighing normative arguments drawn from…

计算机与社会 · 计算机科学 2025-08-13 David-Doron Yaacov

As large language models (LLMs) increasingly participate in tasks with ethical and societal stakes, a critical question arises: do they exhibit an emergent "moral mind" - a consistent structure of moral preferences guiding their decisions -…

计算机与社会 · 计算机科学 2025-04-28 Avner Seror

AI and humans bring complementary skills to group deliberations. Modeling this group decision making is especially challenging when the deliberations include an element of risk and an exploration-exploitation process of appraising the…

人机交互 · 计算机科学 2022-01-11 Wei Ye , Francesco Bullo , Noah Friedkin , Ambuj K Singh

In the diverse array of work investigating the nature of human values from psychology, philosophy and social sciences, there is a clear consensus that values guide behaviour. More recently, a recognition that values provide a means to…

人工智能 · 计算机科学 2026-02-09 Nardine Osman , Mark d'Inverno

We present what we call the Interpretation Problem, whereby any rule in symbolic form is open to infinite interpretation in ways that we might disapprove of and argue that any attempt to build morality into machines is subject to it. We…

人工智能 · 计算机科学 2023-02-08 Cosmin Badea , Gregory Artus

Value alignment is the task of creating autonomous systems whose values align with those of humans. Past work has shown that stories are a potentially rich source of information on human values; however, past work has been limited to…

计算与语言 · 计算机科学 2022-12-13 Md Sultan Al Nahian , Spencer Frazier , Brent Harrison , Mark Riedl

Human moral judgment is context-dependent and modulated by interpersonal relationships. As large language models (LLMs) increasingly function as decision-support systems, determining whether they encode these social nuances is critical. We…

计算与语言 · 计算机科学 2026-04-24 Jiseon Kim , Jea Kwon , Luiz Felipe Vecchietti , Wenchao Dong , Jaehong Kim , Meeyoung Cha