English
Related papers

Related papers: What Would Jiminy Cricket Do? Towards Agents That …

200 papers

Using the example of the film 2001: A Space Odyssey, this chapter illustrates the challenges posed by an AI capable of making decisions that go against human interests. But are human decisions always rational and ethical? In reality, the…

Computers and Society · Computer Science 2025-12-05 Charlotte Jacquemot

Due to their unique persuasive power, language-capable robots must be able to both act in line with human moral norms and clearly and appropriately communicate those norms. These requirements are complicated by the possibility that humans…

Robotics · Computer Science 2021-04-15 Nichole D. Starr , Bertram Malle , Tom Williams

As artificial intelligence rapidly transforms society, developers and policymakers struggle to anticipate which applications will face public moral resistance. We propose that these judgments are not idiosyncratic but systematic and…

Computers and Society · Computer Science 2025-10-08 Kimmo Eriksson , Simon Karlsson , Irina Vartanova , Pontus Strimling

Reinforcement learning algorithms can train agents that solve problems in complex, interesting environments. Normally, the complexity of the trained agent is closely related to the complexity of the environment. This suggests that a highly…

Artificial Intelligence · Computer Science 2018-03-16 Trapit Bansal , Jakub Pachocki , Szymon Sidor , Ilya Sutskever , Igor Mordatch

How do people build up trust with artificial agents? Here, we study a key component of interpersonal trust: people's ability to evaluate the competence of another agent across repeated interactions. Prior work has largely focused on…

Human-Computer Interaction · Computer Science 2022-05-25 Erik Brockbank , Haoliang Wang , Justin Yang , Suvir Mirchandani , Erdem Bıyık , Dorsa Sadigh , Judith E. Fan

In this paper, we analyze the performance of an agent developed according to a well-accepted appraisal theory of human emotion with respect to how it modulates play in the context of a social dilemma. We ask if the agent will be capable of…

Artificial Intelligence · Computer Science 2021-07-19 Moojan Ghafurian , Neil Budnarain , Jesse Hoey

It is clear that one of the primary tools we can use to mitigate the potential risk from a misbehaving AI system is the ability to turn the system off. As the capabilities of AI systems improve, it is important to ensure that such systems…

Artificial Intelligence · Computer Science 2017-06-19 Dylan Hadfield-Menell , Anca Dragan , Pieter Abbeel , Stuart Russell

Recent advancements in deep reinforcement learning have brought forth an impressive display of highly skilled artificial agents capable of complex intelligent behavior. In video games, these artificial agents are increasingly deployed as…

Machine Learning · Statistics 2022-03-14 Ian Colbert , Mehdi Saeedi

Adam Smith developed a version of moral philosophy where better decisions are made by interrogating an impartial spectator within us. We discuss the possibility of using an external non-human-based substitute tool that would augment our…

General Economics · Economics 2023-05-22 Nikodem Tomczak

The tragedy of the commons illustrates a fundamental social dilemma where individual rational actions lead to collectively undesired outcomes, threatening the sustainability of shared resources. Strategies to escape this dilemma, however,…

Computer Science and Game Theory · Computer Science 2026-03-26 Arend Hintze , Christoph Adami

Artificial intelligence (AI) is a digital technology that will be of major importance for the development of humanity in the near future. AI has raised fundamental questions about what we should do with such systems, what the systems…

Computers and Society · Computer Science 2025-08-26 Vincent C. Müller

Dynamic game theory is an increasingly popular tool for modeling multi-agent, e.g. human-robot, interactions. Game-theoretic models presume that each agent wishes to minimize a private cost function that depends on others' actions. These…

Robotics · Computer Science 2025-10-17 Cade Armstrong , Ryan Park , Xinjie Liu , Kushagra Gupta , David Fridovich-Keil

AI-based systems can increasingly perform work tasks autonomously. In safety-critical tasks, human oversight of these systems is required to mitigate risks and to ensure responsibility in case something goes wrong. Since people often…

Human-Computer Interaction · Computer Science 2026-02-12 Cedric Faas , Richard Uth , Sarah Sterz , Markus Langer , Anna Maria Feit

Autonomous intelligent agents are playing increasingly important roles in our lives. They contain information about us and start to perform tasks on our behalves. Chatbots are an example of such agents that need to engage in a complex…

Artificial Intelligence · Computer Science 2019-09-19 Abeer Dyoub , Stefania Costantini , Francesca A. Lisi

In this essay, I argue that explicit ethical machines, whose moral principles are inferred through a bottom-up approach, are unable to replicate human-like moral reasoning and cannot be considered moral agents. By utilizing Alan Turing's…

Computers and Society · Computer Science 2024-07-25 Massimo Passamonti

As AI systems become an increasing part of people's everyday lives, it becomes ever more important that they understand people's ethical norms. Motivated by descriptive ethics, a field of study that focuses on people's descriptive judgments…

Computation and Language · Computer Science 2021-03-25 Nicholas Lourie , Ronan Le Bras , Yejin Choi

As artificial intelligence (AI) systems become increasingly ubiquitous, the topic of AI governance for ethical decision-making by AI has captured public imagination. Within the AI research community, this topic remains less familiar to many…

Artificial Intelligence · Computer Science 2018-12-10 Han Yu , Zhiqi Shen , Chunyan Miao , Cyril Leung , Victor R. Lesser , Qiang Yang

Do We Need Role Models? How do Role Models Shape Collective Morality? To explore the questions, we build a multi-agent simulation powered by a Large Language Model, where agents with diverse intrinsic drives, ranging from cooperative to…

Multiagent Systems · Computer Science 2026-03-17 Junjie Liao , Huacong Tang , Zhou Ziheng , Yizhou Wang , Fangwei Zhong

Road vehicle travel at a reasonable speed involves some risk, even when using computer-controlled driving with failure-free hardware and perfect sensing. A fully-automated vehicle must continuously decide how to allocate this risk without a…

Computers and Society · Computer Science 2020-10-30 Noah J. Goodall

This chapter explores moral responsibility for civilian harms by human-artificial intelligence (AI) teams. Although militaries may have some bad apples responsible for war crimes and some mad apples unable to be responsible for their…

Computers and Society · Computer Science 2023-09-07 Susannah Kate Devitt