中文
相关论文

相关论文: Mathematical foundations of moral preferences

200 篇论文

AI systems are often used to make or contribute to important decisions in a growing range of applications, including criminal justice, hiring, and medicine. Since these decisions impact human lives, it is important that the AI systems act…

One common assumption in game theory is that any player optimizes a utility function that takes into account only its own payoff. However, it has long been observed that in real life players may adopt an altruistic or even spiteful…

计算机科学与博弈论 · 计算机科学 2025-11-25 Michail Fasoulakis , Leonidas Bakopoulos , Charilaos Akasiadis , Georgios Chalkiadakis

As AI systems increasingly permeate everyday life, designers and developers face mounting pressure to balance innovation with ethical design choices. To date, the operationalisation of AI ethics has predominantly depended on frameworks that…

人机交互 · 计算机科学 2025-09-18 Benjamin J. Carroll , Jianlong Zhou , Paul F. Burke , Sabine Ammon

Indirect reciprocity is a mechanism for cooperation in social dilemma situations, in which an individual is motivated to help another to acquire a good reputation and receive help from others afterwards. Ingroup favoritism is another aspect…

物理与社会 · 物理学 2012-11-20 Mitsuhiro Nakamura , Naoki Masuda

Assistance games (also known as cooperative inverse reinforcement learning games) have been proposed as a model for beneficial AI, wherein a robotic agent must act on behalf of a human principal but is initially uncertain about the humans…

人工智能 · 计算机科学 2020-07-21 Arnaud Fickinger , Simon Zhuang , Dylan Hadfield-Menell , Stuart Russell

I study altruistic choices through the lens of a cognitively noisy decision-maker. I introduce a theoretical framework that demonstrates how increased cognitive noise can directionally affect altruistic decisions and put its implications to…

综合经济学 · 经济学 2025-01-03 Niklas M. Witzig

We elicit incomplete preferences over monetary gambles with subjective uncertainty. Subjects rank gambles, and these rankings are used to estimate preferences; payments are based on estimated preferences. About 40\% of subjects express…

综合经济学 · 经济学 2022-10-06 Kirby Nielsen , Luca Rigotti

A stylized experiment, the public goods game, has taught us the peculiar reproducible fact that humans tend to contribute more to shared resources than expected from economically rational assumptions. There have been two competing…

物理与社会 · 物理学 2024-12-03 Chen Shen , Zhixue He , Hao Guo , Shuyue Hu , Jun Tanimoto , Lei Shi , Petter Holme

As large language models (LLMs) increasingly integrate into our daily lives, it becomes crucial to understand their implicit biases and moral tendencies. To address this, we introduce a Moral Foundations LLM dataset (MFD-LLM) grounded in…

计算机与社会 · 计算机科学 2025-04-10 Monika Jotautaite , Mary Phuong , Chatrik Singh Mangat , Maria Angelica Martinez

Several rules for social choice are examined from a unifying point of view that looks at them as procedures for revising a system of degrees of belief in accordance with certain specified logical constraints. Belief is here a social…

人工智能 · 计算机科学 2015-05-06 Rosa Camps , Xavier Mora , Laia Saumell

As large language models (LLMs) increasingly participate in tasks with ethical and societal stakes, a critical question arises: do they exhibit an emergent "moral mind" - a consistent structure of moral preferences guiding their decisions -…

计算机与社会 · 计算机科学 2025-04-28 Avner Seror

Artificial intelligence (AI) is advancing at a pace that raises urgent questions about how to align machine decision-making with human moral values. This working paper investigates how leading AI systems prioritize moral outcomes and what…

人工智能 · 计算机科学 2025-09-15 Eoin O'Doherty , Nicole Weinrauch , Andrew Talone , Uri Klempner , Xiaoyuan Yi , Xing Xie , Yi Zeng

Recent work has constructed economic mechanisms that are both truthful and differentially private. In these mechanisms, privacy is treated separately from the truthfulness; it is not incorporated in players' utility functions (and doing so…

计算机科学与博弈论 · 计算机科学 2012-11-14 Yiling Chen , Stephen Chong , Ian A. Kash , Tal Moran , Salil Vadhan

A rapidly growing literature on lying in behavioral economics and psychology shows that individuals often do not lie even when lying maximizes their utility. In this work, we attempt to incorporate these findings into the theory of…

计算机科学与博弈论 · 计算机科学 2021-11-23 Shahar Dobzinski , Sigal Oren

The ethics of artificial intelligence (AI) systems has risen as an imminent concern across scholarly communities. This concern has propagated a great interest in algorithmic fairness. Large research agendas are now devoted to increasing…

计算机与社会 · 计算机科学 2023-12-21 Kimi Wenzel , Geoff Kaufman , Laura Dabbish

While people generally trust AI to make decisions in various aspects of their lives, concerns arise when AI is involved in decisions with significant moral implications. The absence of a precise mathematical framework for moral reasoning…

人工智能 · 计算机科学 2024-07-11 Vincent Conitzer

Prosociality is fundamental to human social life, and, accordingly, much research has attempted to explain human prosocial behavior. Capraro and Rand (Judgment and Decision Making, 13, 99-111, 2018) recently provided experimental evidence…

物理与社会 · 物理学 2018-06-18 Ben M. Tappin , Valerio Capraro

We introduce a new computational model of moral decision making, drawing on a recent theory of commonsense moral learning via social dynamics. Our model describes moral dilemmas as a utility function that computes trade-offs in values over…

人工智能 · 计算机科学 2018-01-16 Richard Kim , Max Kleiman-Weiner , Andres Abeliuk , Edmond Awad , Sohan Dsouza , Josh Tenenbaum , Iyad Rahwan

An ambitious goal for machine learning is to create agents that behave ethically: The capacity to abide by human moral norms would greatly expand the context in which autonomous agents could be practically and safely deployed, e.g. fully…

人工智能 · 计算机科学 2021-07-21 Adrien Ecoffet , Joel Lehman

Moral Foundations Theory proposes that individuals with conflicting political views base their behavior on different principles chosen from a small group of universal moral foundations. This study proposes using a set of widely accepted…

社会与信息网络 · 计算机科学 2024-07-01 Ruben Interian