中文
相关论文

相关论文: Learning the Value Systems of Societies from Prefe…

200 篇论文

Recent calls for pluralistic alignment emphasize that AI systems should address the diverse needs of all people. Yet, efforts in this space often require sorting people into fixed buckets of pre-specified diversity-defining dimensions…

计算与语言 · 计算机科学 2025-06-03 Liwei Jiang , Taylor Sorensen , Sydney Levine , Yejin Choi

The conceptual framework proposed in this paper centers on the development of a deliberative moral reasoning system - one designed to process complex moral situations by generating, filtering, and weighing normative arguments drawn from…

计算机与社会 · 计算机科学 2025-08-13 David-Doron Yaacov

Building trust in AI-based systems is deemed critical for their adoption and appropriate use. Recent research has thus attempted to evaluate how various attributes of these systems affect user trust. However, limitations regarding the…

人机交互 · 计算机科学 2022-04-29 Michaela Benk , Suzanne Tolmeijer , Florian von Wangenheim , Andrea Ferrario

Recently, there have been increasing calls for computer science curricula to complement existing technical training with topics related to Fairness, Accountability, Transparency, and Ethics. In this paper, we present Value Card, an…

计算机与社会 · 计算机科学 2023-01-11 Hong Shen , Wesley Hanwen Deng , Aditi Chattopadhyay , Zhiwei Steven Wu , Xu Wang , Haiyi Zhu

We identify "values" as actions that classifiers take that speak to open questions of significant social concern. Investigating a classifier's values builds on studies of social bias that uncover how classifiers participate in social…

计算机与社会 · 计算机科学 2024-02-08 Will Penman , Joshua Babu , Abhinaya Raghunathan

Conversational agents (CAs) based on generative artificial intelligence frequently face challenges ensuring ethical interactions that align with human values. Current value alignment efforts largely rely on top-down approaches, such as…

计算机与社会 · 计算机科学 2025-07-30 Lenart Motnikar , Katharina Baum , Alexander Kagan , Sarah Spiekermann-Hoff

Scholars investigating ethical AI, especially in high stakes settings like child welfare, have arguably been seeking ways to embed notions of justice into the design of these critical technologies. These efforts often operationalize justice…

计算机与社会 · 计算机科学 2025-11-13 Maria Y. Rodriguez , Seventy Hall , Pranav Sankhe , Melanie Sage , Winnie Chen , Atri Rudra , Kenny Joseph

With the rapid development of artificial intelligence (AI), ethical issues surrounding AI have attracted increasing attention. In particular, autonomous vehicles may face moral dilemmas in accident scenarios, such as staying the course…

密码学与安全 · 计算机科学 2019-09-30 Teng Wang , Jun Zhao , Han Yu , Jinyan Liu , Xinyu Yang , Xuebin Ren , Shuyu Shi

Observers can glean information from others' emotional expressions through the act of drawing inferences from another individual's emotional expressions. It is important for socially aware artificial systems to be capable of doing that as…

多智能体系统 · 计算机科学 2023-11-14 Jieting Luo , Mehdi Dastani , Thomas Studer , Beishui Liao

The emergence and growth of research on issues of ethics in AI, and in particular algorithmic fairness, has roots in an essential observation that structural inequalities in society are reflected in the data used to train predictive models…

计算机与社会 · 计算机科学 2020-02-28 Caitlin Kuhlman , Latifa Jackson , Rumi Chunara

The rise of artificial intelligence (A.I.) based systems is already offering substantial benefits to the society as a whole. However, these systems may also enclose potential conflicts and unintended consequences. Notably, people will tend…

计算机与社会 · 计算机科学 2020-12-23 Pedro Fernandes , Francisco C. Santos , Manuel Lopes

A crucial consideration when developing and deploying Large Language Models (LLMs) is the human values to which these models are aligned. In the constitutional framework of alignment models are aligned to a set of principles (the…

机器学习 · 计算机科学 2026-01-27 Henry Bell , Lara Neubauer da Costa Schertel , Bochu Ding , Brandon Fain

The AI landscape demands a broad set of legal, ethical, and societal considerations to be accounted for in order to develop ethical AI (eAI) solutions which sustain human values and rights. Currently, a variety of guidelines and a handful…

计算机与社会 · 计算机科学 2021-12-03 Anna Felländer , Jonathan Rebane , Stefan Larsson , Mattias Wiggberg , Fredrik Heintz

In high-stakes AI-supported decisions, considerations are not purely technical but involve moral judgments about fairness, responsibility, and harm. While prior research has focused mainly on functional or behavioral alignment, this paper…

人机交互 · 计算机科学 2026-04-17 Christiane Ernst , Luis Gutmann , Domenique Zipperling , Kathrin Figl , Niklas Kühl

How should we decide which fairness criteria or definitions to adopt in machine learning systems? To answer this question, we must study the fairness preferences of actual users of machine learning systems. Stringent parity constraints on…

人工智能 · 计算机科学 2020-12-09 Angie Peng , Jeff Naecker , Ben Hutchinson , Andrew Smart , Nyalleng Moorosi

Artificial intelligence is humanity's most promising technology because of the remarkable capabilities offered by foundation models. Yet, the same technology brings confusion and consternation: foundation models are poorly understood and…

人工智能 · 计算机科学 2025-07-01 Rishi Bommasani

Solving the AI alignment problem requires having clear, defensible values towards which AI systems can align. Currently, targets for alignment remain underspecified and do not seem to be built from a philosophically robust structure. We…

计算机与社会 · 计算机科学 2023-11-29 Betty Li Hou , Brian Patrick Green

Nudging is a behavioral strategy aimed at influencing people's thoughts and actions. Nudging techniques can be found in many situations in our daily lives, and these nudging techniques can targeted at human fast and unconscious thinking,…

Value alignment is the task of creating autonomous systems whose values align with those of humans. Past work has shown that stories are a potentially rich source of information on human values; however, past work has been limited to…

计算与语言 · 计算机科学 2022-12-13 Md Sultan Al Nahian , Spencer Frazier , Brent Harrison , Mark Riedl

Despite the prevalence of voting systems in the real world there is no consensus among researchers of how people vote strategically, even in simple voting settings. This paper addresses this gap by comparing different approaches that have…

计算机科学与博弈论 · 计算机科学 2019-09-24 Roy Fairstein , Adam Lauz , Kobi Gal , Reshef Meir