中文
相关论文

相关论文: Aligning AI With Shared Human Values

200 篇论文

Recent advances in Large Language Models (LLMs) have enabled human-like responses across various tasks, raising questions about their ethical decision-making capabilities and potential biases. This study systematically evaluates how nine…

计算机与社会 · 计算机科学 2025-11-03 Wentao Xu , Yile Yan , Yuqi Zhu

Recent research illustrates how AI can be developed and deployed in a manner detached from the concrete social context of application. By abstracting from the contexts of AI application, practitioners also disengage from the distinct…

人工智能 · 计算机科学 2025-09-29 Alexander Martin Mussgnug

Moral norms vary across cultures. A recent line of work suggests that English large language models contain human-like moral biases, but these studies typically do not examine moral variation in a diverse cultural setting. We investigate…

计算与语言 · 计算机科学 2023-06-06 Aida Ramezani , Yang Xu

Social alignment in AI systems aims to ensure that these models behave according to established societal values. However, unlike humans, who derive consensus on value judgments through social interaction, current language models (LMs) are…

计算与语言 · 计算机科学 2023-10-31 Ruibo Liu , Ruixin Yang , Chenyan Jia , Ge Zhang , Denny Zhou , Andrew M. Dai , Diyi Yang , Soroush Vosoughi

We present a general approach to automating ethical decisions, drawing on machine learning and computational social choice. In a nutshell, we propose to learn a model of societal preferences, and, when faced with a specific ethical dilemma…

Commonsense knowledge, a major constituent of artificial intelligence (AI), is primarily evaluated in practice by human-prescribed ground-truth labels. An important, albeit implicit, assumption of these labels is that they accurately…

人工智能 · 计算机科学 2026-01-23 Tuan Dung Nguyen , Duncan J. Watts , Mark E. Whiting

Humans can make moral inferences from multiple sources of input. In contrast, automated moral inference in artificial intelligence typically relies on language models with textual input. However, morality is conveyed through modalities…

计算机视觉与模式识别 · 计算机科学 2025-04-17 Warren Zhu , Aida Ramezani , Yang Xu

As AI systems become more advanced, ensuring their alignment with a diverse range of individuals and societal values becomes increasingly critical. But how can we capture fundamental human values and assess the degree to which AI systems…

人机交互 · 计算机科学 2025-11-05 Hua Shen , Tiffany Knearem , Reshmi Ghosh , Yu-Ju Yang , Nicholas Clark , Tanushree Mitra , Yun Huang

The transition of Artificial Intelligence (AI) from a lab-based science to live human contexts brings into sharp focus many historic, socio-cultural biases, inequalities, and moral dilemmas. Many questions that have been raised regarding…

计算机与社会 · 计算机科学 2024-06-19 Kaska Porayska-Pomsta , Wayne Holmes , Selena Nemorin

Does AI understand human values? While this remains an open philosophical question, we take a pragmatic stance by introducing VAPT, the Value-Alignment Perception Toolkit, for studying how LLMs reflect people's values and how people judge…

人机交互 · 计算机科学 2026-04-15 Bhada Yun , Renn Su , April Yi Wang

The rapid progress in Large Language Models (LLMs) could transform many fields, but their fast development creates significant challenges for oversight, ethical creation, and building user trust. This comprehensive review looks at key trust…

计算机与社会 · 计算机科学 2024-07-22 Md Meftahul Ferdaus , Mahdi Abdelguerfi , Elias Ioup , Kendall N. Niles , Ken Pathak , Steven Sloan

Ethical AI spans a gamut of considerations. Among these, the most popular ones, fairness and interpretability, have remained largely distinct in technical pursuits. We discuss and elucidate the differences between fairness and…

计算机与社会 · 计算机科学 2021-06-28 Deepak P , Sanil V , Joemon M. Jose

The recent surge in interest in ethics in artificial intelligence may leave many educators wondering how to address moral, ethical, and philosophical issues in their AI courses. As instructors we want to develop curriculum that not only…

人工智能 · 计算机科学 2017-01-27 Emanuelle Burton , Judy Goldsmith , Sven Koenig , Benjamin Kuipers , Nicholas Mattei , Toby Walsh

The past years have presented a surge in (AI) development, fueled by breakthroughs in deep learning, increased computational power, and substantial investments in the field. Given the generative capabilities of more recent AI systems, the…

AI agents are increasingly deployed and used to make automated decisions that affect our lives on a daily basis. It is imperative to ensure that these systems embed ethical principles and respect human values. We focus on how we can attest…

人工智能 · 计算机科学 2019-09-11 Xavier Ferrer Aran , Jose M. Such , Natalia Criado

In the age of big data, companies and governments are increasingly using algorithms to inform hiring decisions, employee management, policing, credit scoring, insurance pricing, and many more aspects of our lives. AI systems can help us…

计算机与社会 · 计算机科学 2020-08-25 Evgeni Aizenberg , Jeroen van den Hoven

The integration of Artificial Intelligence (AI) into construction project management (CPM) is accelerating, with Large Language Models (LLMs) emerging as accessible decision-support tools. This study aims to critically evaluate the ethical…

人工智能 · 计算机科学 2025-09-08 Somtochukwu Azie , Yiping Meng

The ethics of Machine Learning has become an unavoidable topic in the AI Community. The deployment of machine learning systems in multiple social contexts has resulted in a closer ethical scrutiny of the design, development, and application…

计算机与社会 · 计算机科学 2022-01-19 Miguel Sicart , Irina Shklovski , Mirabelle Jones

We find ourselves surrounded by a rapidly increasing number of autonomous and semi-autonomous systems. Two grand challenges arise from this development: Machine Ethics and Machine Explainability. Machine Ethics, on the one hand, is…

人工智能 · 计算机科学 2019-01-04 Kevin Baum , Holger Hermanns , Timo Speith

This paper elaborates on the concept of moral exercises as a means to help AI actors cultivate virtues that enable effective human oversight of AI systems. We explore the conceptual framework and significance of moral exercises, situating…

计算机与社会 · 计算机科学 2025-05-23 Silvia Crafa , Teresa Scantamburlo