English
Related papers

Related papers: Foundational Moral Values for AI Alignment

200 papers

Risk thresholds provide a measure of the level of risk exposure that a society or individual is willing to withstand, ultimately shaping how we determine the safety of technological systems. Against the backdrop of the Cold War, the first…

Computers and Society · Computer Science 2025-04-22 Heidy Khlaaf , Sarah Myers West

As large language models (LLMs) become increasingly integrated into society, their alignment with human morals is crucial. To better understand this alignment, we created a large corpus of human- and LLM-generated responses to various moral…

Human-Computer Interaction · Computer Science 2024-10-11 Basile Garcia , Crystal Qian , Stefano Palminteri

As a capability coming from computation, how does AI differ fundamentally from the capabilities delivered by rule-based software program? The paper examines the behavior of artificial intelligence (AI) from engineering points of view to…

Human-Computer Interaction · Computer Science 2025-11-19 Bifei Mao , Lanqing Hong

We describe cases where real recommender systems were modified in the service of various human values such as diversity, fairness, well-being, time well spent, and factual accuracy. From this we identify the current practice of values…

Information Retrieval · Computer Science 2021-07-26 Jonathan Stray , Ivan Vendrov , Jeremy Nixon , Steven Adler , Dylan Hadfield-Menell

To implement fair machine learning in a sustainable way, choosing the right fairness objective is key. Since fairness is a concept of justice which comes in various, sometimes conflicting definitions, this is not a trivial task though. The…

Artificial Intelligence · Computer Science 2021-05-04 Boris Ruf , Marcin Detyniecki

This work establishes a foundational framework for standardizing AI evaluation RCTs (sometimes called human uplift studies). Drawing on established experimental practices from disciplines with established RCT traditions, including software…

Apologies are a powerful tool used in human-human interactions to provide affective support, regulate social processes, and exchange information following a trust violation. The emerging field of AI apology investigates the use of apologies…

Computers and Society · Computer Science 2024-12-23 Hadassah Harland , Richard Dazeley , Hashini Senaratne , Peter Vamplew , Francisco Cruz , Bahareh Nakisa

This study is focused on the ethics of Artificial Intelligence and its application in the United States, the paper highlights the impact AI has in every sector of the US economy and multiple facets of the technological space and the…

Artificial Intelligence · Computer Science 2023-10-10 Esther Taiwo , Ahmed Akinsola , Edward Tella , Kolade Makinde , Mayowa Akinwande

As artificial intelligence (AI) improves, traditional alignment strategies may falter in the face of unpredictable self-improvement, hidden subgoals, and the sheer complexity of intelligent systems. Inspired by contemplative wisdom…

Artificial Intelligence · Computer Science 2025-08-19 Ruben Laukkonen , Fionn Inglis , Shamil Chandaria , Lars Sandved-Smith , Edmundo Lopez-Sola , Jakob Hohwy , Jonathan Gold , Adam Elwood

In the rapidly evolving domain of Artificial Intelligence (AI), the complex interaction between innovation and regulation has become an emerging focus of our society. Despite tremendous advancements in AI's capabilities to excel in specific…

Artificial Intelligence · Computer Science 2025-02-07 Nan Sun , Yuantian Miao , Hao Jiang , Ming Ding , Jun Zhang

This paper argues that Machine Learning (ML) algorithms must be educated. ML-trained algorithms moral decisions are ubiquitous in human society. Sometimes reverting the societal advances governments, NGOs and civil society have achieved…

Machine Learning · Computer Science 2023-05-23 Susana Perez Blazquez , Inas Hipolito

Our interdisciplinary study investigates how effectively U.S. laws confront the challenges posed by Generative AI to human values. Through an analysis of diverse hypothetical scenarios crafted during an expert workshop, we have identified…

Computers and Society · Computer Science 2023-09-06 Inyoung Cheong , Aylin Caliskan , Tadayoshi Kohno

Though intelligent agents are supposed to improve human experience (or make it more efficient), it is hard from a human perspective to grasp the ethical values which are explicitly or implicitly embedded in an agent behaviour. This is the…

Artificial Intelligence · Computer Science 2025-08-12 Gianluca Bontempi

Researchers, practitioners, and policymakers with an interest in AI ethics need more integrative approaches for studying and intervening in AI systems across many contexts and scales of activity. This paper presents AI value chains as an…

Computers and Society · Computer Science 2024-09-19 Blair Attard-Frost , David Gray Widder

We review practical challenges in building and deploying ethical AI at the scale of contemporary industrial and societal uses. Apart from the purely technical concerns that are the usual focus of academic research, the operational…

Machine Learning · Computer Science 2021-08-16 Jiahao Chen , Victor Storchan , Eren Kurshan

This article analyzes the impact of artificial intelligence (AI) on contemporary society and the importance of adopting an ethical approach to its development and implementation within organizations. It examines the technocritical…

Computers and Society · Computer Science 2024-05-06 Ernesto Giralt Hernández

This paper introduces the Flourishing AI Benchmark (FAI Benchmark), a novel evaluation framework that assesses AI alignment with human flourishing across seven dimensions: Character and Virtue, Close Social Relationships, Happiness and Life…

Text-based online counselling scales across geographical and stigma barriers, yet faces practitioner shortages, lacks non-verbal cues and suffers inconsistent quality assurance. Whilst artificial intelligence offers promising solutions, its…

Computers and Society · Computer Science 2026-01-15 Philipp Steigerwald , Jennifer Burghardt , Eric Rudolph , Jens Albrecht

An important step in the development of value alignment (VA) systems in AI is understanding how VA can reflect valid ethical principles. We propose that designers of VA systems incorporate ethics by utilizing a hybrid approach in which both…

Artificial Intelligence · Computer Science 2020-12-23 Tae Wan Kim , John Hooker , Thomas Donaldson