English
Related papers

Related papers: AI Safety, Alignment, and Ethics (AI SAE)

200 papers

The rapid integration of Artificial Intelligence (AI) in Higher Education (HE) is transforming personalized learning, administrative automation, and decision-making. However, this progress presents a duality, as AI adoption also introduces…

Computers and Society · Computer Science 2025-03-10 Prashant Mahajan

An important step in the development of value alignment (VA) systems in AI is understanding how VA can reflect valid ethical principles. We propose that designers of VA systems incorporate ethics by utilizing a hybrid approach in which both…

Artificial Intelligence · Computer Science 2020-12-23 Tae Wan Kim , John Hooker , Thomas Donaldson

Although AI has significant potential to transform society, there are serious concerns about its ability to behave and make decisions responsibly. Many ethical regulations, principles, and guidelines for responsible AI have been issued…

Artificial Intelligence · Computer Science 2022-09-21 Qinghua Lu , Liming Zhu , Xiwei Xu , Jon Whittle

Why should moral philosophers, moral psychologists, and machine ethicists care about computational complexity? Debates on whether artificial intelligence (AI) can or should be used to solve problems in ethical domains have mainly been…

Computational Complexity · Computer Science 2023-02-09 Jakob Stenseke

As Artificial Intelligence (AI) systems exert a growing influence on society, real-life incidents begin to underline the importance of AI Ethics. Though calls for more ethical AI systems have been voiced by scholars and the general public…

Computers and Society · Computer Science 2020-04-22 Ville Vakkuri , Kai-Kristian Kemell

We show how to assess a language model's knowledge of basic concepts of morality. We introduce the ETHICS dataset, a new benchmark that spans concepts in justice, well-being, duties, virtues, and commonsense morality. Models predict…

Computers and Society · Computer Science 2023-02-20 Dan Hendrycks , Collin Burns , Steven Basart , Andrew Critch , Jerry Li , Dawn Song , Jacob Steinhardt

Scientific research organizations that are developing and deploying Artificial Intelligence (AI) systems are at the intersection of technological progress and ethical considerations. The push for Responsible AI (RAI) in such institutions…

Artificial Intelligence · Computer Science 2023-12-18 Muneera Bano , Didar Zowghi , Pip Shea , Georgina Ibarra

A series of recent developments points towards auditing as a promising mechanism to bridge the gap between principles and practice in AI ethics. Building on ongoing discussions concerning ethics-based auditing, we offer three contributions.…

Computers and Society · Computer Science 2021-05-04 Jakob Mokander , Luciano Floridi

We consider two fundamental and related issues currently faced by Artificial Intelligence (AI) development: the lack of ethics and interpretability of AI decisions. Can interpretable AI decisions help to address ethics in AI? Using a…

Artificial Intelligence · Computer Science 2021-09-21 Jean-Marie John-Mathews

This paper addresses the question of how to align AI systems with human values and situates it within a wider body of thought regarding technology and value. Far from existing in a vacuum, there has long been an interest in the ability of…

Computers and Society · Computer Science 2021-01-19 Iason Gabriel , Vafa Ghazavi

As artificial intelligence (AI) reshapes industries and societies, ensuring its trustworthiness-through mitigating ethical risks like bias, opacity, and accountability deficits-remains a global challenge. International Organization for…

Computers and Society · Computer Science 2025-04-24 Sridharan Sankaran

To realize the potential benefits and mitigate potential risks of AI, it is necessary to develop a framework of governance that conforms to ethics and fundamental human values. Although several organizations have issued guidelines and…

Computers and Society · Computer Science 2023-07-14 Hyesun Choung , Prabu David , John S. Seberger

As AI systems increasingly permeate everyday life, designers and developers face mounting pressure to balance innovation with ethical design choices. To date, the operationalisation of AI ethics has predominantly depended on frameworks that…

Human-Computer Interaction · Computer Science 2025-09-18 Benjamin J. Carroll , Jianlong Zhou , Paul F. Burke , Sabine Ammon

This paper examines the challenges associated with achieving life-long superalignment in AI systems, particularly large language models (LLMs). Superalignment is a theoretical framework that aspires to ensure that superintelligent AI…

Computers and Society · Computer Science 2024-03-25 Gokul Puthumanaillam , Manav Vora , Pranay Thangeda , Melkior Ornik

As AI systems become more prevalent, concerns about their development, operation, and societal impact intensify. Establishing ethical, social, and safety standards amidst evolving AI capabilities poses significant challenges. Global…

Human-Computer Interaction · Computer Science 2025-03-24 Yutaka Matsubara , Akihisa Morikawa , Daichi Mizuguchi , Kiyoshi Fujiwara

Responsible AI is widely considered as one of the greatest scientific challenges of our time and is key to increase the adoption of AI. Recently, a number of AI ethics principles frameworks have been published. However, without further…

Artificial Intelligence · Computer Science 2023-09-29 Qinghua Lu , Liming Zhu , Xiwei Xu , Jon Whittle , Didar Zowghi , Aurelie Jacquet

Artificial intelligence (AI) trends vary significantly across global regions, shaping the trajectory of innovation, regulation, and societal impact. This variation influences how different regions approach AI development, balancing…

Computers and Society · Computer Science 2025-04-02 Vikram Kulothungan , Deepti Gupta

The neutrality thesis holds that technology cannot be laden with values. This long-standing view has faced critiques, but much of the argumentation against neutrality has focused on traditional, non-smart technologies like bridges and…

Artificial Intelligence · Computer Science 2024-08-23 Torben Swoboda , Lode Lauwaert

This chapter explores the symbiotic relationship between Artificial Intelligence (AI) and trust in networked systems, focusing on how these two elements reinforce each other in strategic cybersecurity contexts. AI's capabilities in data…

Artificial Intelligence · Computer Science 2024-11-21 Yunfei Ge , Quanyan Zhu

AI-driven digital ecosystems span diverse stakeholders including technology firms, regulators, accelerators and civil society, yet often lack cohesive ethical governance. This paper proposes a four-pillar framework (SCOR) to embed…

Computers and Society · Computer Science 2025-09-16 Mohammad Saleh Torkestani , Taha Mansouri