中文
相关论文

相关论文: Smart But Not Moral? Moral Alignment In Human-AI D…

200 篇论文

Human-AI collaboration is increasingly relevant in consequential areas where AI recommendations support human discretion. However, human-AI teams' effectiveness, capability, and fairness highly depend on human perceptions of AI. Positive…

人机交互 · 计算机科学 2025-06-06 Domenique Zipperling , Luca Deck , Julia Lanzl , Niklas Kühl

Public-sector bureaucracies seek to reap the benefits of artificial intelligence (AI), but face important concerns about accountability and transparency when using AI systems. In particular, perception or actuality of AI agency might create…

计算机与社会 · 计算机科学 2025-08-22 Chris Schmitz , Joanna Bryson

Artificially intelligent systems, given a set of non-trivial ethical rules to follow, will inevitably be faced with scenarios which call into question the scope of those rules. In such cases, human reasoners typically will engage in…

人工智能 · 计算机科学 2019-11-06 John Licato , Zaid Marji , Sophia Abraham

Artificial intelligence (AI) assisting with antimicrobial prescribing raises significant moral questions. Utilising ethical frameworks alongside AI-driven systems, while considering infection specific complexities, can support moral…

计算机与社会 · 计算机科学 2022-12-06 William J Bolton , Cosmin Badea , Pantelis Georgiou , Alison Holmes , Timothy M Rawson

When we consult with a doctor, lawyer, or financial advisor, we generally assume that they are acting in our best interests. But what should we assume when it is an artificial intelligence (AI) system that is acting on our behalf? Early…

计算机与社会 · 计算机科学 2020-03-26 Anthony Aguirre , Gaia Dempsey , Harry Surden , Peter B. Reiner

The challenge of aligning artificial intelligence (AI) with human values persists due to the abstract and often conflicting nature of moral principles and the opacity of existing approaches. This paper introduces CogniAlign, a multi-agent…

计算机与社会 · 计算机科学 2026-04-14 Hasin Jawad Ali , Ilhamul Azam , Ajwad Abrar , Md. Kamrul Hasan , Hasan Mahmud

Many studies have identified particular features of artificial intelligences (AI), such as their autonomy and emotion expression, that affect the extent to which they are treated as subjects of moral consideration. However, there has not…

人机交互 · 计算机科学 2024-03-15 Ali Ladak , Jamie Harris , Jacy Reese Anthis

Case studies commonly form the pedagogical backbone in law, ethics, and many other domains that face complex and ambiguous societal questions informed by human values. Similar complexities and ambiguities arise when we consider how AI…

人工智能 · 计算机科学 2023-11-28 K. J. Kevin Feng , Quan Ze Chen , Inyoung Cheong , King Xia , Amy X. Zhang

Though intelligent agents are supposed to improve human experience (or make it more efficient), it is hard from a human perspective to grasp the ethical values which are explicitly or implicitly embedded in an agent behaviour. This is the…

人工智能 · 计算机科学 2025-08-12 Gianluca Bontempi

While much research in artificial intelligence (AI) has focused on scaling capabilities, the accelerating pace of development makes countervailing work on producing harmless, "aligned" systems increasingly urgent. Yet research on alignment…

人工智能 · 计算机科学 2025-12-12 Dani Roytburg , Beck Miller

As human science pushes the boundaries towards the development of artificial intelligence (AI), the sweep of progress has caused scholars and policymakers alike to question the legality of applying or utilising AI in various human…

人机交互 · 计算机科学 2022-10-11 Dr Brendan Walker-Munro , Dr Zena Assaad

Previous research has shown that the fairness and the legitimacy of a moral decision-maker are important for people's acceptance of and compliance with the decision-maker. As technology rapidly advances, there have been increasing hopes and…

机器人学 · 计算机科学 2022-10-11 Boyoung Kim , Elizabeth Phillips

Fairness in AI and machine learning systems has become a fundamental problem in the accountability of AI systems. While the need for accountability of AI models is near ubiquitous, healthcare in particular is a challenging field where…

机器学习 · 计算机科学 2021-02-09 Ming Yuan , Vikas Kumar , Muhammad Aurangzeb Ahmad , Ankur Teredesai

Ethical AI spans a gamut of considerations. Among these, the most popular ones, fairness and interpretability, have remained largely distinct in technical pursuits. We discuss and elucidate the differences between fairness and…

计算机与社会 · 计算机科学 2021-06-28 Deepak P , Sanil V , Joemon M. Jose

As AI systems increasingly operate with autonomy and adaptability, the traditional boundaries of moral responsibility in techno-social systems are being challenged. This paper explores the evolving discourse on the delegation of…

计算机与社会 · 计算机科学 2024-11-26 Gordana Dodig-Crnkovic , Gianfranco Basti , Tobias Holstein

The deployment of large language models (LLMs) in mental health and other sensitive domains raises urgent questions about ethical reasoning, fairness, and responsible alignment. Yet, existing benchmarks for moral and clinical…

计算与语言 · 计算机科学 2025-09-16 Sai Kartheek Reddy Kasu

The guiding principle of AI alignment is to train large language models (LLMs) to be harmless, helpful, and honest (HHH). At the same time, there are mounting concerns that LLMs exhibit a left-wing political bias. Yet, the commitment to AI…

计算与语言 · 计算机科学 2025-07-22 Thilo Hagendorff

Humans display significant uncertainty when confronted with moral dilemmas, yet the extent of such uncertainty in machines and AI agents remains underexplored. Recent studies have confirmed the overly confident tendencies of…

人工智能 · 计算机科学 2025-11-18 Jea Kwon , Luiz Felipe Vecchietti , Sungwon Park , Meeyoung Cha

The impact of Artificial Intelligence does not depend only on fundamental research and technological developments, but for a large part on how these systems are introduced into society and used in everyday situations. Even though AI is…

计算机与社会 · 计算机科学 2022-02-16 Virginia Dignum

If an AI system makes decisions over time, how should we evaluate how aligned it is with a group of stakeholders (who may have conflicting values and preferences)? In this position paper, we advocate for consideration of temporal aspects…

人工智能 · 计算机科学 2024-11-19 Toryn Q. Klassen , Parand A. Alamdari , Sheila A. McIlraith