中文
相关论文

相关论文: Reasonable Machines: A Research Manifesto

200 篇论文

Mechanistic interpretability (MI) aims to explain how neural networks work by uncovering their underlying mechanisms. As the field grows in influence, it is increasingly important to examine not just models themselves, but the assumptions,…

Explainability and its emerging counterpart contestability have become important normative and design principles for trustworthy AI as they enable users and subjects to understand and challenge AI decisions. However, realizing these…

计算机与社会 · 计算机科学 2025-08-15 Timothée Schmude , Mireia Yurrita , Kars Alfrink , Thomas Le Goff , Sebastian Tschiatschek , Tiphaine Viard

The recent rapid advancements in artificial intelligence research and deployment have sparked more discussion about the potential ramifications of socially- and emotionally-intelligent AI. The question is not if research can produce such…

人工智能 · 计算机科学 2021-07-30 Desmond C. Ong

The critical inquiry pervading the realm of Philosophy, and perhaps extending its influence across all Humanities disciplines, revolves around the intricacies of morality and normativity. Surprisingly, in recent years, this thematic thread…

人工智能 · 计算机科学 2024-06-19 Nicholas Kluge Corrêa

Accountability is widely understood as a goal for well governed computer systems, and is a sought-after value in many governance contexts. But how can it be achieved? Recent work on standards for governable artificial intelligence systems…

计算机与社会 · 计算机科学 2021-08-23 Joshua A. Kroll

Various forms of implications of artificial intelligence that either exacerbate or decrease racial systemic injustice have been explored in this applied research endeavor. Taking each thematic area of identifying, analyzing, and debating an…

计算机与社会 · 计算机科学 2022-01-05 Alia Abbas

Recent work on interpretability in machine learning and AI has focused on the building of simplified models that approximate the true criteria used to make decisions. These models are a useful pedagogical device for teaching trained…

人工智能 · 计算机科学 2018-11-06 Brent Mittelstadt , Chris Russell , Sandra Wachter

Advanced communication protocols are critical to enable the coexistence of autonomous robots with humans. Thus, the development of explanatory capabilities is an urgent first step toward autonomous robots. This survey provides an overview…

人工智能 · 计算机科学 2021-05-07 Tatsuya Sakai , Takayuki Nagai

Validity, reliability, and fairness are core ethical principles embedded in classical argument-based assessment validation theory. These principles are also central to the Standards for Educational and Psychological Testing (2014) which…

计算机与社会 · 计算机科学 2024-11-06 Jill Burstein , Geoffrey T. LaFlair

In this paper I argue that the search for explainable models and interpretable decisions in AI must be reformulated in terms of the broader project of offering a pragmatic and naturalistic account of understanding in AI. Intuitively, the…

人工智能 · 计算机科学 2020-06-23 Andrés Páez

Local governments increasingly use artificial intelligence (AI) for automated decision-making. Contestability, making systems responsive to dispute, is a way to ensure they respect human rights to autonomy and dignity. We investigate the…

人机交互 · 计算机科学 2023-02-10 Kars Alfrink , Ianus Keller , Neelke Doorn , Gerd Kortuem

Artificial Knowledge (AK) systems are transforming decision-making across critical domains such as healthcare, finance, and criminal justice. However, their growing opacity presents governance challenges that current regulatory approaches,…

计算机与社会 · 计算机科学 2025-05-29 Dalit Ken-Dror Feldman , Daniel Benoliel

As chatbots increasingly blur the boundary between automated systems and human conversation, the foundations of trust in these systems warrant closer examination. While regulatory and policy frameworks tend to define trust in normative…

人工智能 · 计算机科学 2026-03-11 Aditya Gulati , Nuria Oliver

As artificial intelligence systems increasingly inform high-stakes decisions across sectors, transparency has become foundational to responsible and trustworthy AI implementation. Leveraging our role as a leading institute in advancing AI…

机器学习 · 计算机科学 2025-08-01 Dhanesh Ramachandram , Himanshu Joshi , Judy Zhu , Dhari Gandhi , Lucas Hartman , Ananya Raval

Trustworthy Artificial Intelligence (TAI) integrates ethics that align with human values, looking at their influence on AI behaviour and decision-making. Primarily dependent on self-assessment, TAI evaluation aims to ensure ethical…

计算机与社会 · 计算机科学 2024-09-13 Louise McCormack , Malika Bendechache

Our research endeavors to advance the concept of responsible artificial intelligence (AI), a topic of increasing importance within EU policy discussions. The EU has recently issued several publications emphasizing the necessity of trust in…

人工智能 · 计算机科学 2024-03-12 Sabrina Goellner , Marina Tropmann-Frick , Bostjan Brumen

The design of current natural language oriented robot architectures enables certain architectural components to circumvent moral reasoning capabilities. One example of this is reflexive generation of clarification requests as soon as…

人工智能 · 计算机科学 2020-07-20 Ryan Blake Jackson , Tom Williams

The Rational Speech Act (RSA) model provides a flexible framework to model pragmatic reasoning in computational terms. However, state-of-the-art RSA models are still fairly distant from modern machine learning techniques and present a…

计算与语言 · 计算机科学 2024-04-05 Gaia Carenini , Luca Bischetti , Walter Schaeken , Valentina Bambini

Agentic AI seeks to endow systems with sustained autonomy, reasoning, and interaction capabilities. To realize this vision, its assumptions about agency must be complemented by explicit models of cognition, cooperation, and governance. This…

人工智能 · 计算机科学 2026-02-11 Virginia Dignum , Frank Dignum

Scientific research organizations that are developing and deploying Artificial Intelligence (AI) systems are at the intersection of technological progress and ethical considerations. The push for Responsible AI (RAI) in such institutions…

人工智能 · 计算机科学 2023-12-18 Muneera Bano , Didar Zowghi , Pip Shea , Georgina Ibarra
‹ 上一页 1 8 9 10 下一页 ›