中文
相关论文

相关论文: Diagnosing AI Explanation Methods with Folk Concep…

200 篇论文

Answer Set Programming (ASP) is a popular declarative reasoning and problem solving approach in symbolic AI. Its rule-based formalism makes it inherently attractive for explainable and interpretive reasoning, which is gaining importance…

人工智能 · 计算机科学 2026-01-22 Thomas Eiter , Tobias Geibinger , Zeynep G. Saribatur

Explainable components in XAI algorithms often come from a familiar set of models, such as linear models or decision trees. We formulate an approach where the type of explanation produced is guided by a specification. Specifications are…

机器学习 · 计算机科学 2020-12-15 Harish Naik , György Turán

Artificial intelligence (AI) has a history of nearly a century from its inception to the present day. We have summarized the development trends and discovered universal rules, including both success and failure. We have analyzed the reasons…

人工智能 · 计算机科学 2023-03-07 Lin Zhang

The ability to use symbols is the pinnacle of human intelligence, but has yet to be fully replicated in machines. Here we argue that the path towards symbolically fluent artificial intelligence (AI) begins with a reinterpretation of what…

人工智能 · 计算机科学 2022-01-24 Adam Santoro , Andrew Lampinen , Kory Mathewson , Timothy Lillicrap , David Raposo

Mechanistic interpretability is the program of explaining what AI systems are doing in terms of their internal mechanisms. I analyze some aspects of the program, along with setting out some concrete challenges and assessing progress to…

人工智能 · 计算机科学 2025-01-28 David J. Chalmers

Every step we take in the digital world leaves behind a record of our behavior; a digital footprint. Research has suggested that algorithms can translate these digital footprints into accurate estimates of psychological characteristics,…

人工智能 · 计算机科学 2021-12-17 Yanou Ramon , Sandra C. Matz , R. A. Farrokhnia , David Martens

Explainable AI (XAI) is a promising means of supporting human-AI collaborations for high-stakes visual detection tasks, such as damage detection tasks from satellite imageries, as fully-automated approaches are unlikely to be perfectly safe…

人机交互 · 计算机科学 2021-11-05 Donghoon Shin , Sachin Grover , Kenneth Holstein , Adam Perer

Large-scale foundation models exhibit \emph{behavioral shifts} when subjected to interventions such as scaling, fine-tuning, reinforcement learning with human feedback, or in-context learning. Current explainability methods are structurally…

In recent years, the community of 'explainable artificial intelligence' (XAI) has created a vast body of methods to bridge a perceived gap between model 'complexity' and 'interpretability'. However, a concrete problem to be solved by XAI…

机器学习 · 计算机科学 2023-06-05 Rick Wilming , Leo Kieslich , Benedict Clark , Stefan Haufe

Explainable Artificial Intelligence (XAI) aims to provide insights into the decision-making process of AI models, allowing users to understand their results beyond their decisions. A significant goal of XAI is to improve the performance of…

人工智能 · 计算机科学 2023-06-12 Andrea Apicella , Luca Di Lorenzo , Francesco Isgrò , Andrea Pollastro , Roberto Prevete

As AI systems advance beyond human capabilities, scalable oversight becomes critical: how can we supervise AI that exceeds our abilities? A key challenge is that human evaluators may form incorrect beliefs about AI behavior in complex…

人工智能 · 计算机科学 2025-10-22 Leon Lang , Patrick Forré

From its inception, AI has had a rather ambivalent relationship with humans -- swinging between their augmentation and replacement. Now, as AI technologies enter our everyday lives at an ever increasing pace, there is a greater need for AI…

人工智能 · 计算机科学 2024-05-28 Sarath Sreedharan , Anagha Kulkarni , Subbarao Kambhampati

eXplainable Artificial Intelligence (XAI) has garnered significant attention for enhancing transparency and trust in machine learning models. However, the scopes of most existing explanation techniques focus either on offering a holistic…

机器学习 · 计算机科学 2024-12-12 Fanyu Meng , Xin Liu , Zhaodan Kong , Xin Chen

Despite significant advancements in XAI, scholars note a persistent lack of solid conceptual foundations and integration with broader scientific discourse on explanation. In response, emerging research draws on explanatory strategies from…

机器学习 · 计算机科学 2026-05-22 Marcin Rabiza

With the availability of large databases and recent improvements in deep learning methodology, the performance of AI systems is reaching or even exceeding the human level on an increasing number of complex tasks. Impressive examples of this…

人工智能 · 计算机科学 2017-08-29 Wojciech Samek , Thomas Wiegand , Klaus-Robert Müller

Artificial intelligence explanations can make complex predictive models more comprehensible. To be effective, however, they should anticipate and mitigate possible misinterpretations, e.g., arising when users infer incorrect information…

人机交互 · 计算机科学 2025-08-06 Yueqing Xuan , Kacper Sokol , Mark Sanderson , Jeffrey Chan

With humans interacting with AI-based systems at an increasing rate, it is necessary to ensure the artificial systems are acting in a manner which reflects understanding of the human. In the case of humans and artificial AI agents operating…

人机交互 · 计算机科学 2023-02-03 Andrew Fuchs , Andrea Passarella , Marco Conti

Artificial intelligence systems exhibit many useful capabilities, but they appear to lack understanding. This essay describes how we could go about constructing a machine capable of understanding. As John Locke (1689) pointed out words are…

人工智能 · 计算机科学 2024-05-06 Herbert L. Roitblat

Explainable artificial intelligence (XAI) aims to provide human-interpretable insights into the behavior of deep neural networks (DNNs), typically by estimating a simplified causal structure of the model. In existing work, this causal…

计算机视觉与模式识别 · 计算机科学 2026-03-11 Robin Hesse , Simone Schaub-Meyer , Janina Hesse , Bernt Schiele , Stefan Roth

eXplainable AI focuses on generating explanations for the output of an AI algorithm to a user, usually a decision-maker. Such user needs to interpret the AI system in order to decide whether to trust the machine outcome. When addressing…

人机交互 · 计算机科学 2020-05-28 Irene Celino
‹ 上一页 1 8 9 10 下一页 ›