中文
相关论文

相关论文: Revisiting the robustness of post-hoc interpretabi…

200 篇论文

Explainable artificial intelligence (XAI) methods have become increasingly important in the context of explainable intrusion detection systems (X-IDSs) for improving the interpretability and trustworthiness of X-IDSs. However, existing…

密码学与安全 · 计算机科学 2025-05-14 Mohammed Alquliti , Erisa Karafili , BooJoong Kang

The recent spike in certified Artificial Intelligence (AI) tools for healthcare has renewed the debate around adoption of this technology. One thread of such debate concerns Explainable AI (XAI) and its promise to render AI devices more…

人工智能 · 计算机科学 2023-02-27 Giovanni Cinà , Tabea E. Röber , Rob Goedhart , Ş. İlker Birbil

In many scenarios, the interpretability of machine learning models is a highly required but difficult task. To explain the individual predictions of such models, local model-agnostic approaches have been proposed. However, the process…

机器学习 · 统计学 2025-10-22 Gianluigi Lopardo , Frederic Precioso , Damien Garreau

Most existing interpretable methods explain a black-box model in a post-hoc manner, which uses simpler models or data analysis techniques to interpret the predictions after the model is learned. However, they (a) may derive contradictory…

机器学习 · 计算机科学 2020-01-22 Mengzhuo Guo , Qingpeng Zhang , Xiuwu Liao , Daniel Dajun Zeng

As machine learning and algorithmic decision making systems are increasingly being leveraged in high-stakes human-in-the-loop settings, there is a pressing need to understand the rationale of their predictions. Researchers have responded to…

机器学习 · 计算机科学 2020-12-07 Jonathan Dinu , Jeffrey Bigham , J. Zico Kolter

Explainable AI (XAI) aims to improve user understanding and decisions when using AI models. However, despite innovations in XAI, recent user evaluations reveal that this goal remains elusive. Understanding human cognition can help explain…

人工智能 · 计算机科学 2026-05-01 Louth Bin Rawshan , Zhuoyu Wang , Brian Y. Lim

Explainable artificial intelligence (XAI) aims to develop transparent explanatory approaches for "black-box" deep learning models. However,it remains difficult for existing methods to achieve the trade-off of the three key criteria in…

计算机视觉与模式识别 · 计算机科学 2023-12-27 Changqi Sun , Hao Xu , Yuntian Chen , Dongxiao Zhang

Post-hoc model-agnostic interpretation methods such as partial dependence plots can be employed to interpret complex machine learning models. While these interpretation methods can be applied regardless of model complexity, they can produce…

机器学习 · 统计学 2022-01-24 Christoph Molnar , Giuseppe Casalicchio , Bernd Bischl

The field of explainable artificial intelligence (XAI) aims to explain how black-box machine learning models work. Much of the work centers around the holy grail of providing post-hoc feature attributions to any model architecture. While…

机器学习 · 计算机科学 2023-11-15 Brian Barr , Noah Fatsi , Leif Hancox-Li , Peter Richter , Daniel Proano , Caleb Mok

Explainable Artificial Intelligence (XAI) techniques are used to provide transparency to complex, opaque predictive models. However, these techniques are often designed for image and text data, and it is unclear how fit-for-purpose they are…

计算机与社会 · 计算机科学 2024-10-18 Mythreyi Velmurugan , Chun Ouyang , Yue Xu , Renuka Sindhgatta , Bemali Wickramanayake , Catarina Moreira

EXplainable AI (XAI) methods have been proposed to interpret how a deep neural network predicts inputs through model saliency explanations that highlight the parts of the inputs deemed important to arrive a decision at a specific target.…

计算机视觉与模式识别 · 计算机科学 2020-09-23 Yi-Shan Lin , Wen-Chuan Lee , Z. Berkay Celik

Deep neural networks for medical image diagnosis often achieve high predictive accuracy while relying on spurious or clinically irrelevant visual cues, limiting their trustworthiness in practice. Post-hoc explanation methods are widely used…

计算机视觉与模式识别 · 计算机科学 2026-05-12 Zubair Faruqui , Rahul Dubey

Counterfactual post-hoc interpretability approaches have been proven to be useful tools to generate explanations for the predictions of a trained blackbox classifier. However, the assumptions they make about the data and the classifier make…

机器学习 · 计算机科学 2019-06-13 Thibault Laugel , Marie-Jeanne Lesot , Christophe Marsala , Marcin Detyniecki

The ubiquity of machine learning based predictive models in modern society naturally leads people to ask how trustworthy those models are? In predictive modeling, it is quite common to induce a trade-off between accuracy and…

机器学习 · 计算机科学 2019-04-05 John Mitros , Brian Mac Namee

Perturbation-based post-hoc image explanation methods are commonly used to explain image prediction models. These methods perturb parts of the input to measure how those parts affect the output. Since the methods only require the input and…

计算机视觉与模式识别 · 计算机科学 2025-11-18 Gustav Grund Pihlgren , Kary Främling

We argue that robustness of explanations---i.e., that similar inputs should give rise to similar explanations---is a key desideratum for interpretability. We introduce metrics to quantify robustness and demonstrate that current methods do…

机器学习 · 计算机科学 2018-06-22 David Alvarez-Melis , Tommi S. Jaakkola

Explainable AI (XAI) methods generally fall into two categories. Post-hoc approaches generate explanations for pre-trained models and are compatible with various neural network architectures. These methods often use feature importance…

计算机视觉与模式识别 · 计算机科学 2025-05-20 Piotr Borycki , Magdalena Trędowicz , Szymon Janusz , Jacek Tabor , Przemysław Spurek , Arkadiusz Lewicki , Łukasz Struski

Explainable Artificial Intelligence (XAI) has aided machine learning (ML) researchers with the power of scrutinizing the decisions of the black-box models. XAI methods enable looking deep inside the models' behavior, eventually generating…

密码学与安全 · 计算机科学 2025-10-07 Maraz Mia , Mir Mehedi A. Pritom

Explainable AI has become a common term in the literature, scrutinized by computer scientists and statisticians and highlighted by psychological or philosophical researchers. One major effort many researchers tackle is constructing general…

Reliable probability estimates are critical in many machine learning applications, yet modern classifiers are often poorly calibrated. Post-hoc calibration provides a simple and widely used solution, but the large number of proposed…

机器学习 · 计算机科学 2026-05-29 Eugène Berta , David Holzmüller , Francis Bach , Michael I. Jordan