English

In Defence of Post-hoc Explainability

Machine Learning 2025-10-31 v2 Artificial Intelligence

Abstract

This position paper defends post-hoc explainability methods as legitimate tools for scientific knowledge production in machine learning. Addressing criticism of these methods' reliability and epistemic status, we develop a philosophical framework grounded in mediated understanding and bounded factivity. We argue that scientific insights can emerge through structured interpretation of model behaviour without requiring complete mechanistic transparency, provided explanations acknowledge their approximative nature and undergo rigorous empirical validation. Through analysis of recent biomedical ML applications, we demonstrate how post-hoc methods, when properly integrated into scientific practice, generate novel hypotheses and advance phenomenal understanding.

Keywords

Cite

@article{arxiv.2412.17883,
  title  = {In Defence of Post-hoc Explainability},
  author = {Nick Oh},
  journal= {arXiv preprint arXiv:2412.17883},
  year   = {2025}
}

Comments

v1 presented at the Interpretable AI: Past, Present, and Future Workshop at NeurIPS 2024 (non-archival)

R2 v1 2026-06-28T20:47:18.093Z