English

Low-cost concept-based localized explanations: How far can we get with training-free approaches?

Artificial Intelligence 2026-06-27 v1 Computation and Language Computer Vision and Pattern Recognition

Abstract

Concept-based Explainable AI (C-XAI) seeks human-understandable explanations grounded in semantic concepts, yet validation is limited by the scarcity of fine-grained concept annotations. We evaluate whether mid-scale Multimodal Large Language Models (MLLMs) can perform localized concept naming under strict zero-shot conditions by assigning labels to bounding-box regions at both object and part levels. We propose a reproducible zero-shot evaluation protocol for Concept Naming (CoNa) with (i) closed-set, category-constrained prompting for moderate vocabularies and (ii) Open-CoNa, an embedding-similarity-based strategy for large label spaces. Experiments with four MLLMs (7B-32B) show consistent performance trends across datasets, reaching 62%-88% object-level exact-match accuracy, highlighting the potential of training-free concept annotation from localized regions. We discuss limitations and failure modes and release a reproducible framework to support future low-cost C-XAI research.

Keywords

Cite

@article{arxiv.2606.29069,
  title  = {Low-cost concept-based localized explanations: How far can we get with training-free approaches?},
  author = {Darian Fernández-Gutiérrez and Rafael Bello and Marilyn Bello and Natalia Díaz-Rodríguez},
  journal= {arXiv preprint arXiv:2606.29069},
  year   = {2026}
}

Comments

6 pages, 2 figures, 4 tables. Accepted at the 2026 IEEE International Conference on Artificial Intelligence (CAI), 8-10 May 2026, Granada, Spain. Code: https://github.com/darianfgUgr/CoNa