English

Sensivity of LLMs' Explanations to the Training Randomness:Context, Class & Task Dependencies

Computation and Language 2026-03-10 v1

Abstract

Transformer models are now a cornerstone in natural language processing. Yet, explaining their decisions remains a challenge. It was shown recently that the same model trained on the same data with a different randomness can lead to very different explanations. In this paper, we investigate how the (syntactic) context, the classes to be learned and the tasks influence this explanations' sensitivity to randomness. We show that they all have statistically significant impact: smallest for the (syntactic) context, medium for the classes and largest for the tasks.

Keywords

Cite

@article{arxiv.2603.08241,
  title  = {Sensivity of LLMs' Explanations to the Training Randomness:Context, Class & Task Dependencies},
  author = {Romain Loncour and Jérémie Bogaert and François-Xavier Standaert},
  journal= {arXiv preprint arXiv:2603.08241},
  year   = {2026}
}

Comments

6 pages, 6 figures

R2 v1 2026-07-01T11:10:07.269Z