English

Interpretable Text Embeddings and Text Similarity Explanation: A Survey

Computation and Language 2025-10-03 v2 Artificial Intelligence Information Retrieval

Abstract

Text embeddings are a fundamental component in many NLP tasks, including classification, regression, clustering, and semantic search. However, despite their ubiquitous application, challenges persist in interpreting embeddings and explaining similarities between them. In this work, we provide a structured overview of methods specializing in inherently interpretable text embeddings and text similarity explanation, an underexplored research area. We characterize the main ideas, approaches, and trade-offs. We compare means of evaluation, discuss overarching lessons learned and finally identify opportunities and open challenges for future research.

Keywords

Cite

@article{arxiv.2502.14862,
  title  = {Interpretable Text Embeddings and Text Similarity Explanation: A Survey},
  author = {Juri Opitz and Lucas Möller and Andrianos Michail and Sebastian Padó and Simon Clematide},
  journal= {arXiv preprint arXiv:2502.14862},
  year   = {2025}
}

Comments

EMNLP 2025 (main)

R2 v1 2026-06-28T21:51:50.358Z