English

Automatic Design of Semantic Similarity Ensembles Using Grammatical Evolution

Computation and Language 2025-04-28 v8 Artificial Intelligence

Abstract

Semantic similarity measures are a key component in natural language processing tasks such as document analysis, requirement matching, and user input interpretation. However, the performance of individual measures varies considerably across datasets. To address this, ensemble approaches that combine multiple measures are often employed. This paper presents an automated strategy based on grammatical evolution for constructing semantic similarity ensembles. The method evolves aggregation functions that maximize correlation with human-labeled similarity scores. Experiments on standard benchmark datasets demonstrate that the proposed approach outperforms existing ensemble techniques in terms of accuracy. The results confirm the effectiveness of grammatical evolution in designing adaptive and accurate similarity models. The source code that illustrates our approach can be downloaded from https://github.com/jorge-martinez-gil/sesige.

Keywords

Cite

@article{arxiv.2307.00925,
  title  = {Automatic Design of Semantic Similarity Ensembles Using Grammatical Evolution},
  author = {Jorge Martinez-Gil},
  journal= {arXiv preprint arXiv:2307.00925},
  year   = {2025}
}

Comments

38 pages

R2 v1 2026-06-28T11:20:37.974Z