中文

跨语言词嵌入评估的局限性

计算与语言 2018-06-07 v1

摘要

本工作旨在探索现有跨语言词嵌入评估方法的潜在局限性,针对内在与外在跨语言评估方法之间缺乏相关性的问题进行探讨。为证明该假设,我们构建了用于外在和内在评估任务的英语-俄语数据集,并比较了5种不同跨语言模型在其上的表现。结果表明,即使在不同内在基准上的分数也互不相关。我们可以得出结论,除非了解母语者在认知中如何处理语义,否则将人类参考作为跨语言词嵌入的真实基准是不合适的。

关键词

引用

@article{arxiv.1806.02253,
  title  = {The Limitations of Cross-language Word Embeddings Evaluation},
  author = {Amir Bakarov and Roman Suvorov and Ilya Sochenkov},
  journal= {arXiv preprint arXiv:1806.02253},
  year   = {2018}
}

备注

In Proceedings of the 7th Joint Conference on Lexical and Computational Semantics (*SEM 2018)