English

CompiLIG at SemEval-2017 Task 1: Cross-Language Plagiarism Detection Methods for Semantic Textual Similarity

Computation and Language 2017-04-06 v1

Abstract

We present our submitted systems for Semantic Textual Similarity (STS) Track 4 at SemEval-2017. Given a pair of Spanish-English sentences, each system must estimate their semantic similarity by a score between 0 and 5. In our submission, we use syntax-based, dictionary-based, context-based, and MT-based methods. We also combine these methods in unsupervised and supervised way. Our best run ranked 1st on track 4a with a correlation of 83.02% with human annotations.

Keywords

Cite

@article{arxiv.1704.01346,
  title  = {CompiLIG at SemEval-2017 Task 1: Cross-Language Plagiarism Detection Methods for Semantic Textual Similarity},
  author = {Jeremy Ferrero and Frederic Agnes and Laurent Besacier and Didier Schwab},
  journal= {arXiv preprint arXiv:1704.01346},
  year   = {2017}
}