English

Evaluating Neural Morphological Taggers for Sanskrit

Computation and Language 2020-05-25 v1

Abstract

Neural sequence labelling approaches have achieved state of the art results in morphological tagging. We evaluate the efficacy of four standard sequence labelling models on Sanskrit, a morphologically rich, fusional Indian language. As its label space can theoretically contain more than 40,000 labels, systems that explicitly model the internal structure of a label are more suited for the task, because of their ability to generalise to labels not seen during training. We find that although some neural models perform better than others, one of the common causes for error for all of these models is mispredictions due to syncretism.

Keywords

Cite

@article{arxiv.2005.10893,
  title  = {Evaluating Neural Morphological Taggers for Sanskrit},
  author = {Ashim Gupta and Amrith Krishna and Pawan Goyal and Oliver Hellwig},
  journal= {arXiv preprint arXiv:2005.10893},
  year   = {2020}
}

Comments

Accepted to SIGMORPHON Workshop at ACL 2020

R2 v1 2026-06-23T15:43:38.398Z