English

Development of a Hindi Lemmatizer

Computation and Language 2013-07-16 v2

Abstract

We live in a translingual society, in order to communicate with people from different parts of the world we need to have an expertise in their respective languages. Learning all these languages is not at all possible; therefore we need a mechanism which can do this task for us. Machine translators have emerged as a tool which can perform this task. In order to develop a machine translator we need to develop several different rules. The very first module that comes in machine translation pipeline is morphological analysis. Stemming and lemmatization comes under morphological analysis. In this paper we have created a lemmatizer which generates rules for removing the affixes along with the addition of rules for creating a proper root word.

Keywords

Cite

@article{arxiv.1305.6211,
  title  = {Development of a Hindi Lemmatizer},
  author = {Snigdha Paul and Nisheeth Joshi and Iti Mathur},
  journal= {arXiv preprint arXiv:1305.6211},
  year   = {2013}
}

Comments

International Journal of Computational Linguistics and Natural Language Processing, Vol 2, Issue 5, 2013

R2 v1 2026-06-22T00:23:10.399Z