English

LegalNLP -- Natural Language Processing methods for the Brazilian Legal Language

Computation and Language 2021-11-01 v1 Machine Learning

Abstract

We present and make available pre-trained language models (Phraser, Word2Vec, Doc2Vec, FastText, and BERT) for the Brazilian legal language, a Python package with functions to facilitate their use, and a set of demonstrations/tutorials containing some applications involving them. Given that our material is built upon legal texts coming from several Brazilian courts, this initiative is extremely helpful for the Brazilian legal field, which lacks other open and specific tools and language models. Our main objective is to catalyze the use of natural language processing tools for legal texts analysis by the Brazilian industry, government, and academia, providing the necessary tools and accessible material.

Keywords

Cite

@article{arxiv.2110.15709,
  title  = {LegalNLP -- Natural Language Processing methods for the Brazilian Legal Language},
  author = {Felipe Maia Polo and Gabriel Caiaffa Floriano Mendonça and Kauê Capellato J. Parreira and Lucka Gianvechio and Peterson Cordeiro and Jonathan Batista Ferreira and Leticia Maria Paz de Lima and Antônio Carlos do Amaral Maia and Renato Vicente},
  journal= {arXiv preprint arXiv:2110.15709},
  year   = {2021}
}