LegalNLP——面向巴西法律语言的自然语言处理方法
计算与语言
2021-11-01 v1 机器学习
摘要
我们提出并公开了针对巴西法律语言的预训练语言模型(Phraser、Word2Vec、Doc2Vec、FastText和BERT),一个包含便于使用这些模型的函数的Python包,以及一组包含相关应用的演示/教程。鉴于我们的材料基于来自多个巴西法院的法律文本,这一举措对缺乏其他开放且专用工具和语言模型的巴西法律领域极为有益。我们的主要目标是提供必要的工具和可访问的材料,以促进巴西工业界、政府和学术界使用自然语言处理工具进行法律文本分析。
引用
@article{arxiv.2110.15709,
title = {LegalNLP -- Natural Language Processing methods for the Brazilian Legal Language},
author = {Felipe Maia Polo and Gabriel Caiaffa Floriano Mendonça and Kauê Capellato J. Parreira and Lucka Gianvechio and Peterson Cordeiro and Jonathan Batista Ferreira and Leticia Maria Paz de Lima and Antônio Carlos do Amaral Maia and Renato Vicente},
journal= {arXiv preprint arXiv:2110.15709},
year = {2021}
}