English

A Summary of the First Workshop on Language Technology for Language Documentation and Revitalization

Computation and Language 2020-04-29 v1

Abstract

Despite recent advances in natural language processing and other language technology, the application of such technology to language documentation and conservation has been limited. In August 2019, a workshop was held at Carnegie Mellon University in Pittsburgh to attempt to bring together language community members, documentary linguists, and technologists to discuss how to bridge this gap and create prototypes of novel and practical language revitalization technologies. This paper reports the results of this workshop, including issues discussed, and various conceived and implemented technologies for nine languages: Arapaho, Cayuga, Inuktitut, Irish Gaelic, Kidaw'ida, Kwak'wala, Ojibwe, San Juan Quiahije Chatino, and Seneca.

Cite

@article{arxiv.2004.13203,
  title  = {A Summary of the First Workshop on Language Technology for Language Documentation and Revitalization},
  author = {Graham Neubig and Shruti Rijhwani and Alexis Palmer and Jordan MacKenzie and Hilaria Cruz and Xinjian Li and Matthew Lee and Aditi Chaudhary and Luke Gessler and Steven Abney and Shirley Anugrah Hayati and Antonios Anastasopoulos and Olga Zamaraeva and Emily Prud'hommeaux and Jennette Child and Sara Child and Rebecca Knowles and Sarah Moeller and Jeffrey Micher and Yiyuan Li and Sydney Zink and Mengzhou Xia and Roshan S Sharma and Patrick Littell},
  journal= {arXiv preprint arXiv:2004.13203},
  year   = {2020}
}

Comments

Accepted at SLTU-CCURL 2020

R2 v1 2026-06-23T15:08:22.214Z