This paper describes the Notre Dame Natural Language Processing Group's (NDNLP) submission to the WNGT 2019 shared task (Hayashi et al., 2019). We investigated the impact of auto-sizing (Murray and Chiang, 2015; Murray et al., 2019) to the Transformer network (Vaswani et al., 2017) with the goal of substantially reducing the number of parameters in the model. Our method was able to eliminate more than 25% of the model's parameters while suffering a decrease of only 1.1 BLEU.
@article{arxiv.1910.07134,
title = {Efficiency through Auto-Sizing: Notre Dame NLP's Submission to the WNGT 2019 Efficiency Task},
author = {Kenton Murray and Brian DuSell and David Chiang},
journal= {arXiv preprint arXiv:1910.07134},
year = {2019}
}
Comments
The 3rd Workshop on Neural Generation and Translation (WNGT 2019)