English

Specializing Multi-domain NMT via Penalizing Low Mutual Information

Computation and Language 2022-10-25 v1 Artificial Intelligence

Abstract

Multi-domain Neural Machine Translation (NMT) trains a single model with multiple domains. It is appealing because of its efficacy in handling multiple domains within one model. An ideal multi-domain NMT should learn distinctive domain characteristics simultaneously, however, grasping the domain peculiarity is a non-trivial task. In this paper, we investigate domain-specific information through the lens of mutual information (MI) and propose a new objective that penalizes low MI to become higher. Our method achieved the state-of-the-art performance among the current competitive multi-domain NMT models. Also, we empirically show our objective promotes low MI to be higher resulting in domain-specialized multi-domain NMT.

Keywords

Cite

@article{arxiv.2210.12910,
  title  = {Specializing Multi-domain NMT via Penalizing Low Mutual Information},
  author = {Jiyoung Lee and Hantae Kim and Hyunchang Cho and Edward Choi and Cheonbok Park},
  journal= {arXiv preprint arXiv:2210.12910},
  year   = {2022}
}

Comments

Accepted in EMNLP 2022

R2 v1 2026-06-28T04:18:57.988Z