English

Gender Bias in Multilingual Neural Machine Translation: The Architecture Matters

Computation and Language 2020-12-25 v1

Abstract

Multilingual Neural Machine Translation architectures mainly differ in the amount of sharing modules and parameters among languages. In this paper, and from an algorithmic perspective, we explore if the chosen architecture, when trained with the same data, influences the gender bias accuracy. Experiments in four language pairs show that Language-Specific encoders-decoders exhibit less bias than the Shared encoder-decoder architecture. Further interpretability analysis of source embeddings and the attention shows that, in the Language-Specific case, the embeddings encode more gender information, and its attention is more diverted. Both behaviors help in mitigating gender bias.

Keywords

Cite

@article{arxiv.2012.13176,
  title  = {Gender Bias in Multilingual Neural Machine Translation: The Architecture Matters},
  author = {Marta R. Costa-jussà and Carlos Escolano and Christine Basta and Javier Ferrando and Roser Batlle and Ksenia Kharitonova},
  journal= {arXiv preprint arXiv:2012.13176},
  year   = {2020}
}

Comments

12 pages, 5 figures, 3 tables