Triangular Architecture for Rare Language Translation

  • 2018-07-11 04:56:06
  • Shuo Ren, Wenhu Chen, Shujie Liu, Mu Li, Ming Zhou, Shuai Ma
  • 0

Abstract

Neural Machine Translation (NMT) performs poor on the low-resource languagepair $(X,Z)$, especially when $Z$ is a rare language. By introducing anotherrich language $Y$, we propose a novel triangular training architecture (TA-NMT)to leverage bilingual data $(Y,Z)$ (may be small) and $(X,Y)$ (can be rich) toimprove the translation performance of low-resource pairs. In this triangulararchitecture, $Z$ is taken as the intermediate latent variable, and translationmodels of $Z$ are jointly optimized with a unified bidirectional EM algorithmunder the goal of maximizing the translation likelihood of $(X,Y)$. Empiricalresults demonstrate that our method significantly improves the translationquality of rare languages on MultiUN and IWSLT2012 datasets, and achieves evenbetter performance combining back-translation methods.

 

Quick Read (beta)

loading the full paper ...