2000 character limit reached
Learning Coupled Policies for Simultaneous Machine Translation using Imitation Learning (2002.04306v2)
Published 11 Feb 2020 in cs.CL, cs.AI, and cs.LG
Abstract: We present a novel approach to efficiently learn a simultaneous translation model with coupled programmer-interpreter policies. First, wepresent an algorithmic oracle to produce oracle READ/WRITE actions for training bilingual sentence-pairs using the notion of word alignments. This oracle actions are designed to capture enough information from the partial input before writing the output. Next, we perform a coupled scheduled sampling to effectively mitigate the exposure bias when learning both policies jointly with imitation learning. Experiments on six language-pairs show our method outperforms strong baselines in terms of translation quality while keeping the translation delay low.
- Philip Arthur (9 papers)
- Trevor Cohn (105 papers)
- Gholamreza Haffari (141 papers)