Encoding Source Language with Convolutional Neural Network for Machine Translation (1503.01838v5)

Published 6 Mar 2015 in cs.CL, cs.LG, and cs.NE

Abstract: The recently proposed neural network joint model (NNJM) (Devlin et al., 2014) augments the n-gram target LLM with a heuristically chosen source context window, achieving state-of-the-art performance in SMT. In this paper, we give a more systematic treatment by summarizing the relevant source information through a convolutional architecture guided by the target information. With different guiding signals during decoding, our specifically designed convolution+gating architectures can pinpoint the parts of a source sentence that are relevant to predicting a target word, and fuse them with the context of entire source sentence to form a unified representation. This representation, together with target language words, are fed to a deep neural network (DNN) to form a stronger NNJM. Experiments on two NIST Chinese-English translation tasks show that the proposed model can achieve significant improvements over the previous NNJM by up to +1.08 BLEU points on average

Authors (6)

Fandong Meng (174 papers)
Zhengdong Lu (35 papers)
Mingxuan Wang (83 papers)
Hang Li (277 papers)
Wenbin Jiang (18 papers)
Qun Liu (230 papers)

Citations (105)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Encoding Source Language with Convolutional Neural Network for Machine Translation (1503.01838v5)

Summary

Related Papers