Inference Strategies for Machine Translation with Conditional Masking (2010.02352v2)

Published 5 Oct 2020 in cs.CL

Abstract: Conditional masked LLM (CMLM) training has proven successful for non-autoregressive and semi-autoregressive sequence generation tasks, such as machine translation. Given a trained CMLM, however, it is not clear what the best inference strategy is. We formulate masked inference as a factorization of conditional probabilities of partial sequences, show that this does not harm performance, and investigate a number of simple heuristics motivated by this perspective. We identify a thresholding strategy that has advantages over the standard "mask-predict" algorithm, and provide analyses of its behavior on machine translation tasks.

PDF Abstract

Summarize Bookmark Chat (Pro)

Authors (3)

Julia Kreutzer (44 papers)
George Foster (24 papers)
Colin Cherry (38 papers)

Citations (5)

View on Semantic Scholar

Inference Strategies for Machine Translation with Conditional Masking (2010.02352v2)

Related Papers