NEXUS Network: Connecting the Preceding and the Following in Dialogue Generation (1810.00671v2)

Published 27 Sep 2018 in cs.CL and cs.AI

Abstract: Sequence-to-Sequence (seq2seq) models have become overwhelmingly popular in building end-to-end trainable dialogue systems. Though highly efficient in learning the backbone of human-computer communications, they suffer from the problem of strongly favoring short generic responses. In this paper, we argue that a good response should smoothly connect both the preceding dialogue history and the following conversations. We strengthen this connection through mutual information maximization. To sidestep the non-differentiability of discrete natural language tokens, we introduce an auxiliary continuous code space and map such code space to a learnable prior distribution for generation purpose. Experiments on two dialogue datasets validate the effectiveness of our model, where the generated responses are closely related to the dialogue context and lead to more interactive conversations.

PDF Abstract

Summarize PDF Markdown Bookmark Chat (Pro)

Authors (4)

Hui Su (38 papers)
Xiaoyu Shen (73 papers)
Wenjie Li (183 papers)
Dietrich Klakow (114 papers)

Citations (29)

View on Semantic Scholar

NEXUS Network: Connecting the Preceding and the Following in Dialogue Generation (1810.00671v2)

Related Papers