Adversarial Training for Community Question Answer Selection Based on Multi-scale Matching (1804.08058v2)

Published 22 Apr 2018 in cs.CL

Abstract: Community-based question answering (CQA) websites represent an important source of information. As a result, the problem of matching the most valuable answers to their corresponding questions has become an increasingly popular research topic. We frame this task as a binary (relevant/irrelevant) classification problem, and present an adversarial training framework to alleviate label imbalance issue. We employ a generative model to iteratively sample a subset of challenging negative samples to fool our classification model. Both models are alternatively optimized using REINFORCE algorithm. The proposed method is completely different from previous ones, where negative samples in training set are directly used or uniformly down-sampled. Further, we propose using Multi-scale Matching which explicitly inspects the correlation between words and ngrams of different levels of granularity. We evaluate the proposed method on SemEval 2016 and SemEval 2017 datasets and achieves state-of-the-art or similar performance.

PDF Abstract

Summarize Bookmark Chat (Pro)

Authors (7)

Xiao Yang (158 papers)
Madian Khabsa (38 papers)
Miaosen Wang (11 papers)
Wei Wang (1793 papers)
Ahmed Awadallah (27 papers)
Daniel Kifer (65 papers)
C. Lee Giles (69 papers)

Citations (22)

View on Semantic Scholar

Adversarial Training for Community Question Answer Selection Based on Multi-scale Matching (1804.08058v2)

Related Papers