Deep Discriminant Analysis for i-vector Based Robust Speaker Recognition (1805.01344v1)

Published 3 May 2018 in cs.SD and eess.AS

Abstract: Linear Discriminant Analysis (LDA) has been used as a standard post-processing procedure in many state-of-the-art speaker recognition tasks. Through maximizing the inter-speaker difference and minimizing the intra-speaker variation, LDA projects i-vectors to a lower-dimensional and more discriminative sub-space. In this paper, we propose a neural network based compensation scheme(termed as deep discriminant analysis, DDA) for i-vector based speaker recognition, which shares the spirit with LDA. Optimized against softmax loss and center loss at the same time, the proposed method learns a more compact and discriminative embedding space. Compared with the Gaussian distribution assumption of data and the learnt linear projection in LDA, the proposed method doesn't pose any assumptions on data and can learn a non-linear projection function. Experiments are carried out on a short-duration text-independent dataset based on the SRE Corpus, noticeable performance improvement can be observed against the normal LDA or PLDA methods.

Authors (4)

Shuai Wang (466 papers)
Zili Huang (18 papers)
Yanmin Qian (99 papers)
Kai Yu (202 papers)

Citations (8)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Deep Discriminant Analysis for i-vector Based Robust Speaker Recognition (1805.01344v1)

Summary

Related Papers