Information Maximized Variational Domain Adversarial Learning for Speaker Verification

Domain mismatch is a common problem in speaker ver- ification. This paper proposes an information-maximized variational domain adversarial neural network (InfoVDANN) to reduce domain mismatch by incorporating an InfoVAE into domain adversarial training (DAT). DAT aims to pro- duce speaker discriminative and domain-invariant features. The InfoVAE has two roles. First, it performs variational regularization on the learned features so that they follow a Gaussian distribution, which is essential for the standard PLDA backend. Second, it preserves mutual information be- tween the features and the training set to extract extra speaker discriminative information. Experiments on both SRE16 and SRE18-CMN2 show that the InfoVDANN outperforms the recent VDANN, which suggests that increasing the mutual information between the latent features and input features enables the InfoVDANN to extract extra speaker information that is otherwise not possible.

5091-TuMakChien.pdf

5091-TuMakChien.pdf (391)

Thumbs Up

CITE

Documents

Presentation Slides

Information Maximized Variational Domain Adversarial Learning for Speaker Verification

5091-TuMakChien.pdf

QUESTIONS?