Speaker recognition method based on Gaussian mixture model embedded with time delay neural network

yuhua,Hongxia Dai,wangjilin,Li Zhao,Xin Wei

Speaker recognition method based on Gaussian mixture model embedded with time delay neural network

2009

The invention discloses a speaker recognition method based on a Gaussian mixture model (GMM) embedded with a time delay neural network (TDNN). In the speaker recognition method, the advantages of the TDNN and the GMM are fully considered, the TDNN is embedded into the GMM, and solves a residual of input and output vectors of the TDNN by fully utilizing the time sequence of an input characteristic vector through the conversion of a time delay network, and the residual modifies the training of the GMM through an expectation maximization method; besides, a likelihood probability is acquired by a modified GMM model parameter and the residual, and a TDNN parameter is modified by an inertial backward inversion method so as to ensure that parameters of the GMM and the TDNN are alternately updated. An experiment shows that: a recognition rate of the method is improved to a certain extent compared with that of a baseline GMM under various signal to noise ratios.

Keywords:

Correction
Source
Cite
Save
Machine Reading By IdeaReader

References

Citations