Combining Evidence from a Generative and a Discriminative Model in Phoneme Recognition

We investigate the use of the log-likelihood of the features obtained from a generative Gaussian mixture model, and the posterior probability of phonemes from a discriminative multilayered perceptron in multi-stream combination for recognition of phonemes. Multi-stream combination techniques, namely early integration and late integration are used to combine the evidence from these models. By using multi-stream combination, we obtain a phoneme recognition accuracy of 74\% on the standard TIMIT database, an absolute improvement of 2.5\% over the single best stream.


Presented at:
Proceedings of Interspeech
Year:
2008
Note:
IDIAP-RR 08-20
Laboratories:




 Record created 2010-02-11, last modified 2018-03-17

n/a:
Download fulltextPDF
External links:
Download fulltextURL
Download fulltextRelated documents
Rate this document:

Rate this document:
1
2
3
 
(Not yet reviewed)