Exploiting Contextual Information for Improved Phoneme Recognition

Pinto, Joel Praveen; Hermansky, Hynek; Yegnanarayana, B.; Magimai.-Doss, Mathew

doi:10.1109/ICASSP.2008.4518643

conference paper

Exploiting Contextual Information for Improved Phoneme Recognition

Pinto, Joel Praveen

•

Hermansky, Hynek

•

Yegnanarayana, B.

more

2008

2008 IEEE International Conference on Acoustics, Speech and Signal Processing

In this paper, we investigate the significance of contextual information in a phoneme recognition system using the hidden Markov model - artificial neural network paradigm. Contextual information is probed at the feature level as well as at the output of the multilayerd perceptron. At the feature level, we analyse and compare different methods to model sub-phonemic classes. To exploit the contextual information at the output of the multilayered perceptron, we propose the hierarchical estimation of phoneme posterior probabilities. The best phoneme (excluding silence) recognition accuracy of 73.4% on the TIMIT database is comparable to that of the state-of-the-art systems, but more emphasis is on analysis of the contextual information.

Name

pinto-icassp-phnrecog-2008.pdf

Access type

openaccess

Size

61.79 KB

Format

Adobe PDF

Checksum (MD5)

eed1b52ff399cdcb9b6c9a697dad9338