Speachreading using shape and intensity information

Luettin, Juergen; Thacker, Neil A.; Beet, Steve W.

doi:10.21437/ICSLP.1996-15

Luettin, Juergen; Thacker, Neil A.; Beet, Steve W.

1996

Download

Formats

Format
BibTeX
MARCXML
TextMARC
MARC
DublinCore
EndNote
NLM
RefWorks
RIS

Files

Abstract

We describe a speechreading system that uses both, shape information from the lip contours and intensity information from the mouth area. Shape information is obtained by tracking and parameterising the inner and outer lip boundary in an image sequence. Intensity information is extracted from a grey level model, based on principal component analysis. In comparison to other approaches, the intensity area deforms with the shape model to ensure that similar object features are represented after non-rigid deformation of the lips. We describe speaker independent recognition experiments based on these features and Hidden Markov Models. Preliminary results suggest that similar performance can be achieved by using either shape or intensity information and slightly higher performance by their combined use.

Details

Title Speachreading using shape and intensity information

Author(s) Luettin, Juergen ; Thacker, Neil A. ; Beet, Steve W.

Published in Proceedings of the 4th International Conference on Spoken Language Processing (ICSLP'96)

Volume 1

Pages 58-61

Conference 4th International Conference on Spoken Language Processing (ICSLP'96)

Date 1996

Keywords

vision

DOI https://doi.org/10.21437/ICSLP.1996-15

Additional link URL

Laboratories LIDIAP

Record Appears in Scientific production and competences > STI - School of Engineering > IEM - Institut d'Electricité et de Microtechnique > LIDIAP - L'IDIAP Laboratory
Scientific production and competences > Euler Center for Signal Processing
Conference Papers
Work produced at EPFL
Published

Record creation date 2006-03-10

Files

Abstract

Details

PDF