Joint Phoneme Segmentation Inference and Classification using CRFs

Palaz, Dimitri; Magimai.-Doss, Mathew; Collobert, Ronan

doi:10.1109/GlobalSIP.2014.7032185

Palaz, Dimitri; Magimai.-Doss, Mathew; Collobert, Ronan

2014

Download

Formats

Format
BibTeX
MARC
MARCXML
DublinCore
EndNote
NLM
RefWorks
RIS

Files

Abstract

State-of-the-art phoneme sequence recognition systems are based on hybrid hidden Markov model/artificial neural networks (HMM/ANN) framework. In this framework, the local classifier, ANN, is typically trained using Viterbi expectation-maximization algorithm, which involves two separate steps: phoneme sequence segmentation and training of ANN. In this paper, we propose a CRF based phoneme sequence recognition approach that simultaneously infers the phoneme segmentation and classifies the phoneme sequence. More specifically, the phoneme sequence recognition system consists of a local classifier ANN followed by a conditional random field (CRF) whose parameters are trained jointly, using a cost function that discriminates the true phoneme sequence against all competing sequences. In order to efficiently train such a system we introduce a novel CRF based segmentation using acyclic graph. We study the viability of the proposed approach on TIMIT phoneme recognition task. Our studies show that the proposed approach is capable of achieving performance similar to standard hybrid HMM/ANN and ANN/CRF systems where the ANN is trained with manual segmentation.

Details

Title Joint Phoneme Segmentation Inference and Classification using CRFs

Author(s) Palaz, Dimitri ; Magimai.-Doss, Mathew ; Collobert, Ronan

Published in 2014 IEEE Global Conference on Signal and Information Processing (GlobalSIP)

Pages 587-591

Conference Global Conference on Signal and Information Processing

Date 2014

DOI https://doi.org/10.1109/GlobalSIP.2014.7032185

Laboratories LIDIAP

Record Appears in Scientific production and competences > STI - School of Engineering > IEM - Institut d'Electricité et de Microtechnique > LIDIAP - L'IDIAP Laboratory
Scientific production and competences > Euler Center for Signal Processing
Conference Papers
Work produced at EPFL

Record creation date 2014-11-19

Actions

Preview

Select file: