Repository logo

Infoscience

  • English
  • French
Log In
Logo EPFL, École polytechnique fédérale de Lausanne

Infoscience

  • English
  • French
Log In
  1. Home
  2. Academic and Research Output
  3. Conferences, Workshops, Symposiums, and Seminars
  4. Front-end for Far-field Speech Recognition based on Frequency Domain Linear Prediction
 
conference paper

Front-end for Far-field Speech Recognition based on Frequency Domain Linear Prediction

Ganapathy, Sriram  
•
Thomas, Samuel
•
Hermansky, Hynek  
2008
Interspeech 2008
Interspeech 2008

Automatic Speech Recognition (ASR) systems usually fail when they encounter speech from far-field microphone in reverberant environments. This is due to the application of short-term feature extraction techniques which do not compensate for the artifacts introduced by long room impulse responses. In this paper, we propose a front-end, based on Frequency Domain Linear Prediction (FDLP), that tries to remove reverberation artifacts present in far-field speech. Long temporal segments of far-field speech are analyzed in narrow frequency sub-bands to extract FDLP envelopes and residual signals. Filtering the residual signals with gain normalized inverse FDLP filters result in a set of sub-band signals which are synthesized to reconstruct the signal back. ASR experiments on far-field speech data processed by the proposed front-end show significant improvements (relative reduction of $30 %$ in word error rate) compared to other robust feature extraction techniques.

  • Files
  • Details
  • Metrics
Loading...
Thumbnail Image
Name

tsamuel-interspeech-1-2008.pdf

Access type

openaccess

Size

375.29 KB

Format

Adobe PDF

Checksum (MD5)

304ffb9447521f531bb1023d13e832ae

Logo EPFL, École polytechnique fédérale de Lausanne
  • Contact
  • infoscience@epfl.ch

  • Follow us on Facebook
  • Follow us on Instagram
  • Follow us on LinkedIn
  • Follow us on X
  • Follow us on Youtube
AccessibilityLegal noticePrivacy policyCookie settingsEnd User AgreementGet helpFeedback

Infoscience is a service managed and provided by the Library and IT Services of EPFL. © EPFL, tous droits réservés