Fast latent semantic indexing of spoken documents by using self-organizing maps

This paper describes a new latent semantic indexing (LSI) method for spoken audio documents. The framework is indexing broadcast news from radio and TV as a combination of large vocabulary continuous speech recognition (LVCSR), natural language processing (NLP) and information retrieval (IR). For indexing, the documents are presented as vectors of word counts, whose dimensionality is rapidly reduced by random mapping (RM). The obtained vectors are projected into the latent semantic subspace determined by SVD, where the vectors are then smoothed by a self-organizing map (SOM). The smoothing by the closest document clusters is important here, because the documents are often short and have a high word error rate (WER). As the clusters in the semantic subspace reflect the news topics, the SOMs provide an easy way to visualize the index and query results and to explore the database. Test results are reported for TREC's spoken document retrieval databases.


Published in:
Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing ICASSP'2000
Presented at:
Proceedings of the IEEE International Conference on Acoustics, Speech and Signal Processing ICASSP'2000
Year:
2000
Publisher:
Istanbul, Turkey
Keywords:
Note:
IDIAP-RR 99-20
Laboratories:




 Record created 2006-03-10, last modified 2018-01-27

External links:
Download fulltextURL
Download fulltextRelated documents
Download fulltextn/a
Rate this document:

Rate this document:
1
2
3
 
(Not yet reviewed)