On Automatic Annotation of Images with Latent Space Models

Monay, Florent; Gatica-Perez, Daniel

report

Monay, Florent

•

Gatica-Perez, Daniel

2003

Image auto-annotation, i.e., the association of words to whole images, has attracted considerable attention. In particular, unsupervised, probabilistic latent variable models of text and image features have shown encouraging results, but their performance with respect to other approaches remains unknown. In this paper, we apply and compare two simple latent space models commonly used in text analysis, namely Latent Semantic Analysis (LSA) and Probabilistic LSA (PLSA). Annotation strategies for each model are discussed. Remarkably, we found that, on a 8000-image dataset, a classic LSA model defined on keywords and a very basic image representation performed as well as much more complex, state-of-the-art methods. Furthermore, non-probabilistic methods (LSA and direct image matching) outperformed PLSA on the same dataset.

Name

rr03-31.pdf

Access type

openaccess

Size

351.65 KB

Format

Adobe PDF

Checksum (MD5)

9a5071fb6cb87016eb7b9e7648c0372c