On Image Auto-Annotation with Latent Space Models

Monay, Florent; Gatica-Perez, Daniel

doi:10.1145/957013.957070

conference paper

On Image Auto-Annotation with Latent Space Models

Monay, Florent

•

Gatica-Perez, Daniel

2003

MULTIMEDIA '03: Proceedings of the eleventh ACM international conference on Multimedia

ACM Int. Conf. on Multimedia (ACM MM)

Image auto-annotation, i.e., the association of words to whole images, has attracted considerable attention. In particular, unsupervised, probabilistic latent variable models of text and image features have shown encouraging results, but their performance with respect to other approaches remains unknown. In this paper, we apply and compare two simple latent space models commonly used in text analysis, namely Latent Semantic Analysis (LSA) and Probabilistic LSA (PLSA). Annotation strategies for each model are discussed. Remarkably, we found that, on a 8000-image dataset, a classic LSA model defined on keywords and a very basic image representation performed as well as much more complex, state-of-the-art methods. Furthermore, non-probabilistic methods (LSA and direct image matching) outperformed PLSA on the same dataset.

Name

monay-acm-sp054.pdf

Access type

openaccess

Size

229.82 KB

Format

Adobe PDF

Checksum (MD5)

7efcbe9d5936e978c1499fa7410937c4