login

Automatic Image Annotation Using Auxiliary Text Information

Edinburgh Research ExplorerPublished 1 June 2008Open access
Yansong Feng, Mirella Lapata
Citations69
View PDF

TL;DR

This paper creates a database of pictures that are naturally embedded into news articles and proposes to use their captions as a proxy for annotation keywords, showing that an image annotation model can be developed on this dataset alone without the overhead of manual annotation.

Abstract

The availability of databases of images labeled with keywords is necessary for developing and evaluating image annotation models. Dataset collection is however a costly and time consuming task. In this paper we exploit the vast resource of images available on the web. We create a database of pictures that are naturally embedded into news articles and propose to use their captions as a proxy for annotation

Keywords

Computer ScienceBiochemistry, Genetics and Molecular Biology