Learning Semantics for Image Annotation

05/15/2017
by   Amara Tariq, et al.
0

Image search and retrieval engines rely heavily on textual annotation in order to match word queries to a set of candidate images. A system that can automatically annotate images with meaningful text can be highly beneficial for such engines. Currently, the approaches to develop such systems try to establish relationships between keywords and visual features of images. In this paper, We make three main contributions to this area: (i) We transform this problem from the low-level keyword space to the high-level semantics space that we refer to as the " image theme", (ii) Instead of treating each possible keyword independently, we use latent Dirichlet allocation to learn image themes from the associated texts in a training phase. Images are then annotated with image themes rather than keywords, using a modified continuous relevance model, which takes into account the spatial coherence and the visual continuity among images of common theme. (iii) To achieve more coherent annotations among images of common theme, we have integrated ConceptNet in learning the semantics of images, and hence augment image descriptions beyond annotations provided by humans. Images are thus further annotated by a few most significant words of the prominent image theme. Our extensive experiments show that a coherent theme-based image annotation using high-level semantics results in improved precision and recall as compared with equivalent classical keyword annotation systems.

READ FULL TEXT

page 10

page 11

research
06/20/2013

Analysing Word Importance for Image Annotation

Image annotation provides several keywords automatically for a given ima...
research
01/26/2020

An Effective Automatic Image Annotation Model Via Attention Model and Data Equilibrium

Nowadays, a huge number of images are available. However, retrieving a r...
research
05/06/2017

Image Annotation using Multi-Layer Sparse Coding

Automatic annotation of images with descriptive words is a challenging p...
research
12/06/2018

Representing pictures with emotions

Modern research in content-based image retrieval systems (CIBR) has beco...
research
12/18/2012

A Multi-View Embedding Space for Modeling Internet Images, Tags, and their Semantics

This paper investigates the problem of modeling Internet images and asso...
research
05/27/2019

Dynamically Visual Disambiguation of Keyword-based Image Search

Due to the high cost of manual annotation, learning directly from the we...
research
09/23/2022

Best Prompts for Text-to-Image Models and How to Find Them

Recent progress in generative models, especially in text-guided diffusio...

Please sign up or login with your details

Forgot password? Click here to reset