Automated Image Captioning for Rapid Prototyping and Resource Constrained Environments

06/04/2016
by   Karan Sharma, et al.
0

Significant performance gains in deep learning coupled with the exponential growth of image and video data on the Internet have resulted in the recent emergence of automated image captioning systems. Ensuring scalability of automated image captioning systems with respect to the ever increasing volume of image and video data is a significant challenge. This paper provides a valuable insight in that the detection of a few significant (top) objects in an image allows one to extract other relevant information such as actions (verbs) in the image. We expect this insight to be useful in the design of scalable image captioning systems. We address two parameters by which the scalability of image captioning systems could be quantified, i.e., the traditional algorithmic time complexity which is important given the resource limitations of the user device and the system development time since the programmers' time is a critical resource constraint in many real-world scenarios. Additionally, we address the issue of how word embeddings could be used to infer the verb (action) from the nouns (objects) in a given image in a zero-shot manner. Our results show that it is possible to attain reasonably good performance on predicting actions and captioning images using our approaches with the added advantage of simplicity of implementation.

READ FULL TEXT

page 1

page 6

research
01/31/2022

Deep Learning Approaches on Image Captioning: A Review

Automatic image captioning, which involves describing the contents of an...
research
10/06/2018

A Comprehensive Survey of Deep Learning for Image Captioning

Generating a description of an image is called image captioning. Image c...
research
01/31/2020

iCap: Interative Image Captioning with Predictive Text

In this paper we study a brand new topic of interactive image captioning...
research
09/05/2020

An Efficient Technique for Image Captioning using Deep Neural Network

With the huge expansion of internet and trillions of gigabytes of data g...
research
05/26/2019

A Survey on Biomedical Image Captioning

Image captioning applied to biomedical images can assist and accelerate ...
research
09/19/2020

Misinformation and its stakeholders in Europe: a web-based analysis

The rise of the internet and computational power in recent years allowed...
research
03/01/2017

Evolving Deep Neural Networks

The success of deep learning depends on finding an architecture to fit t...

Please sign up or login with your details

Forgot password? Click here to reset