Learning to Answer Questions From Image Using Convolutional Neural Network

06/01/2015
by   Lin Ma, et al.
0

In this paper, we propose to employ the convolutional neural network (CNN) for the image question answering (QA). Our proposed CNN provides an end-to-end framework with convolutional architectures for learning not only the image and question representations, but also their inter-modal interactions to produce the answer. More specifically, our model consists of three CNNs: one image CNN to encode the image content, one sentence CNN to compose the words of the question, and one multimodal convolution layer to learn their joint representation for the classification in the space of candidate answer words. We demonstrate the efficacy of our proposed model on the DAQUAR and COCO-QA datasets, which are two benchmark datasets for the image QA, with the performances significantly outperforming the state-of-the-art.

READ FULL TEXT
research
04/23/2015

Multimodal Convolutional Neural Networks for Matching Image and Sentence

In this paper, we propose multimodal convolutional neural networks (m-CN...
research
11/18/2015

Image Question Answering using Convolutional Neural Network with Dynamic Parameter Prediction

We tackle image question answering (ImageQA) problem by learning a convo...
research
11/30/2019

A Hybrid Approach Towards Two Stage Bengali Question Classification Utilizing Smart Data Balancing Technique

Question classification (QC) is the primary step of the Question Answeri...
research
04/19/2016

M^2S-Net: Multi-Modal Similarity Metric Learning based Deep Convolutional Network for Answer Selection

Recent works using artificial neural networks based on distributed word ...
research
08/28/2018

A Quantum Many-body Wave Function Inspired Language Modeling Approach

The recently proposed quantum language model (QLM) aimed at a principled...
research
11/15/2018

Improving Skin Condition Classification with a Question Answering Model

We present a skin condition classification methodology based on a sequen...
research
08/05/2018

Combining Graph-based Dependency Features with Convolutional Neural Network for Answer Triggering

Answer triggering is the task of selecting the best-suited answer for a ...

Please sign up or login with your details

Forgot password? Click here to reset