Video Captioning Using Weak Annotation

09/02/2020
by   Jingyi Hou, et al.
0

Video captioning has shown impressive progress in recent years. One key reason of the performance improvements made by existing methods lie in massive paired video-sentence data, but collecting such strong annotation, i.e., high-quality sentences, is time-consuming and laborious. It is the fact that there now exist an amazing number of videos with weak annotation that only contains semantic concepts such as actions and objects. In this paper, we investigate using weak annotation instead of strong annotation to train a video captioning model. To this end, we propose a progressive visual reasoning method that progressively generates fine sentences from weak annotations by inferring more semantic concepts and their dependency relationships for video captioning. To model concept relationships, we use dependency trees that are spanned by exploiting external knowledge from large sentence corpora. Through traversing the dependency trees, the sentences are generated to train the captioning model. Accordingly, we develop an iterative refinement algorithm that refines sentences via spanning dependency trees and fine-tunes the captioning model using the refined sentences in an alternative training manner. Experimental results demonstrate that our method using weak annotation is very competitive to the state-of-the-art methods using strong annotation.

READ FULL TEXT

page 1

page 3

page 9

research
11/17/2015

Deep Compositional Captioning: Describing Novel Object Categories without Paired Training Data

While recent deep neural network models have achieved promising results ...
research
08/25/2016

Title Generation for User Generated Videos

A great video title describes the most salient event compactly and captu...
research
11/27/2018

Unsupervised Image Captioning

Deep neural networks have achieved great successes on the image captioni...
research
06/04/2019

Relational Reasoning using Prior Knowledge for Visual Captioning

Exploiting relationships among objects has achieved remarkable progress ...
research
07/27/2023

Exploring Annotation-free Image Captioning with Retrieval-augmented Pseudo Sentence Generation

Training an image captioner without annotated image-sentence pairs has g...
research
06/05/2015

Sentence Directed Video Object Codetection

We tackle the problem of video object codetection by leveraging the weak...
research
08/06/2019

Cascaded Revision Network for Novel Object Captioning

Image captioning, a challenging task where the machine automatically des...

Please sign up or login with your details

Forgot password? Click here to reset