Few-Shot Learning for Video Object Detection in a Transfer-Learning Scheme

03/26/2021
by   Zhongjie Yu, et al.
0

Different from static images, videos contain additional temporal and spatial information for better object detection. However, it is costly to obtain a large number of videos with bounding box annotations that are required for supervised deep learning. Although humans can easily learn to recognize new objects by watching only a few video clips, deep learning usually suffers from overfitting. This leads to an important question: how to effectively learn a video object detector from only a few labeled video clips? In this paper, we study the new problem of few-shot learning for video object detection. We first define the few-shot setting and create a new benchmark dataset for few-shot video object detection derived from the widely used ImageNet VID dataset. We employ a transfer-learning framework to effectively train the video object detector on a large number of base-class objects and a few video clips of novel-class objects. By analyzing the results of two methods under this framework (Joint and Freeze) on our designed weak and strong base datasets, we reveal insufficiency and overfitting problems. A simple but effective method, called Thaw, is naturally developed to trade off the two problems and validate our analysis. Extensive experiments on our proposed benchmark datasets with different scenarios demonstrate the effectiveness of our novel analysis in this new few-shot video object detection problem.

READ FULL TEXT

page 4

page 8

research
04/30/2021

Few-Shot Video Object Detection

We introduce Few-Shot Video Object Detection (FSVOD) with three importan...
research
05/20/2021

Generalized Few-Shot Object Detection without Forgetting

Recently few-shot object detection is widely adopted to deal with data-l...
research
10/22/2020

Restoring Negative Information in Few-Shot Object Detection

Few-shot learning has recently emerged as a new challenge in the deep le...
research
10/12/2022

BoxMask: Revisiting Bounding Box Supervision for Video Object Detection

We present a new, simple yet effective approach to uplift video object d...
research
10/06/2020

Representation learning from videos in-the-wild: An object-centric approach

We propose a method to learn image representations from uncurated videos...
research
12/20/2022

Bridging Images and Videos: A Simple Learning Framework for Large Vocabulary Video Object Detection

Scaling object taxonomies is one of the important steps toward a robust ...
research
08/03/2021

ODIP: Towards Automatic Adaptation for Object Detection by Interactive Perception

Object detection plays a deep role in visual systems by identifying inst...

Please sign up or login with your details

Forgot password? Click here to reset