Relation Distillation Networks for Video Object Detection

08/26/2019
by   Jiajun Deng, et al.
0

It has been well recognized that modeling object-to-object relations would be helpful for object detection. Nevertheless, the problem is not trivial especially when exploring the interactions between objects to boost video object detectors. The difficulty originates from the aspect that reliable object relations in a video should depend on not only the objects in the present frame but also all the supportive objects extracted over a long range span of the video. In this paper, we introduce a new design to capture the interactions across the objects in spatio-temporal context. Specifically, we present Relation Distillation Networks (RDN) --- a new architecture that novelly aggregates and propagates object relation to augment object features for detection. Technically, object proposals are first generated via Region Proposal Networks (RPN). RDN then, on one hand, models object relation via multi-stage reasoning, and on the other, progressively distills relation through refining supportive object proposals with high objectness scores in a cascaded manner. The learnt relation verifies the efficacy on both improving object detection in each frame and box linking across frames. Extensive experiments are conducted on ImageNet VID dataset, and superior results are reported when comparing to state-of-the-art methods. More remarkably, our RDN achieves 81.8 When further equipped with linking and rescoring, we obtain to-date the best reported mAP of 83.8

READ FULL TEXT
research
03/31/2020

Long Short-Term Relation Networks for Video Action Detection

It has been well recognized that modeling human-object or object-object ...
research
07/15/2021

What and When to Look?: Temporal Span Proposal Network for Video Visual Relation Detection

Identifying relations between objects is central to understanding the sc...
research
01/30/2018

Object Detection in Videos by Short and Long Range Object Linking

We address the problem of detecting objects in videos with the interest ...
research
08/18/2021

Social Fabric: Tubelet Compositions for Video Relation Detection

This paper strives to classify and detect the relationship between objec...
research
01/20/2021

Video Relation Detection with Trajectory-aware Multi-modal Features

Video relation detection problem refers to the detection of the relation...
research
11/30/2017

Relation Networks for Object Detection

Although it is well believed for years that modeling relations between o...
research
10/23/2020

Object-aware Feature Aggregation for Video Object Detection

We present an Object-aware Feature Aggregation (OFA) module for video ob...

Please sign up or login with your details

Forgot password? Click here to reset