3D-Augmented Contrastive Knowledge Distillation for Image-based Object Pose Estimation

06/02/2022
by   Zhidan Liu, et al.
0

Image-based object pose estimation sounds amazing because in real applications the shape of object is oftentimes not available or not easy to take like photos. Although it is an advantage to some extent, un-explored shape information in 3D vision learning problem looks like "flaws in jade". In this paper, we deal with the problem in a reasonable new setting, namely 3D shape is exploited in the training process, and the testing is still purely image-based. We enhance the performance of image-based methods for category-agnostic object pose estimation by exploiting 3D knowledge learned by a multi-modal method. Specifically, we propose a novel contrastive knowledge distillation framework that effectively transfers 3D-augmented image representation from a multi-modal model to an image-based model. We integrate contrastive learning into the two-stage training procedure of knowledge distillation, which formulates an advanced solution to combine these two approaches for cross-modal tasks. We experimentally report state-of-the-art results compared with existing category-agnostic image-based methods by a large margin (up to +5 on ObjectNet3D dataset), demonstrating the effectiveness of our method.

READ FULL TEXT

page 1

page 7

research
05/30/2022

Knowledge Distillation for 6D Pose Estimation by Keypoint Distribution Alignment

Knowledge distillation facilitates the training of a compact student net...
research
01/18/2022

Cross-modal Contrastive Distillation for Instructional Activity Anticipation

In this study, we aim to predict the plausible future action steps given...
research
05/28/2019

Probabilistic Category-Level Pose Estimation via Segmentation and Predicted-Shape Priors

We introduce a new method for category-level pose estimation which produ...
research
05/12/2021

PoseContrast: Class-Agnostic Object Viewpoint Estimation in the Wild with Pose-Aware Contrastive Learning

Motivated by the need of estimating the pose (viewpoint) of arbitrary ob...
research
09/22/2021

KD-VLP: Improving End-to-End Vision-and-Language Pretraining with Object Knowledge Distillation

Self-supervised vision-and-language pretraining (VLP) aims to learn tran...
research
05/27/2023

An Image Based Visual Servo Method for Probe-and-Drogue Autonomous Aerial Refueling

With the high focus on autonomous aerial refueling recently, it becomes ...
research
04/20/2018

An Approximate Shading Model with Detail Decomposition for Object Relighting

We present an object relighting system that allows an artist to select a...

Please sign up or login with your details

Forgot password? Click here to reset