Unseen Object Segmentation in Videos via Transferable Representations

01/08/2019
by   Yi-Wen Chen, et al.
0

In order to learn object segmentation models in videos, conventional methods require a large amount of pixel-wise ground truth annotations. However, collecting such supervised data is time-consuming and labor-intensive. In this paper, we exploit existing annotations in source images and transfer such visual information to segment videos with unseen object categories. Without using any annotations in the target video, we propose a method to jointly mine useful segments and learn feature representations that better adapt to the target frames. The entire process is decomposed into two tasks: 1) solving a submodular function for selecting object-like segments, and 2) learning a CNN model with a transferable module for adapting seen categories in the source domain to the unseen target video. We present an iterative update scheme between two tasks to self-learn the final solution for object segmentation. Experimental results on numerous benchmark datasets show that the proposed method performs favorably against the state-of-the-art algorithms.

READ FULL TEXT

page 12

page 14

research
01/16/2019

Domain Adaptation for Structured Output via Discriminative Representations

Predicting structured outputs such as semantic segmentation relies on ex...
research
01/16/2019

Domain Adaptation for Structured Output via Discriminative Patch Representations

Predicting structured outputs such as semantic segmentation relies on ex...
research
02/20/2019

Learning Transferable Self-attentive Representations for Action Recognition in Untrimmed Videos with Weak Supervision

Action recognition in videos has attracted a lot of attention in the pas...
research
06/06/2018

Fast and Accurate Online Video Object Segmentation via Tracking Parts

Online video object segmentation is a challenging task as it entails to ...
research
11/04/2022

Domain Adaptive Video Semantic Segmentation via Cross-Domain Moving Object Mixing

The network trained for domain adaptation is prone to bias toward the ea...
research
08/13/2019

Frame-to-Frame Aggregation of Active Regions in Web Videos for Weakly Supervised Semantic Segmentation

When a deep neural network is trained on data with only image-level labe...
research
06/10/2022

Learning self-calibrated optic disc and cup segmentation from multi-rater annotations

The segmentation of optic disc(OD) and optic cup(OC) from fundus images ...

Please sign up or login with your details

Forgot password? Click here to reset