Towards Accurate Pixel-wise Object Tracking by Attention Retrieval

08/06/2020
by   Zhipeng Zhang, et al.
9

The encoding of the target in object tracking moves from the coarse bounding-box to fine-grained segmentation map recently. Revisiting de facto real-time approaches that are capable of predicting mask during tracking, we observed that they usually fork a light branch from the backbone network for segmentation. Although efficient, directly fusing backbone features without considering the negative influence of background clutter tends to introduce false-negative predictions, lagging the segmentation accuracy. To mitigate this problem, we propose an attention retrieval network (ARN) to perform soft spatial constraints on backbone features. We first build a look-up-table (LUT) with the ground-truth mask in the starting frame, and then retrieves the LUT to obtain an attention map for spatial constraints. Moreover, we introduce a multi-resolution multi-stage segmentation network (MMS) to further weaken the influence of background clutter by reusing the predicted mask to filter backbone features. Our approach set a new state-of-the-art on recent pixel-wise object tracking benchmark VOT2020 while running at 40 fps. Notably, the proposed model surpasses SiamMask by 11.7/4.2/5.5 points on VOT2020, DAVIS2016, and DAVIS2017, respectively. We will release our code at https://github.com/researchmm/TracKit.

READ FULL TEXT

page 4

page 8

research
11/20/2017

Pixel-wise object tracking

In this paper, we propose a novel pixel-wise visual object tracking fram...
research
09/21/2020

Discriminative Segmentation Tracking Using Dual Memory Banks

Existing template-based trackers usually localize the target in each fra...
research
05/08/2020

TSDM: Tracking by SiamRPN++ with a Depth-refiner and a Mask-generator

In a generic object tracking, depth (D) information provides informative...
research
08/25/2023

Integrating Boxes and Masks: A Multi-Object Framework for Unified Visual Tracking and Segmentation

Tracking any given object(s) spatially and temporally is a common purpos...
research
12/07/2022

BoxPolyp:Boost Generalized Polyp Segmentation Using Extra Coarse Bounding Box Annotations

Accurate polyp segmentation is of great importance for colorectal cancer...
research
07/04/2020

Alpha-Refine: Boosting Tracking Performance by Precise Bounding Box Estimation

In recent years, the multiple-stage strategy has become a popular trend ...
research
05/03/2018

Visual Object Tracking: The Initialisation Problem

Model initialisation is an important component of object tracking. Track...

Please sign up or login with your details

Forgot password? Click here to reset