Summarize and Search: Learning Consensus-aware Dynamic Convolution for Co-Saliency Detection

10/01/2021
by   Ni Zhang, et al.
0

Humans perform co-saliency detection by first summarizing the consensus knowledge in the whole group and then searching corresponding objects in each image. Previous methods usually lack robustness, scalability, or stability for the first process and simply fuse consensus features with image features for the second process. In this paper, we propose a novel consensus-aware dynamic convolution model to explicitly and effectively perform the "summarize and search" process. To summarize consensus image features, we first summarize robust features for every single image using an effective pooling method and then aggregate cross-image consensus cues via the self-attention mechanism. By doing this, our model meets the scalability and stability requirements. Next, we generate dynamic kernels from consensus features to encode the summarized consensus knowledge. Two kinds of kernels are generated in a supplementary way to summarize fine-grained image-specific consensus object cues and the coarse group-wise common knowledge, respectively. Then, we can effectively perform object searching by employing dynamic convolution at multiple scales. Besides, a novel and effective data synthesis method is also proposed to train our network. Experimental results on four benchmark datasets verify the effectiveness of our proposed method. Our code and saliency maps are available at <https://github.com/nnizhang/CADC>.

READ FULL TEXT

page 3

page 6

page 7

page 8

research
04/28/2020

Gradient-Induced Co-Saliency Detection

Co-saliency detection (Co-SOD) aims to segment the common salient foregr...
research
04/30/2023

Discriminative Co-Saliency and Background Mining Transformer for Co-Salient Object Detection

Most previous co-salient object detection works mainly focus on extracti...
research
07/05/2022

SESS: Saliency Enhancing with Scaling and Sliding

High-quality saliency maps are essential in several machine learning app...
research
12/09/2020

DS-Net: Dynamic Spatiotemporal Network for Video Salient Object Detection

As moving objects always draw more attention of human eyes, the temporal...
research
10/09/2022

CAGroup3D: Class-Aware Grouping for 3D Object Detection on Point Clouds

We present a novel two-stage fully sparse convolutional 3D object detect...
research
09/08/2019

AtLoc: Attention Guided Camera Localization

Deep learning has achieved impressive results in camera localization, bu...
research
06/10/2022

Out of Sight, Out of Mind: A Source-View-Wise Feature Aggregation for Multi-View Image-Based Rendering

To estimate the volume density and color of a 3D point in the multi-view...

Please sign up or login with your details

Forgot password? Click here to reset