Semantic-enhanced Image Clustering

08/21/2022
by   Shaotian Cai, et al.
0

Image clustering is an important, and open challenge task in computer vision. Although many methods have been proposed to solve the image clustering task, they only explore images and uncover clusters according to the image features, thus are unable to distinguish visually similar but semantically different images. In this paper, we propose to investigate the task of image clustering with the help of visual-language pre-training model. Different from the zero-shot setting in which the class names are known, we only know the number of clusters in this setting. Therefore, how to map images to a proper semantic space and how to cluster images from both image and semantic spaces are two key problems. To solve the above problems, we propose a novel image clustering method guided by the visual-language pre-training model CLIP, named as Semantic-enhanced Image Clustering (SIC). In this new method, we propose a method to map the given images to a proper semantic space first and efficient methods to generate pseudo-labels according to the relationships between images and semantics. Finally, we propose to perform clustering with the consistency learning in both image space and semantic space, in a self-supervised learning fashion. Theoretical result on convergence analysis shows that our proposed method can converge in sublinear speed. Theoretical analysis on expectation risk also shows that we can reduce the expectation risk by improving the neighborhood consistency or prediction confidence or reducing neighborhood imbalance. Experimental results on five benchmark datasets clearly show the superiority of our new method.

READ FULL TEXT

page 1

page 2

page 8

research
12/06/2019

ClusterFit: Improving Generalization of Visual Representations

Pre-training convolutional neural networks with weakly-supervised and se...
research
03/17/2021

SPICE: Semantic Pseudo-labeling for Image Clustering

This paper presents SPICE, a Semantic Pseudo-labeling framework for Imag...
research
12/22/2022

Offline Clustering Approach to Self-supervised Learning for Class-imbalanced Image Data

Class-imbalanced datasets are known to cause the problem of model being ...
research
03/02/2023

Geometric Visual Similarity Learning in 3D Medical Image Self-supervised Pre-training

Learning inter-image similarity is crucial for 3D medical images self-su...
research
08/13/2018

A Transfer Learning based Feature-Weak-Relevant Method for Image Clustering

Image clustering is to group a set of images into disjoint clusters in a...
research
05/22/2014

Self-tuned Visual Subclass Learning with Shared Samples An Incremental Approach

Computer vision tasks are traditionally defined and evaluated using sema...
research
10/24/2021

Image-Based CLIP-Guided Essence Transfer

The conceptual blending of two signals is a semantic task that may under...

Please sign up or login with your details

Forgot password? Click here to reset