Semantic-Enhanced Image Clustering

Shaotian Cai,Xiaojun Chen,Longteng Chen,Qin Zhang,Liping Qiu

doi:10.1609/aaai.v37i6.25841

Abstract

Image clustering is an important and open challenging task in computer vision. Although many methods have been proposed to solve the image clustering task, they only explore images and uncover clusters according to the image features, thus being unable to distinguish visually similar but semantically different images. In this paper, we propose to investigate the task of image clustering with the help of visual-language pre-training model. Different from the zero-shot setting, in which the class names are known, we only know the number of clusters in this setting. Therefore, how to map images to a proper semantic space and how to cluster images from both image and semantic spaces are two key problems. To solve the above problems, we propose a novel image clustering method guided by the visual-language pre-training model CLIP, named Semantic-Enhanced Image Clustering (SIC). In this new method, we propose a method to map the given images to a proper semantic space first and efficient methods to generate pseudo-labels according to the relationships between images and semantics. Finally, we propose to perform clustering with consistency learning in both image space and semantic space, in a self-supervised learning fashion. The theoretical result of convergence analysis shows that our proposed method can converge at a sublinear speed. Theoretical analysis of expectation risk also shows that we can reduce the expectation risk by improving neighborhood consistency, increasing prediction confidence, or reducing neighborhood imbalance. Experimental results on five benchmark datasets clearly show the superiority of our new method.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Semantic-Enhanced Image Clustering

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence

Lead the way for us

Journal: Proceedings of the AAAI Conference on Artificial Intelligence	Publication Date: Jun 26, 2023
Citations: 2

Similar Papers

Biases in human number estimation are well-described by clustering algorithms from computer vision
H Y Im ... S.-H Zhong
Journal of Vision | VOL. 13
H Y Im, et. al.H Y Im ... S.-H Zhong
25 Jul 2013
Journal of Vision | VOL. 13

Improve Semantic Correspondence by Filtering the Correlation Scores in both Image Space and Hough Space
Shihua Xiong ... Yonggang Lu
-
Shihua Xiong, et. al.Shihua Xiong ... Yonggang Lu
01 Jan 2020
01 Jan 2020

Semantic Embedding Guided Attention with Explicit Visual Feature Fusion for Video Captioning
Shanshan Dong ... Xinshun Xu
ACM Transactions on Multimedia Computing, Communications, and Applications | VOL. 19
Shanshan Dong, et. al.Shanshan Dong ... Xinshun Xu
06 Feb 2023
ACM Transactions on Multimedia Computing, Communications, and Applications | VOL. 19

Effective image clustering using self-organizing migrating algorithm
Seyed Jalaleddin Mousavirad ... Gerald Schaefer
-
Seyed Jalaleddin Mousavirad, et. al.Seyed Jalaleddin Mousavirad ... Gerald Schaefer
08 Jul 2020
08 Jul 2020

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Semantic-Enhanced Image Clustering

Abstract

Talk to us

Similar Papers

More From: Proceedings of the AAAI Conference on Artificial Intelligence