Abstract:Image clustering is an important and open-challenging task in computer vision. Although many methods have been proposed to solve the image clustering task, they only explore images and uncover clusters according to the image features, thus being unable to distinguish visually similar but semantically different images. In this paper, we propose to investigate the task of image clustering with the help of a visual-language pre-training model. Different from the zero-shot setting, in which the class names are known, we only know the number of clusters in this setting. Therefore, how to map images to a proper semantic space and how to cluster images from both image and semantic spaces are two key problems. To solve the above problems, we propose a novel image clustering method guided by the visual-language pre-training model CLIP, named \textbf{Semantic-Enhanced Image Clustering (SIC)}. In this new method, we propose a method to map the given images to a proper semantic space first and efficient methods to generate pseudo-labels according to the relationships between images and semantics. Finally, we propose performing clustering with consistency learning in both image space and semantic space, in a self-supervised learning fashion. The theoretical result of convergence analysis shows that our proposed method can converge at a sublinear speed. Theoretical analysis of expectation risk also shows that we can reduce the expected risk by improving neighborhood consistency, increasing prediction confidence, or reducing neighborhood imbalance. Experimental results on five benchmark datasets clearly show the superiority of our new method.

Cluster Correction on Polysemy and Synonymy

Document Clustering Based on Word Sense Cluster

Semantic smoothing of document models for agglomerative clustering

Semantic Smoothing for Model-based Document Clustering

Improving Online Clustering of Chinese Technology Web News with Bag-of-Near-Synonyms

WordNet and Semantic similarity based approach for document clustering

A Clustering Algorithm for Short Documents Based On Concept Similarity

Semantic Correlation Network Based Text Clustering

Constrained Coclustering for Textual Documents.

Co-Clustering With Manifold And Double Sparse Representation

Cross-Lingual Document Clustering Based on Similarity Space Model

Pairwise Constrained Clustering with Group Similarity-Based Patterns

Word Clustering for Collocation-Based Word Sense Disambiguation

Semantic Term "Blurring" and Stochastic "Barcoding" for Improved Unsupervised Text Classification

Clustering Combination Method

Improving Document Clustering by Eliminating Unnatural Language

Regularized bi-directional co-clustering

Document Representation with Statistical Word Senses in Cross-Lingual Document Clustering

A Novel Text Clustering Algorithm Based on Inner Product Space Model of Semantic

Resampling methods for document clustering

Semantic-Enhanced Image Clustering