Unsupervised Visual Representation Learning by Online Constrained K-Means

Qian, Qi; Xu, Yuanhong; Hu, Juhua; Li, Hao; **, Rong

Computer Science > Computer Vision and Pattern Recognition

arXiv:2105.11527 (cs)

[Submitted on 24 May 2021 (v1), last revised 28 Mar 2022 (this version, v3)]

Title:Unsupervised Visual Representation Learning by Online Constrained K-Means

Authors:Qi Qian, Yuanhong Xu, Juhua Hu, Hao Li, Rong **

View PDF

Abstract:Cluster discrimination is an effective pretext task for unsupervised representation learning, which often consists of two phases: clustering and discrimination. Clustering is to assign each instance a pseudo label that will be used to learn representations in discrimination. The main challenge resides in clustering since prevalent clustering methods (e.g., k-means) have to run in a batch mode. Besides, there can be a trivial solution consisting of a dominating cluster. To address these challenges, we first investigate the objective of clustering-based representation learning. Based on this, we propose a novel clustering-based pretext task with online \textbf{Co}nstrained \textbf{K}-m\textbf{e}ans (\textbf{CoKe}). Compared with the balanced clustering that each cluster has exactly the same size, we only constrain the minimal size of each cluster to flexibly capture the inherent data structure. More importantly, our online assignment method has a theoretical guarantee to approach the global optimum. By decoupling clustering and discrimination, CoKe can achieve competitive performance when optimizing with only a single view from each instance. Extensive experiments on ImageNet and other benchmark data sets verify both the efficacy and efficiency of our proposal. Code is available at \url{this https URL}.

Comments:	accepted by CVPR'22
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2105.11527 [cs.CV]
	(or arXiv:2105.11527v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2105.11527

Submission history

From: Qi Qian [view email]
[v1] Mon, 24 May 2021 20:38:32 UTC (189 KB)
[v2] Mon, 27 Dec 2021 18:37:45 UTC (117 KB)
[v3] Mon, 28 Mar 2022 20:15:05 UTC (147 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Unsupervised Visual Representation Learning by Online Constrained K-Means

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Unsupervised Visual Representation Learning by Online Constrained K-Means

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators