References & Citations
Computer Science > Computer Vision and Pattern Recognition
Title: Center-wise Local Image Mixture For Contrastive Representation Learning
(Submitted on 5 Nov 2020 (v1), last revised 18 Oct 2021 (this version, v3))
Abstract: Contrastive learning based on instance discrimination trains model to discriminate different transformations of the anchor sample from other samples, which does not consider the semantic similarity among samples. This paper proposes a new kind of contrastive learning method, named CLIM, which uses positives from other samples in the dataset. This is achieved by searching local similar samples of the anchor, and selecting samples that are closer to the corresponding cluster center, which we denote as center-wise local image selection. The selected samples are instantiated via an data mixture strategy, which performs as a smoothing regularization. As a result, CLIM encourages both local similarity and global aggregation in a robust way, which we find is beneficial for feature representation. Besides, we introduce \emph{multi-resolution} augmentation, which enables the representation to be scale invariant. We reach 75.5% top-1 accuracy with linear evaluation over ResNet-50, and 59.3% top-1 accuracy when fine-tuned with only 1% labels.
Submission history
From: Hao Li [view email][v1] Thu, 5 Nov 2020 08:20:31 GMT (1040kb,D)
[v2] Fri, 27 Nov 2020 09:17:24 GMT (2105kb,D)
[v3] Mon, 18 Oct 2021 02:15:36 GMT (1702kb,D)
Link back to: arXiv, form interface, contact.