Trip-ROMA: Self-Supervised Learning with Triplets and Random Mappings

Li, Wenbin; Yang, Xuesong; Kong, Meihao; Wang, Lei; Huo, Jing; Gao, Yang; Luo, Jiebo

Computer Science > Computer Vision and Pattern Recognition

arXiv:2107.10419 (cs)

[Submitted on 22 Jul 2021 (v1), last revised 24 Aug 2023 (this version, v3)]

Title:Trip-ROMA: Self-Supervised Learning with Triplets and Random Mappings

Authors:Wenbin Li, Xuesong Yang, Meihao Kong, Lei Wang, Jing Huo, Yang Gao, Jiebo Luo

View PDF

Abstract:Contrastive self-supervised learning (SSL) methods, such as MoCo and SimCLR, have achieved great success in unsupervised visual representation learning. They rely on a large number of negative pairs and thus require either large memory banks or large batches. Some recent non-contrastive SSL methods, such as BYOL and SimSiam, attempt to discard negative pairs and have also shown remarkable performance. To avoid collapsed solutions caused by not using negative pairs, these methods require non-trivial asymmetry designs. However, in small data regimes, we can not obtain a sufficient number of negative pairs or effectively avoid the over-fitting problem when negatives are not used at all. To address this situation, we argue that negative pairs are still important but one is generally sufficient for each positive pair. We show that a simple Triplet-based loss (Trip) can achieve surprisingly good performance without requiring large batches or asymmetry designs. Moreover, to alleviate the over-fitting problem in small data regimes and further enhance the effect of Trip, we propose a simple plug-and-play RandOm MApping (ROMA) strategy by randomly mapping samples into other spaces and requiring these randomly projected samples to satisfy the same relationship indicated by the triplets. Integrating the triplet-based loss with random mapping, we obtain the proposed method Trip-ROMA. Extensive experiments, including unsupervised representation learning and unsupervised few-shot learning, have been conducted on ImageNet-1K and seven small datasets. They successfully demonstrate the effectiveness of Trip-ROMA and consistently show that ROMA can further effectively boost other SSL methods. Code is available at this https URL.

Comments:	Accepted to Transactions on Machine Learning Research (TMLR) 2023
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2107.10419 [cs.CV]
	(or arXiv:2107.10419v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2107.10419

Submission history

From: Wenbin Li [view email]
[v1] Thu, 22 Jul 2021 02:06:38 UTC (1,107 KB)
[v2] Tue, 31 Aug 2021 06:35:35 UTC (1,109 KB)
[v3] Thu, 24 Aug 2023 03:09:41 UTC (1,780 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Trip-ROMA: Self-Supervised Learning with Triplets and Random Mappings

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Trip-ROMA: Self-Supervised Learning with Triplets and Random Mappings

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators