Demystifying Contrastive Self-Supervised Learning: Invariances, Augmentations and Dataset Biases

Purushwalkam, Senthil; Gupta, Abhinav

Computer Science > Computer Vision and Pattern Recognition

arXiv:2007.13916 (cs)

[Submitted on 28 Jul 2020 (v1), last revised 29 Jul 2020 (this version, v2)]

Title:Demystifying Contrastive Self-Supervised Learning: Invariances, Augmentations and Dataset Biases

Authors:Senthil Purushwalkam, Abhinav Gupta

View PDF

Abstract:Self-supervised representation learning approaches have recently surpassed their supervised learning counterparts on downstream tasks like object detection and image classification. Somewhat mysteriously the recent gains in performance come from training instance classification models, treating each image and it's augmented versions as samples of a single class. In this work, we first present quantitative experiments to demystify these gains. We demonstrate that approaches like MOCO and PIRL learn occlusion-invariant representations. However, they fail to capture viewpoint and category instance invariance which are crucial components for object recognition. Second, we demonstrate that these approaches obtain further gains from access to a clean object-centric training dataset like Imagenet. Finally, we propose an approach to leverage unstructured videos to learn representations that possess higher viewpoint invariance. Our results show that the learned representations outperform MOCOv2 trained on the same data in terms of invariances encoded and the performance on downstream image classification and semantic segmentation tasks.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2007.13916 [cs.CV]
	(or arXiv:2007.13916v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2007.13916

Submission history

From: Senthil Purushwalkam [view email]
[v1] Tue, 28 Jul 2020 00:11:31 UTC (161 KB)
[v2] Wed, 29 Jul 2020 05:38:11 UTC (162 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Demystifying Contrastive Self-Supervised Learning: Invariances, Augmentations and Dataset Biases

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Demystifying Contrastive Self-Supervised Learning: Invariances, Augmentations and Dataset Biases

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators