Self-supervised learning of Split Invariant Equivariant representations

Garrido, Quentin; Najman, Laurent; Lecun, Yann

Computer Science > Computer Vision and Pattern Recognition

arXiv:2302.10283 (cs)

[Submitted on 14 Feb 2023 (v1), last revised 19 Jun 2023 (this version, v2)]

Title:Self-supervised learning of Split Invariant Equivariant representations

Authors:Quentin Garrido (FAIR, LIGM), Laurent Najman (LIGM), Yann Lecun (FAIR, CIMS)

View PDF

Abstract:Recent progress has been made towards learning invariant or equivariant representations with self-supervised learning. While invariant methods are evaluated on large scale datasets, equivariant ones are evaluated in smaller, more controlled, settings. We aim at bridging the gap between the two in order to learn more diverse representations that are suitable for a wide range of tasks. We start by introducing a dataset called 3DIEBench, consisting of renderings from 3D models over 55 classes and more than 2.5 million images where we have full control on the transformations applied to the objects. We further introduce a predictor architecture based on hypernetworks to learn equivariant representations with no possible collapse to invariance. We introduce SIE (Split Invariant-Equivariant) which combines the hypernetwork-based predictor with representations split in two parts, one invariant, the other equivariant, to learn richer representations. We demonstrate significant performance gains over existing methods on equivariance related tasks from both a qualitative and quantitative point of view. We further analyze our introduced predictor and show how it steers the learned latent space. We hope that both our introduced dataset and approach will enable learning richer representations without supervision in more complex scenarios. Code and data are available at this https URL.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2302.10283 [cs.CV]
	(or arXiv:2302.10283v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2302.10283
Journal reference:	The Fortieth International Conference on Machine Learning, 2023, Honolulu, United States

Submission history

From: Quentin Garrido [view email] [via CCSD proxy]
[v1] Tue, 14 Feb 2023 07:53:18 UTC (7,779 KB)
[v2] Mon, 19 Jun 2023 12:21:08 UTC (6,147 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Self-supervised learning of Split Invariant Equivariant representations

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Self-supervised learning of Split Invariant Equivariant representations

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators