Learning Representations on the Unit Sphere: Investigating Angular Gaussian and von Mises-Fisher Distributions for Online Continual Learning

Michel, Nicolas; Chierchia, Giovanni; Negrel, Romain; Bercher, Jean-François

Computer Science > Machine Learning

arXiv:2306.03364v2 (cs)

[Submitted on 6 Jun 2023 (v1), revised 3 Oct 2023 (this version, v2), latest version 16 Feb 2024 (v4)]

Title:Learning Representations on the Unit Sphere: Investigating Angular Gaussian and von Mises-Fisher Distributions for Online Continual Learning

Authors:Nicolas Michel, Giovanni Chierchia, Romain Negrel, Jean-François Bercher

View PDF

Abstract:We use the maximum a posteriori estimation principle for learning representations distributed on the unit sphere. We propose to use the angular Gaussian distribution, which corresponds to a Gaussian projected on the unit-sphere and derive the associated loss function. We also consider the von Mises-Fisher distribution, which is the conditional of a Gaussian in the unit-sphere. The learned representations are pushed toward fixed directions, which are the prior means of the Gaussians; allowing for a learning strategy that is resilient to data drift. This makes it suitable for online continual learning, which is the problem of training neural networks on a continuous data stream, where multiple classification tasks are presented sequentially so that data from past tasks are no longer accessible, and data from the current task can be seen only once. To address this challenging scenario, we propose a memory-based representation learning technique equipped with our new loss functions. Our approach does not require negative data or knowledge of task boundaries and performs well with smaller batch sizes while being computationally efficient. We demonstrate with extensive experiments that the proposed method outperforms the current state-of-the-art methods on both standard evaluation scenarios and realistic scenarios with blurry task boundaries. For reproducibility, we use the same training pipeline for every compared method and share the code at this https URL.

Comments:	17 pages, under review
Subjects:	Machine Learning (cs.LG); Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2306.03364 [cs.LG]
	(or arXiv:2306.03364v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2306.03364

Submission history

From: Nicolas Michel Mr [view email]
[v1] Tue, 6 Jun 2023 02:38:01 UTC (2,165 KB)
[v2] Tue, 3 Oct 2023 20:43:50 UTC (2,265 KB)
[v3] Thu, 5 Oct 2023 07:06:27 UTC (2,265 KB)
[v4] Fri, 16 Feb 2024 17:08:51 UTC (2,306 KB)

Computer Science > Machine Learning

Title:Learning Representations on the Unit Sphere: Investigating Angular Gaussian and von Mises-Fisher Distributions for Online Continual Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Learning Representations on the Unit Sphere: Investigating Angular Gaussian and von Mises-Fisher Distributions for Online Continual Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators