Bias in Motion: Theoretical Insights into the Dynamics of Bias in SGD Training

Jain, Anchit; Nobahari, Rozhin; Baratin, Aristide; Mannelli, Stefano Sarao

Computer Science > Machine Learning

arXiv:2405.18296 (cs)

[Submitted on 28 May 2024]

Title:Bias in Motion: Theoretical Insights into the Dynamics of Bias in SGD Training

Authors:Anchit Jain, Rozhin Nobahari, Aristide Baratin, Stefano Sarao Mannelli

View PDF HTML (experimental)

Abstract:Machine learning systems often acquire biases by leveraging undesired features in the data, impacting accuracy variably across different sub-populations. Current understanding of bias formation mostly focuses on the initial and final stages of learning, leaving a gap in knowledge regarding the transient dynamics. To address this gap, this paper explores the evolution of bias in a teacher-student setup modeling different data sub-populations with a Gaussian-mixture model. We provide an analytical description of the stochastic gradient descent dynamics of a linear classifier in this setting, which we prove to be exact in high dimension. Notably, our analysis reveals how different properties of sub-populations influence bias at different timescales, showing a shifting preference of the classifier during training. Applying our findings to fairness and robustness, we delineate how and when heterogeneous data and spurious features can generate and amplify bias. We empirically validate our results in more complex scenarios by training deeper networks on synthetic and real datasets, including CIFAR10, MNIST, and CelebA.

Subjects:	Machine Learning (cs.LG); Disordered Systems and Neural Networks (cond-mat.dis-nn); Machine Learning (stat.ML)
Cite as:	arXiv:2405.18296 [cs.LG]
	(or arXiv:2405.18296v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2405.18296

Submission history

From: Anchit Jain [view email]
[v1] Tue, 28 May 2024 15:50:10 UTC (1,133 KB)

Computer Science > Machine Learning

Title:Bias in Motion: Theoretical Insights into the Dynamics of Bias in SGD Training

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Bias in Motion: Theoretical Insights into the Dynamics of Bias in SGD Training

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators