Applied Federated Learning: Architectural Design for Robust and Efficient Learning in Privacy Aware Settings

Stojkovic, Branislav; Woodbridge, Jonathan; Fang, Zhihan; Cai, Jerry; Petrov, Andrey; Iyer, Sathya; Huang, Daoyu; Yau, Patrick; Kumar, Arvind Sastha; Jawa, Hitesh; Guha, Anamita

Computer Science > Machine Learning

arXiv:2206.00807v2 (cs)

[Submitted on 2 Jun 2022 (v1), last revised 7 Jun 2022 (this version, v2)]

Title:Applied Federated Learning: Architectural Design for Robust and Efficient Learning in Privacy Aware Settings

Authors:Branislav Stojkovic, Jonathan Woodbridge, Zhihan Fang, Jerry Cai, Andrey Petrov, Sathya Iyer, Daoyu Huang, Patrick Yau, Arvind Sastha Kumar, Hitesh Jawa, Anamita Guha

View PDF

Abstract:The classical machine learning paradigm requires the aggregation of user data in a central location where machine learning practitioners can preprocess data, calculate features, tune models and evaluate performance. The advantage of this approach includes leveraging high performance hardware (such as GPUs) and the ability of machine learning practitioners to do in depth data analysis to improve model performance. However, these advantages may come at a cost to data privacy. User data is collected, aggregated, and stored on centralized servers for model development. Centralization of data poses risks, including a heightened risk of internal and external security incidents as well as accidental data misuse. Federated learning with differential privacy is designed to avoid the server-side centralization pitfall by bringing the ML learning step to users' devices. Learning is done in a federated manner where each mobile device runs a training loop on a local copy of a model. Updates from on-device models are sent to the server via encrypted communication and through differential privacy to improve the global model. In this paradigm, users' personal data remains on their devices. Surprisingly, model training in this manner comes at a fairly minimal degradation in model performance. However, federated learning comes with many other challenges due to its distributed nature, heterogeneous compute environments and lack of data visibility. This paper explores those challenges and outlines an architectural design solution we are exploring and testing to productionize federated learning at Meta scale.

Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2206.00807 [cs.LG]
	(or arXiv:2206.00807v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2206.00807

Submission history

From: Branislav Stojkovic [view email]
[v1] Thu, 2 Jun 2022 00:30:04 UTC (658 KB)
[v2] Tue, 7 Jun 2022 16:35:27 UTC (658 KB)

Computer Science > Machine Learning

Title:Applied Federated Learning: Architectural Design for Robust and Efficient Learning in Privacy Aware Settings

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Applied Federated Learning: Architectural Design for Robust and Efficient Learning in Privacy Aware Settings

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators