A Comprehensive Empirical Study of Bugs in Open-Source Federated Learning Frameworks

Shao, Weijie; Gao, Yuyang; Song, Fu; Chen, Sen; Fan, Lingling; He, **gZhu

Computer Science > Software Engineering

arXiv:2308.05014 (cs)

[Submitted on 9 Aug 2023 (v1), last revised 6 Oct 2023 (this version, v2)]

Title:A Comprehensive Empirical Study of Bugs in Open-Source Federated Learning Frameworks

Authors:Weijie Shao, Yuyang Gao, Fu Song, Sen Chen, Lingling Fan, **gZhu He

View PDF

Abstract:Federated learning (FL) is a distributed machine learning (ML) paradigm, allowing multiple clients to collaboratively train shared machine learning (ML) models without exposing clients' data privacy. It has gained substantial popularity in recent years, especially since the enforcement of data protection laws and regulations in many countries. To foster the application of FL, a variety of FL frameworks have been proposed, allowing non-experts to easily train ML models. As a result, understanding bugs in FL frameworks is critical for facilitating the development of better FL frameworks and potentially encouraging the development of bug detection, localization and repair tools. Thus, we conduct the first empirical study to comprehensively collect, taxonomize, and characterize bugs in FL frameworks. Specifically, we manually collect and classify 1,119 bugs from all the 676 closed issues and 514 merged pull requests in 17 popular and representative open-source FL frameworks on GitHub. We propose a classification of those bugs into 12 bug symptoms, 12 root causes, and 18 fix patterns. We also study their correlations and distributions on 23 functionalities. We identify nine major findings from our study, discuss their implications and future research directions based on our findings.

Subjects:	Software Engineering (cs.SE); Machine Learning (cs.LG)
Cite as:	arXiv:2308.05014 [cs.SE]
	(or arXiv:2308.05014v2 [cs.SE] for this version)
	https://doi.org/10.48550/arXiv.2308.05014

Submission history

From: Fu Song [view email]
[v1] Wed, 9 Aug 2023 15:14:16 UTC (265 KB)
[v2] Fri, 6 Oct 2023 09:04:19 UTC (1,055 KB)

Computer Science > Software Engineering

Title:A Comprehensive Empirical Study of Bugs in Open-Source Federated Learning Frameworks

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Software Engineering

Title:A Comprehensive Empirical Study of Bugs in Open-Source Federated Learning Frameworks

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators