How to Collaborate: Towards Maximizing the Generalization Performance in Cross-Silo Federated Learning

Sun, Yuchang; Kountouris, Marios; Zhang, Jun

Computer Science > Machine Learning

arXiv:2401.13236 (cs)

[Submitted on 24 Jan 2024 (v1), last revised 28 Nov 2024 (this version, v2)]

Title:How to Collaborate: Towards Maximizing the Generalization Performance in Cross-Silo Federated Learning

Authors:Yuchang Sun, Marios Kountouris, Jun Zhang

View PDF HTML (experimental)

Abstract:Federated learning (FL) has attracted vivid attention as a privacy-preserving distributed learning framework. In this work, we focus on cross-silo FL, where clients become the model owners after training and are only concerned about the model's generalization performance on their local data. Due to the data heterogeneity issue, asking all the clients to join a single FL training process may result in model performance degradation. To investigate the effectiveness of collaboration, we first derive a generalization bound for each client when collaborating with others or when training independently. We show that the generalization performance of a client can be improved only by collaborating with other clients that have more training data and similar data distribution. Our analysis allows us to formulate a client utility maximization problem by partitioning clients into multiple collaborating groups. A hierarchical clustering-based collaborative training (HCCT) scheme is then proposed, which does not need to fix in advance the number of groups. We further analyze the convergence of HCCT for general non-convex loss functions which unveils the effect of data similarity among clients. Extensive simulations show that HCCT achieves better generalization performance than baseline schemes, whereas it degenerates to independent training and conventional FL in specific scenarios.

Subjects:	Machine Learning (cs.LG); Distributed, Parallel, and Cluster Computing (cs.DC)
Cite as:	arXiv:2401.13236 [cs.LG]
	(or arXiv:2401.13236v2 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2401.13236

Submission history

From: Yuchang Sun [view email]
[v1] Wed, 24 Jan 2024 05:41:34 UTC (824 KB)
[v2] Thu, 28 Nov 2024 13:29:41 UTC (821 KB)

Computer Science > Machine Learning

Title:How to Collaborate: Towards Maximizing the Generalization Performance in Cross-Silo Federated Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:How to Collaborate: Towards Maximizing the Generalization Performance in Cross-Silo Federated Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators