Communication-Efficient Federated Learning through Adaptive Weight Clustering and Server-Side Distillation (2401.14211v3)

Published 25 Jan 2024 in cs.LG and cs.DC

Abstract: Federated Learning (FL) is a promising technique for the collaborative training of deep neural networks across multiple devices while preserving data privacy. Despite its potential benefits, FL is hindered by excessive communication costs due to repeated server-client communication during training. To address this challenge, model compression techniques, such as sparsification and weight clustering are applied, which often require modifying the underlying model aggregation schemes or involve cumbersome hyperparameter tuning, with the latter not only adjusts the model's compression rate but also limits model's potential for continuous improvement over growing data. In this paper, we propose FedCompress, a novel approach that combines dynamic weight clustering and server-side knowledge distillation to reduce communication costs while learning highly generalizable models. Through a comprehensive evaluation on diverse public datasets, we demonstrate the efficacy of our approach compared to baselines in terms of communication costs and inference speed.

References (24)

Authors (4)

Vasileios Tsouvalas (10 papers)
Aaqib Saeed (36 papers)
Tanir Ozcelebi (14 papers)
Nirvana Meratnia (9 papers)

Citations (4)

View on Semantic Scholar

Summary

We haven't generated a summary for this paper yet.

Summarize Now

Tweets

https://twitter.com/HPCPapers/status/1752210639328526646

Communication-Efficient Federated Learning through Adaptive Weight Clustering and Server-Side Distillation (2401.14211v3)

Summary

Related Papers

Tweets