Papers
Topics
Authors
Recent
Search
2000 character limit reached

ARIA: On the Interaction Between Architectures, Initialization and Aggregation Methods for Federated Visual Classification

Published 24 Nov 2023 in cs.CV, cs.AI, and cs.DC | (2311.14625v2)

Abstract: Federated Learning (FL) is a collaborative training paradigm that allows for privacy-preserving learning of cross-institutional models by eliminating the exchange of sensitive data and instead relying on the exchange of model parameters between the clients and a server. Despite individual studies on how client models are aggregated, and, more recently, on the benefits of ImageNet pre-training, there is a lack of understanding of the effect the architecture chosen for the federation has, and of how the aforementioned elements interconnect. To this end, we conduct the first joint ARchitecture-Initialization-Aggregation study and benchmark ARIAs across a range of medical image classification tasks. We find that, contrary to current practices, ARIA elements have to be chosen together to achieve the best possible performance. Our results also shed light on good choices for each element depending on the task, the effect of normalisation layers, and the utility of SSL pre-training, pointing to potential directions for designing FL-specific architectures and training pipelines.

Definition Search Book Streamline Icon: https://streamlinehq.com
References (29)
  1. “Federated learning in medicine: facilitating multi-institutional collaborations without sharing patient data,” Scientific reports, vol. 10, no. 1, pp. 12598, 2020.
  2. “Communication-efficient learning of deep networks from decentralized data,” in Artificial intelligence and statistics. PMLR, 2017, pp. 1273–1282.
  3. “Measuring the effects of non-identical data distribution for federated visual classification,” arXiv preprint arXiv:1909.06335, 2019.
  4. “Scaffold: Stochastic controlled averaging for federated learning,” in International conference on machine learning. PMLR, 2020, pp. 5132–5143.
  5. “Handling data heterogeneity via architectural design for federated visual recognition,” in Thirty-seventh Conference on Neural Information Processing Systems, 2023.
  6. “On the importance and applicability of pre-training for federated learning,” in The Eleventh International Conference on Learning Representations, 2022.
  7. “Where to begin? on the impact of pre-training and initialization in federated learning,” arXiv preprint arXiv:2210.08090, 2022.
  8. “Rethinking architecture design for tackling data heterogeneity in federated learning,” in Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, 2022, pp. 10061–10071.
  9. “Battle of the backbones: A large-scale comparison of pretrained models across computer vision tasks,” arXiv preprint arXiv:2310.19909, 2023.
  10. “Identity mappings in deep residual networks,” in Computer Vision–ECCV 2016: 14th European Conference, Amsterdam, The Netherlands, October 11–14, 2016, Proceedings, Part IV 14. Springer, 2016, pp. 630–645.
  11. “Wide residual networks,” arXiv preprint arXiv:1605.07146, 2016.
  12. “Densely connected convolutional networks,” in Proceedings of the IEEE conference on computer vision and pattern recognition, 2017, pp. 4700–4708.
  13. “Making batch normalization great in federated deep learning,” arXiv preprint arXiv:2303.06530, 2023.
  14. “Characterizing signal propagation to close the performance gap in unnormalized resnets,” arXiv preprint arXiv:2101.08692, 2021.
  15. “Efficientnetv2: Smaller models and faster training,” in International conference on machine learning. PMLR, 2021, pp. 10096–10106.
  16. “A convnet for the 2020s,” in Proceedings of the IEEE/CVF conference on computer vision and pattern recognition, 2022, pp. 11976–11986.
  17. “An image is worth 16x16 words: Transformers for image recognition at scale,” arXiv preprint arXiv:2010.11929, 2020.
  18. “Swin transformer: Hierarchical vision transformer using shifted windows,” in Proceedings of the IEEE/CVF international conference on computer vision, 2021, pp. 10012–10022.
  19. “Emerging properties in self-supervised vision transformers,” in Proceedings of the IEEE/CVF international conference on computer vision, 2021, pp. 9650–9660.
  20. “Adaptive federated optimization,” arXiv preprint arXiv:2003.00295, 2020.
  21. “Medmnist v2-a large-scale lightweight benchmark for 2d and 3d biomedical image classification,” Scientific Data, vol. 10, no. 1, pp. 41, 2023.
  22. “Flamby: Datasets and benchmarks for cross-silo federated learning in realistic healthcare settings,” Advances in Neural Information Processing Systems, vol. 35, pp. 5315–5334, 2022.
  23. “Deeporgan: Multi-level deep convolutional networks for automated pancreas segmentation,” in Medical Image Computing and Computer-Assisted Intervention–MICCAI 2015: 18th International Conference, Munich, Germany, October 5-9, 2015, Proceedings, Part I 18. Springer, 2015, pp. 556–564.
  24. “The kits21 challenge: Automatic segmentation of kidneys, renal tumors, and renal cysts in corticomedullary-phase ct,” arXiv preprint arXiv:2307.01984, 2023.
  25. “A large annotated medical image dataset for the development and evaluation of segmentation algorithms,” arXiv preprint arXiv:1902.09063, 2019.
  26. “Seven-point checklist and skin lesion classification using multitask multimodal neural nets,” IEEE journal of biomedical and health informatics, vol. 23, no. 2, pp. 538–546, 2018.
  27. “Pad-ufes-20: A skin lesion dataset composed of patient data and clinical images collected from smartphones,” Data in brief, vol. 32, pp. 106221, 2020.
  28. “A patient-centric dataset of images and metadata for identifying melanomas using clinical context,” Scientific data, vol. 8, no. 1, pp. 34, 2021.
  29. “Nvidia flare: Federated learning from simulation to real-world,” arXiv preprint arXiv:2210.13291, 2022.
Citations (1)

Summary

No one has generated a summary of this paper yet.

Paper to Video (Beta)

No one has generated a video about this paper yet.

Whiteboard

No one has generated a whiteboard explanation for this paper yet.

Open Problems

We haven't generated a list of open problems mentioned in this paper yet.

Continue Learning

We haven't generated follow-up questions for this paper yet.

Collections

Sign up for free to add this paper to one or more collections.