A comprehensive experimental comparison between federated and centralized learning

被引:0
|
作者
Garst, Swier [1 ]
Dekker, Julian [1 ]
Reinders, Marcel [1 ]
机构
[1] Delft Univ Technol, Intelligent Syst, Mourik Broekmanweg 6, NL-2628 XE Delft, Zuid Holland, Netherlands
来源
DATABASE-THE JOURNAL OF BIOLOGICAL DATABASES AND CURATION | 2025年 / 2025卷
关键词
D O I
10.1093/database/baaf016
中图分类号
Q [生物科学];
学科分类号
07 ; 0710 ; 09 ;
摘要
Federated learning is an upcoming machine learning paradigm which allows data from multiple sources to be used for training of classifiers without the data leaving the source it originally resides. This can be highly valuable for use cases such as medical research, where gathering data at a central location can be quite complicated due to privacy and legal concerns of the data. In such cases, federated learning has the potential to vastly speed up the research cycle. Although federated and central learning have been compared from a theoretical perspective, an extensive experimental comparison of performances and learning behavior still lacks. We have performed a comprehensive experimental comparison between federated and centralized learning. We evaluated various classifiers on various datasets exploring influences of different sample distributions as well as different class distributions across the clients. The results show similar performances under a wide variety of settings between the federated and central learning strategies. Federated learning is able to deal with various imbalances in the data distributions. It is sensitive to batch effects between different datasets when they coincide with location, similar to central learning, but this setting might go unobserved more easily. Federated learning seems to be robust to various challenges such as skewed data distributions, high data dimensionality, multiclass problems, and complex models. Taken together, the insights from our comparison gives much promise for applying federated learning as an alternative to sharing data. Code for reproducing the results in this work can be found at: https://github.com/swiergarst/FLComparison
引用
收藏
页数:10
相关论文
共 50 条
  • [31] Federated Learning on Multimodal Data: A Comprehensive Survey
    Yi-Ming Lin
    Yuan Gao
    Mao-Guo Gong
    Si-Jia Zhang
    Yuan-Qiao Zhang
    Zhi-Yuan Li
    Machine Intelligence Research, 2023, 20 : 539 - 553
  • [32] A Survey on Federated Learning: The Journey From Centralized to Distributed On-Site Learning and Beyond
    AbdulRahman, Sawsan
    Tout, Hanine
    Ould-Slimane, Hakima
    Mourad, Azzam
    Talhi, Chamseddine
    Guizani, Mohsen
    IEEE INTERNET OF THINGS JOURNAL, 2021, 8 (07): : 5476 - 5497
  • [33] Hypernetwork-driven centralized contrastive learning for federated graph classification
    Zhu, Jianian
    Li, Yichen
    Wang, Haozhao
    Qi, Yining
    Li, Ruixuan
    WORLD WIDE WEB-INTERNET AND WEB INFORMATION SYSTEMS, 2024, 27 (05):
  • [34] A Comprehensive Empirical Study of Heterogeneity in Federated Learning
    Abdelmoniem, Ahmed M. M.
    Ho, Chen-Yu
    Papageorgiou, Pantelis
    Canini, Marco
    IEEE INTERNET OF THINGS JOURNAL, 2023, 10 (16) : 14071 - 14083
  • [35] Federated Learning for Internet of Things: A Comprehensive Survey
    Nguyen, Dinh C.
    Ding, Ming
    Pathirana, Pubudu N.
    Seneviratne, Aruna
    Li, Jun
    Poor, H. Vincent
    IEEE COMMUNICATIONS SURVEYS AND TUTORIALS, 2021, 23 (03): : 1622 - 1658
  • [36] Federated Learning on Multimodal Data: A Comprehensive Survey
    Lin, Yi-Ming
    Gao, Yuan
    Gong, Mao-Guo
    Zhang, Si-Jia
    Zhang, Yuan-Qiao
    Li, Zhi-Yuan
    MACHINE INTELLIGENCE RESEARCH, 2023, 20 (04) : 539 - 553
  • [37] Wireless Federated Learning With Hybrid Local and Centralized Training: A Latency Minimization Design
    Huang, Ning
    Dai, Minghui
    Wu, Yuan
    Quek, Tony Q. S.
    Shen, Xuemin
    IEEE JOURNAL OF SELECTED TOPICS IN SIGNAL PROCESSING, 2023, 17 (01) : 248 - 263
  • [38] Optimal Design of Hybrid Federated and Centralized Learning in the Mobile Edge Computing Systems
    Hong, Wei
    Luo, Xueting
    Zhao, Zhongyuan
    Peng, Mugen
    Quek, Tony Q. S.
    2021 IEEE INTERNATIONAL CONFERENCE ON COMMUNICATIONS WORKSHOPS (ICC WORKSHOPS), 2021,
  • [39] A Comparative Study of Federated Versus Centralized Learning for Time Series Prediction in IoT
    Furtado, Lia Sucupira
    da Costa, Leonardo Ferreira
    Goncalves Rocha, Paulo Henrique
    Leal Rego, Paulo Antonio
    INTELLIGENT SYSTEMS, PT II, 2022, 13654 : 297 - 311
  • [40] Model aggregation techniques in federated learning: A comprehensive survey
    Qi, Pian
    Chiaro, Diletta
    Guzzo, Antonella
    Ianni, Michele
    Fortino, Giancarlo
    Piccialli, Francesco
    FUTURE GENERATION COMPUTER SYSTEMS-THE INTERNATIONAL JOURNAL OF ESCIENCE, 2024, 150 : 272 - 293