Federated Learning (FL) represents a distributed, privacy-preserving machine learning (ML) paradigm that enables decentralized model training across multiple clients. While traditional aggregation techniques, such as Federated Averaging (FedAVG), have demonstrated effectiveness, they often struggle in Not Independent and Identically Distributed (non-IID) scenarios, where data distributions vary significantly among clients. To address these limitations, this study introduces FedGP, a novel aggregation strategy based on Genetic Programming (GP). FedGP dynamically evolves aggregation functions, enabling adaptive and personalized model updates that better capture the heterogeneity inherent in distributed data. The proposed method is evaluated on the PathMNIST dataset, employing a comprehensive experimental design comprising 24 configurations, including 8 setups with FedAVG and 16 with FedGP. The comparative analysis highlights FedGP’s superior generalization capabilities and reduced biases, outperforming FedAVG in terms of accuracy. These results position FedGP as a robust and scalable solution for real-world FL applications, particularly in environments characterized by data heterogeneity.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

FedGP: Genetic Programming for Evolutionary Aggregation in Federated Learning with Non-IID Data

  • Elia Pacioni,
  • Francisco Fernández De Vega,
  • Davide Calvaresi

摘要

Federated Learning (FL) represents a distributed, privacy-preserving machine learning (ML) paradigm that enables decentralized model training across multiple clients. While traditional aggregation techniques, such as Federated Averaging (FedAVG), have demonstrated effectiveness, they often struggle in Not Independent and Identically Distributed (non-IID) scenarios, where data distributions vary significantly among clients. To address these limitations, this study introduces FedGP, a novel aggregation strategy based on Genetic Programming (GP). FedGP dynamically evolves aggregation functions, enabling adaptive and personalized model updates that better capture the heterogeneity inherent in distributed data. The proposed method is evaluated on the PathMNIST dataset, employing a comprehensive experimental design comprising 24 configurations, including 8 setups with FedAVG and 16 with FedGP. The comparative analysis highlights FedGP’s superior generalization capabilities and reduced biases, outperforming FedAVG in terms of accuracy. These results position FedGP as a robust and scalable solution for real-world FL applications, particularly in environments characterized by data heterogeneity.