首页 > AI前沿 > Client and Training Data Selection for Computationally Efficient Synchronized Federated Learning

Client and Training Data Selection for Computationally Efficient Synchronized Federated Learning

arXiv机器学习 2026-09-30 16:14 6 阅读 查看原文

Federated learning (FL) is a promising paradigm of machine learning, which preserves user privacy by enabling learning without sharing raw data with a cloud server.

Straggling clients have been a problem for FL as they introduce delays in aggregating the local models and hence, the convergence of the global model.

Therefore, it is important to have a mechanism that ensures fast convergence of the global model as well as good FL participation rate.

Another issue for the convergence of a model in FL is the non-independent and identically distributed (non-iid) data across the clients.

Prior approaches based on probabilistic client selection do not work well under non-iid data especially when the number of clients is small.

We show scenarios where such approaches fail and propose a joint client-training data selection algorithm for fast convergence of FL models.

Our experiments on CIFAR-100 dataset show that convergence of the FL model can be significantly improved over prior works that can consider non-iid data and heterogeneous computation and higher model accuracy.