REVIEW 4 cited by
Non-IID data in Federated Learning: A Survey with Taxonomy, Metrics, Methods, Frameworks and Future Directions
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Recent advances in machine learning have highlighted Federated Learning (FL) as a promising approach that enables multiple distributed users (so-called clients) to collectively train ML models without sharing their private data. While this privacy-preserving method shows potential, it struggles when data across clients is not independent and identically distributed (non-IID) data. The latter remains an unsolved challenge that can result in poorer model performance and slower training times. Despite the significance of non-IID data in FL, there is a lack of consensus among researchers about its classification and quantification. This technical survey aims to fill that gap by providing a detailed taxonomy for non-IID data, partition protocols, and metrics to quantify data heterogeneity. Additionally, we describe popular solutions to address non-IID data and standardized frameworks employed in FL with heterogeneous data. Based on our state-of-the-art survey, we present key lessons learned and suggest promising future research directions.
Forward citations
Cited by 4 Pith papers
-
Privacy-Preserving Federated Averaging with Byzantine Aggregators in Asynchronous Networks
A new protocol enables differentially private federated averaging in asynchronous networks with fully Byzantine aggregators, using replicated servers, LWE masking, and verifiable cluster shuffling.
-
Model Fusion via Retrofitting
A neuron-centric fusion method that clusters intermediate activations of independently trained models into importance-weighted centroids and fits the fused network to them, outperforming baselines in zero-shot non-IID...
-
On the Effectiveness of Adaptation Strategies for VLM-Based Federated Learning in Remote Sensing
In federated remote sensing, LoRA tuning of a frozen CLIP model achieves the best accuracy-to-communication trade-off, while full fine-tuning causes severe catastrophic forgetting of pretrained knowledge.
-
PIcsC: Partitioning-Induced Covariate Shift Correction
A Fisher-information regularizer is proposed to correct partition-induced covariate shift in cross-validation and federated learning, with reported gains of 3-5 points over FedAvg-class baselines.
Discussion (0). Continue with ORCID to comment.