Pith. sign in

REVIEW 3 cited by

Robust Federated Learning in a Heterogeneous Environment

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1906.06629 v2 pith:DIF4O6JC submitted 2019-06-16 cs.LG stat.ML

classification cs.LGstat.ML
keywords datalearningalgorithmfederatedmachinesstatisticaldifferentestimation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We study a recently proposed large-scale distributed learning paradigm, namely Federated Learning, where the worker machines are end users' own devices. Statistical and computational challenges arise in Federated Learning particularly in the presence of heterogeneous data distribution (i.e., data points on different devices belong to different distributions signifying different clusters) and Byzantine machines (i.e., machines that may behave abnormally, or even exhibit arbitrary and potentially adversarial behavior). To address the aforementioned challenges, first we propose a general statistical model for this problem which takes both the cluster structure of the users and the Byzantine machines into account. Then, leveraging the statistical model, we solve the robust heterogeneous Federated Learning problem \emph{optimally}; in particular our algorithm matches the lower bound on the estimation error in dimension and the number of data points. Furthermore, as a by-product, we prove statistical guarantees for an outlier-robust clustering algorithm, which can be considered as the Lloyd algorithm with robust estimation. Finally, we show via synthetic as well as real data experiments that the estimation error obtained by our proposed algorithm is significantly better than the non-Byzantine-robust algorithms; in particular, we gain at least by 53\% and 33\% for synthetic and real data experiments, respectively, in typical settings.

Discussion (0). Sign in to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Encryption-Compatible Clustered Federated Learning via Distributed Expectation-Maximization over Metadata

    cs.LG 2026-07 conditional novelty 6.0 of 10

    Distributed EM over random-network metadata makes clustered federated learning compatible with additive encryption without giving up efficiency.

  2. AFBS:Buffer Gradient Selection in Semi-asynchronous Federated Learning

    cs.LG 2025-06 conditional novelty 6.0 of 10

    AFBS scores buffered gradients by staleness and dataset size, discards low-value ones, and clusters clients through random-projection-encrypted label distributions before aggregation in semi-asynchronous federated learning.

  3. FedBiCross: Personalized One-Shot Federated Learning on Medical Images

    cs.LG 2026-01 unverdicted novelty 5.0 of 10

    FedBiCross clusters clients by model similarity, uses bi-level cross-cluster optimization for adaptive knowledge transfer, and applies personalized distillation to outperform baselines in non-IID data-free one-shot fe...

Pith tools