Pith. sign in

REVIEW 2 cited by

Distributed Robust Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1409.5937 v2 pith:LHXJZVRY submitted 2014-09-21 stat.ML cs.LG

classification stat.MLcs.LG
keywords robustdistributedlearningcomputingframeworknodespointrobustness
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

We propose a framework for distributed robust statistical learning on {\em big contaminated data}. The Distributed Robust Learning (DRL) framework can reduce the computational time of traditional robust learning methods by several orders of magnitude. We analyze the robustness property of DRL, showing that DRL not only preserves the robustness of the base robust learning method, but also tolerates contaminations on a constant fraction of results from computing nodes (node failures). More precisely, even in presence of the most adversarial outlier distribution over computing nodes, DRL still achieves a breakdown point of at least $ \lambda^*/2 $, where $ \lambda^* $ is the break down point of corresponding centralized algorithm. This is in stark contrast with naive division-and-averaging implementation, which may reduce the breakdown point by a factor of $ k $ when $ k $ computing nodes are used. We then specialize the DRL framework for two concrete cases: distributed robust principal component analysis and distributed robust regression. We demonstrate the efficiency and the robustness advantages of DRL through comprehensive simulations and predicting image tags on a large-scale image set.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. One-shot Robust Federated Learning of Independent Component Analysis

    cs.LG 2025-05 conditional novelty 6.0 of 10

    A one-shot federated ICA method that uses k-means to resolve permutation ambiguity and geometric median aggregation to stay accurate when a fraction of clients have very small sample sizes.

  2. Generalized Rank Regression

    stat.ME 2026-05 unverdicted novelty 5.0 of 10

    Generalized Rank Regression extends rank methods to non-monotonic scores, derives Bahadur representation and asymptotic normality, proposes a two-stage sub-gradient algorithm, and shows variance equivalence to composi...

Pith tools