REVIEW 1 cited by
Few-Round Distributed Principal Component Analysis: Closing the Statistical Efficiency Gap by Consensus
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Distributed algorithms and theories are called for in this era of big data. Under weaker local signal-to-noise ratios, we improve upon the celebrated one-round distributed principal component analysis (PCA) algorithm designed in the spirit of divide-and-conquer, by introducing a few additional communication rounds of consensus. The proposed shifted subspace iteration algorithm is able to close the local phase transition gap, reduce the asymptotic variance, and also alleviate the potential bias. Our estimation procedure is easy to implement and tuning-free. The resulting estimator is shown to be statistically efficient after an acceptable number of iterations. We also discuss extensions to distributed elliptical PCA for heavy-tailed data. Empirical experiments on synthetic and benchmark datasets demonstrate our method's statistical advantage over the divide-and-conquer approach.
Forward citations
Cited by 1 Pith paper
-
Debiased distributed PCA under high dimensional spiked model
A debiased, sparsity-adaptive distributed PCA algorithm achieves consistency under finite sixth moments and outperforms prior methods, especially with few machines.
Discussion (0). Continue with ORCID to comment.