Near-optimal one- and two-round protocols for ℓp heavy hitters and Fp estimation in the coordinator and distributed tracking models, including the first near-optimal algorithms for tracking Fp.
When Distributed Computation is Communication Expensive
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
We consider a number of fundamental statistical and graph problems in the message-passing model, where we have $k$ machines (sites), each holding a piece of data, and the machines want to jointly solve a problem defined on the union of the $k$ data sets. The communication is point-to-point, and the goal is to minimize the total communication among the $k$ machines. This model captures all point-to-point distributed computational models with respect to minimizing communication costs. Our analysis shows that exact computation of many statistical and graph problems in this distributed setting requires a prohibitively large amount of communication, and often one cannot improve upon the communication of the simple protocol in which all machines send their data to a centralized server. Thus, in order to obtain protocols that are communication-efficient, one has to allow approximation, or investigate the distribution or layout of the data sets.
citation-role summary
citation-polarity summary
fields
cs.DS 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Simple and Optimal Algorithms for Heavy Hitters and Frequency Moments in Distributed Models
Near-optimal one- and two-round protocols for ℓp heavy hitters and Fp estimation in the coordinator and distributed tracking models, including the first near-optimal algorithms for tracking Fp.