pith. sign in

arxiv: 1810.04146 · v1 · pith:7UOPCXEBnew · submitted 2018-10-09 · 💻 cs.DC

Decoupled Strategy for Imbalanced Workloads in MapReduce Frameworks

classification 💻 cs.DC
keywords mapreducecommunicationdecoupledframeworksimplementationone-sidedperformancestrategy
0
0 comments X
read the original abstract

In this work, we consider the integration of MPI one-sided communication and non-blocking I/O in HPC-centric MapReduce frameworks. Using a decoupled strategy, we aim to overlap the Map and Reduce phases of the algorithm by allowing processes to communicate and synchronize using solely one-sided operations. Hence, we effectively increase the performance in situations where the workload per process is unexpectedly unbalanced. Using a Word-Count implementation and a large dataset from the Purdue MapReduce Benchmarks Suite (PUMA), we demonstrate that our approach can provide up to 23% performance improvement on average compared to a reference MapReduce implementation that uses state-of-the-art MPI collective communication and I/O.

This paper has not been read by Pith yet.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.