Pith. sign in

REVIEW 1 cited by

DCRA: A Distributed Chiplet-based Reconfigurable Architecture for Irregular Applications

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.15443 v2 pith:AHJ2J6FR submitted 2023-11-26 cs.AR cs.DC

classification cs.ARcs.DC
keywords dcraapplicationsarchitecturehelpirregularchiplet-basedconfigurationsdatasets
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In recent years, the growing demand to process large graphs and sparse datasets has led to increased research efforts to develop hardware- and software-based architectural solutions to accelerate them. While some of these approaches achieve scalable parallelization with up to thousands of cores, adaptation of these proposals by the industry remained slow. To help solve this dissonance, we identified a set of questions and considerations that current research has not considered deeply. Starting from a tile-based architecture, we put forward a Distributed Chiplet-based Reconfigurable Architecture (DCRA) for irregular applications that carefully consider fabrication constraints that made prior work either hard or costly to implement or too rigid to be applied. We identify and study pre-silicon, package-time and compile-time configurations that help optimize DCRA for different deployments and target metrics. To enable that, we propose a practical path for manufacturing chip packages by composing variable numbers of DCRA and memory dies, with a software-configurable Torus network to connect them. We evaluate six applications and four datasets, with several configurations and memory technologies, to provide a detailed analysis of the performance, power, and cost of DCRA as a compute node for scale-out sparse data processing. Finally, we present our findings and discuss how DCRA's framework for design exploration can help guide architects to build scalable and cost-efficient systems for irregular applications.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Topology-Aware Virtualization over Inter-Core Connected Neural Processing Units

    cs.AR 2025-06 conditional novelty 7.0 of 10

    vNPU virtualizes inter-core connected NPUs via core-ID redirection, range-based memory translation, and topology mapping, achieving up to 1.92x speedup over MIG.

Pith tools