Pith. sign in

HERO: Heterogeneous Embedded Research Platform for Exploring RISC-V Manycore Accelerators on FPGA

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Heterogeneous embedded systems on chip (HESoCs) co-integrate a standard host processor with programmable manycore accelerators (PMCAs) to combine general-purpose computing with domain-specific, efficient processing capabilities. While leading companies successfully advance their HESoC products, research lags behind due to the challenges of building a prototyping platform that unites an industry-standard host processor with an open research PMCA architecture. In this work we introduce HERO, an FPGA-based research platform that combines a PMCA composed of clusters of RISC-V cores, implemented as soft cores on an FPGA fabric, with a hard ARM Cortex-A multicore host processor. The PMCA architecture mapped on the FPGA is silicon-proven, scalable, configurable, and fully modifiable. HERO includes a complete software stack that consists of a heterogeneous cross-compilation toolchain with support for OpenMP accelerator programming, a Linux driver, and runtime libraries for both host and PMCA. HERO is designed to facilitate rapid exploration on all software and hardware layers: run-time behavior can be accurately analyzed by tracing events, and modifications can be validated through fully automated hard ware and software builds and executed tests. We demonstrate the usefulness of HERO by means of case studies from our research.

fields

cs.NI 1

years

2019 1

verdicts

CONDITIONAL 1

representative citing papers

Network-Accelerated Non-Contiguous Memory Transfers

cs.NI · 2019-08-22 · conditional · novelty 7.0

Offloading MPI datatype unpacking to programmable NIC packet handlers can reach up to 12x faster message processing than host-based unpacking in simulation.

citing papers explorer

Showing 1 of 1 citing paper.

  • Network-Accelerated Non-Contiguous Memory Transfers cs.NI · 2019-08-22 · conditional · none · ref 27 · internal anchor

    Offloading MPI datatype unpacking to programmable NIC packet handlers can reach up to 12x faster message processing than host-based unpacking in simulation.