REVIEW 1 cited by
Efficient Recurrent Neural Networks using Structured Matrices in FPGAs
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
abstract
Recurrent Neural Networks (RNNs) are becoming increasingly important for time series-related applications which require efficient and real-time implementations. The recent pruning based work ESE suffers from degradation of performance/energy efficiency due to the irregular network structure after pruning. We propose block-circulant matrices for weight matrix representation in RNNs, thereby achieving simultaneous model compression and acceleration. We aim to implement RNNs in FPGA with highest performance and energy efficiency, with certain accuracy requirement (negligible accuracy degradation). Experimental results on actual FPGA deployments shows that the proposed framework achieves a maximum energy efficiency improvement of 35.7$\times$ compared with ESE.
Forward citations
Cited by 1 Pith paper
-
Hardware-Efficient Photonic Tensor Core: Accelerating Deep Neural Networks with Structured Compression
A block-circulant photonic tensor core runs structure-compressed image classifiers within about 1.4 to 3.7 percentage points of full-precision digital models while reducing trainable parameters by up to 74.91%.
Discussion (0). Continue with ORCID to comment.