Pith. sign in

REVIEW

PCNN: Pattern-based Fine-Grained Regular Pruning towards Optimizing CNN Accelerators

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2002.04997 v2 pith:TSPNBZTZ submitted 2020-02-11 cs.LG stat.ML

classification cs.LGstat.ML
keywords pcnnpruningcompressionfine-grainedonlyregularsparsityaccelerators
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Weight pruning is a powerful technique to realize model compression. We propose PCNN, a fine-grained regular 1D pruning method. A novel index format called Sparsity Pattern Mask (SPM) is presented to encode the sparsity in PCNN. Leveraging SPM with limited pruning patterns and non-zero sequences with equal length, PCNN can be efficiently employed in hardware. Evaluated on VGG-16 and ResNet-18, our PCNN achieves the compression rate up to 8.4X with only 0.2% accuracy loss. We also implement a pattern-aware architecture in 55nm process, achieving up to 9.0X speedup and 28.39 TOPS/W efficiency with only 3.1% on-chip memory overhead of indices.

Discussion (0). Sign in to comment.

Pith tools