Pith. sign in

REVIEW 1 cited by

Pruning Algorithms to Accelerate Convolutional Neural Networks for Edge Applications: A Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2005.04275 v1 pith:B6XTQDSZ submitted 2020-05-08 cs.LG stat.ML

classification cs.LGstat.ML
keywords pruningmodelsurveycompressionconvolutionaledgemajorneural
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

With the general trend of increasing Convolutional Neural Network (CNN) model sizes, model compression and acceleration techniques have become critical for the deployment of these models on edge devices. In this paper, we provide a comprehensive survey on Pruning, a major compression strategy that removes non-critical or redundant neurons from a CNN model. The survey covers the overarching motivation for pruning, different strategies and criteria, their advantages and drawbacks, along with a compilation of major pruning techniques. We conclude the survey with a discussion on alternatives to pruning and current challenges for the model compression community.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Priority-Aware Model-Distributed Inference at Edge Networks

    cs.DC 2024-12 conditional novelty 4.0 of 10

    Adding priority weights to the multi-source model-distributed inference objective and scheduling each layer-group by a greedy delay-to-priority ratio shortens average inference time for high-priority sources on edge testbeds.

Pith tools