Pith. sign in

REVIEW 1 cited by

Fast Adjustable Threshold For Uniform Neural Network Quantization (Winning solution of LPIRC-II)

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1812.07872 v3 pith:2MCJRLH2 submitted 2018-12-19 cs.LG stat.ML

classification cs.LGstat.ML
keywords quantizationnetworkprocedurefine-tuningneuralwithoutaccuracyadjustable
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Neural network quantization procedure is the necessary step for porting of neural networks to mobile devices. Quantization allows accelerating the inference, reducing memory consumption and model size. It can be performed without fine-tuning using calibration procedure (calculation of parameters necessary for quantization), or it is possible to train the network with quantization from scratch. Training with quantization from scratch on the labeled data is rather long and resource-consuming procedure. Quantization of network without fine-tuning leads to accuracy drop because of outliers which appear during the calibration. In this article we suggest to simplify the quantization procedure significantly by introducing the trained scale factors for quantization thresholds. It allows speeding up the process of quantization with fine-tuning up to 8 epochs as well as reducing the requirements to the set of train images. By our knowledge, the proposed method allowed us to get the first public available quantized version of MNAS without significant accuracy reduction - 74.8% vs 75.3% for original full-precision network. Model and code are ready for use and available at: https://github.com/agoncharenko1992/FAT-fast_adjustable_threshold.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Towards Applying Deep Learning to The Internet of Things: A Model and A Framework

    cs.NI 2024-12 conditional novelty 5.0 of 10

    A conceptual schema and management framework called DLOM2 are proposed to select or create optimized deep learning models for IoT devices, but the design remains unvalidated.

Pith tools