Pith. sign in

REVIEW 1 cited by

Learning to Train a Binary Neural Network

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1809.10463 v1 pith:KIYULAKS submitted 2018-09-27 cs.LG cs.CVstat.ML

classification cs.LGcs.CVstat.ML
keywords neuralbinarynetworkmodelsnetworksdevicesdifferenteveryone
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Convolutional neural networks have achieved astonishing results in different application areas. Various methods which allow us to use these models on mobile and embedded devices have been proposed. Especially binary neural networks seem to be a promising approach for these devices with low computational power. However, understanding binary neural networks and training accurate models for practical applications remains a challenge. In our work, we focus on increasing our understanding of the training process and making it accessible to everyone. We publish our code and models based on BMXNet for everyone to use. Within this framework, we systematically evaluated different network architectures and hyperparameters to provide useful insights on how to train a binary neural network. Further, we present how we improved accuracy by increasing the number of connections in the network.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Effective Training of Convolutional Neural Networks with Low-bitwidth Weights and Activations

    cs.CV 2019-08 conditional novelty 6.0 of 10

    Two-stage and gradually decreasing quantization, stochastic precision sampling, and joint teacher-student distillation each improve low-bit CNN accuracy on ImageNet and CIFAR-100, with the largest gains when combined.

Pith tools