Pith. sign in

REVIEW

Principled Deep Neural Network Training through Linear Programming

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1810.03218 v3 pith:NHKEJDKM submitted 2018-10-07 cs.LG math.OCstat.ML

classification cs.LGmath.OCstat.ML
keywords problemstrainingnetworkbetterdeepneuralpolyhedralrepresentation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Deep learning has received much attention lately due to the impressive empirical performance achieved by training algorithms. Consequently, a need for a better theoretical understanding of these problems has become more evident in recent years. In this work, using a unified framework, we show that there exists a polyhedron which encodes simultaneously all possible deep neural network training problems that can arise from a given architecture, activation functions, loss function, and sample-size. Notably, the size of the polyhedral representation depends only linearly on the sample-size, and a better dependency on several other network parameters is unlikely (assuming $P\neq NP$). Additionally, we use our polyhedral representation to obtain new and better computational complexity results for training problems of well-known neural network architectures. Our results provide a new perspective on training problems through the lens of polyhedral theory and reveal a strong structure arising from these problems.

Discussion (0). Continue with ORCID to comment.

Pith tools