For every depth d below the exact depth of the n-simplex, every depth-d polytope misses the simplex by an empty-corner distance of exactly n+1-2^d.
Lower Bounds on the Depth of Integral ReLU Neural Networks via Lattice Polytopes
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
We prove that the set of functions representable by ReLU neural networks with integer weights strictly increases with the network depth while allowing arbitrary width. More precisely, we show that $\lceil\log_2(n)\rceil$ hidden layers are indeed necessary to compute the maximum of $n$ numbers, matching known upper bounds. Our results are based on the known duality between neural networks and Newton polytopes via tropical geometry. The integrality assumption implies that these Newton polytopes are lattice polytopes. Then, our depth lower bounds follow from a parity argument on the normalized volume of faces of such polytopes.
citation-role summary
citation-polarity summary
fields
math.MG 1years
2025 1verdicts
ACCEPT 1roles
background 1polarities
background 1representative citing papers
citing papers explorer
-
Approximation Depth of Convex Polytopes
For every depth d below the exact depth of the n-simplex, every depth-d polytope misses the simplex by an empty-corner distance of exactly n+1-2^d.