Pith. sign in

REVIEW 1 cited by

Deep ReLU Networks Have Surprisingly Simple Polytopes

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.09145 v2 pith:GZTMFI6V submitted 2023-05-16 cs.LG cs.AIcs.CVcs.MM

classification cs.LGcs.AIcs.CVcs.MM
keywords polytopesnetworkfacesnumberrelunetworkssimpleanalyzing
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

A ReLU network is a piecewise linear function over polytopes. Figuring out the properties of such polytopes is of fundamental importance for the research and development of neural networks. So far, either theoretical or empirical studies on polytopes only stay at the level of counting their number, which is far from a complete characterization. Here, we propose to study the shapes of polytopes via the number of faces of the polytope. Then, by computing and analyzing the histogram of faces across polytopes, we find that a ReLU network has relatively simple polytopes under both initialization and gradient descent, although these polytopes can be rather diverse and complicated by a specific design. This finding can be appreciated as a kind of generalized implicit bias, subjected to the intrinsic geometric constraint in space partition of a ReLU network. Next, we perform a combinatorial analysis to explain why adding depth does not generate a more complicated polytope by bounding the average number of faces of polytopes with the dimensionality. Our results concretely reveal what kind of simple functions a network learns and what will happen when a network goes deep. Also, by characterizing the shape of polytopes, the number of faces can be a novel leverage for other problems, \textit{e.g.}, serving as a generic tool to explain the power of popular shortcut networks such as ResNet and analyzing the impact of different regularization strategies on a network's space partition.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. A Quotient Homology Theory of Representation in Neural Networks

    cs.LG 2025-02 conditional novelty 7.0 of 10

    For ReLU networks, the homology of the output representation is isomorphic to the homology of the input manifold quotiented by the network's overlap decomposition, when polyhedron-manifold intersections are convex.

Pith tools