REVIEW 1 cited by
Depth-Width Tradeoffs in Approximating Natural Functions with Neural Networks
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
abstract
We provide several new depth-based separation results for feed-forward neural networks, proving that various types of simple and natural functions can be better approximated using deeper networks than shallower ones, even if the shallower networks are much larger. This includes indicators of balls and ellipses; non-linear functions which are radial with respect to the $L_1$ norm; and smooth non-linear functions. We also show that these gaps can be observed experimentally: Increasing the depth indeed allows better learning than increasing width, when training neural networks to learn an indicator of a unit ball.
Forward citations
Cited by 1 Pith paper
-
Theoretical Issues in Deep Networks: Approximation, Optimization and Generalization
A synthesis of approximation, optimization, and generalization theory arguing that gradient descent's implicit norm control on weight directions explains why overparameterized deep networks generalize.
Discussion (0). Continue with ORCID to comment.