Pith. sign in

REVIEW

Understanding Deep Learning Generalization by Maximum Entropy

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1711.07758 v1 pith:R2P5ETL5 submitted 2017-11-21 cs.LG

classification cs.LG
keywords entropymaximumgeneralizationlearningdeepfeaturemodelunderstanding
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Deep learning achieves remarkable generalization capability with overwhelming number of model parameters. Theoretical understanding of deep learning generalization receives recent attention yet remains not fully explored. This paper attempts to provide an alternative understanding from the perspective of maximum entropy. We first derive two feature conditions that softmax regression strictly apply maximum entropy principle. DNN is then regarded as approximating the feature conditions with multilayer feature learning, and proved to be a recursive solution towards maximum entropy principle. The connection between DNN and maximum entropy well explains why typical designs such as shortcut and regularization improves model generalization, and provides instructions for future model development.

Discussion (0). Sign in to comment.

Pith tools