Effects of the optimisation of the margin distribution on generalisation in deep architectures

Brendan McCane; Lech Szymanski; Wei Gao; Zhi-Hua Zhou

arxiv: 1704.05646 · v1 · pith:CNRTV2XBnew · submitted 2017-04-19 · 💻 cs.LG

Effects of the optimisation of the margin distribution on generalisation in deep architectures

Lech Szymanski , Brendan McCane , Wei Gao , Zhi-Hua Zhou This is my paper

classification 💻 cs.LG

keywords margindeeparchitecturesgeneralisationlearninglossmaximisationvariance

0 comments

read the original abstract

Despite being so vital to success of Support Vector Machines, the principle of separating margin maximisation is not used in deep learning. We show that minimisation of margin variance and not maximisation of the margin is more suitable for improving generalisation in deep architectures. We propose the Halfway loss function that minimises the Normalised Margin Variance (NMV) at the output of a deep learning models and evaluate its performance against the Softmax Cross-Entropy loss on the MNIST, smallNORB and CIFAR-10 datasets.

This paper has not been read by Pith yet.

Effects of the optimisation of the margin distribution on generalisation in deep architectures

discussion (0)