Neural Conditional Gradients

Christian Bauckhage; Kristian Kersting; Patrick Schramowski

arxiv: 1803.04300 · v2 · pith:2SVBSXY3new · submitted 2018-03-12 · 💻 cs.LG · stat.ML

Neural Conditional Gradients

Patrick Schramowski , Christian Bauckhage , Kristian Kersting This is my paper

classification 💻 cs.LG stat.ML

keywords optimizersconditionalfrank-wolfegradientshand-designedlearnedlearningnetworks

0 comments

read the original abstract

The move from hand-designed to learned optimizers in machine learning has been quite successful for gradient-based and -free optimizers. When facing a constrained problem, however, maintaining feasibility typically requires a projection step, which might be computationally expensive and not differentiable. We show how the design of projection-free convex optimization algorithms can be cast as a learning problem based on Frank-Wolfe Networks: recurrent networks implementing the Frank-Wolfe algorithm aka. conditional gradients. This allows them to learn to exploit structure when, e.g., optimizing over rank-1 matrices. Our LSTM-learned optimizers outperform hand-designed as well learned but unconstrained ones. We demonstrate this for training support vector machines and softmax classifiers.

This paper has not been read by Pith yet.

Neural Conditional Gradients

discussion (0)