arxiv: 0906.4779 · v4 · pith:6F44CYDGnew · submitted 2009-06-25 · 💻 cs.LG · physics.data-an· stat.ML

Minimum Probability Flow Learning

Jascha Sohl-Dickstein , Peter Battaglino , Michael R. DeWeese This is my paper

classification 💻 cs.LG physics.data-anstat.ML

keywords distributionlearningmodeldatadivergencedynamicsestimationising

0 comments

read the original abstract

Fitting probabilistic models to data is often difficult, due to the general intractability of the partition function and its derivatives. Here we propose a new parameter estimation technique that does not require computing an intractable normalization factor or sampling from the equilibrium distribution of the model. This is achieved by establishing dynamics that would transform the observed data distribution into the model distribution, and then setting as the objective the minimization of the KL divergence between the data distribution and the distribution produced by running the dynamics for an infinitesimal time. Score matching, minimum velocity learning, and certain forms of contrastive divergence are shown to be special cases of this learning technique. We demonstrate parameter estimation in Ising models, deep belief networks and an independent component analysis model of natural scenes. In the Ising model case, current state of the art techniques are outperformed by at least an order of magnitude in learning time, with lower error in recovered coupling parameters.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

Generative Modeling by Estimating Gradients of the Data Distribution
cs.LG 2019-07 unverdicted novelty 6.0

Score-based generative modeling via multi-noise-level score matching and annealed Langevin dynamics produces samples on par with GANs and sets a new inception score record on CIFAR-10.