Pith. sign in

REVIEW

Wasserstein Training of Boltzmann Machines

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1507.01972 v1 pith:TJEGAV5B submitted 2015-07-07 stat.ML cs.LG

classification stat.MLcs.LG
keywords boltzmannmetrictrainingwassersteindatadistributionslearnmodel
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The Boltzmann machine provides a useful framework to learn highly complex, multimodal and multiscale data distributions that occur in the real world. The default method to learn its parameters consists of minimizing the Kullback-Leibler (KL) divergence from training samples to the Boltzmann model. We propose in this work a novel approach for Boltzmann training which assumes that a meaningful metric between observations is given. This metric can be represented by the Wasserstein distance between distributions, for which we derive a gradient with respect to the model parameters. Minimization of this new Wasserstein objective leads to generative models that are better when considering the metric and that have a cluster-like structure. We demonstrate the practical potential of these models for data completion and denoising, for which the metric between observations plays a crucial role.

Discussion (0). Continue with ORCID to comment.

Pith tools