Pith. sign in

One-Shot Generalization in Deep Generative Models

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Humans have an impressive ability to reason about new concepts and experiences from just a single example. In particular, humans have an ability for one-shot generalization: an ability to encounter a new concept, understand its structure, and then be able to generate compelling alternative variations of the concept. We develop machine learning systems with this important capacity by developing new deep generative models, models that combine the representational power of deep learning with the inferential power of Bayesian reasoning. We develop a class of sequential generative models that are built on the principles of feedback and attention. These two characteristics lead to generative models that are among the state-of-the art in density estimation and image generation. We demonstrate the one-shot generalization ability of our models using three tasks: unconditional sampling, generating new exemplars of a given concept, and generating new exemplars of a family of concepts. In all cases our models are able to generate compelling and diverse samples---having seen new examples just once---providing an important class of general-purpose models for one-shot machine learning.

fields

cs.LG 1

years

2019 1

verdicts

CONDITIONAL 1

representative citing papers

Learning to Generalize to Unseen Tasks with Bilevel Optimization

cs.LG · 2019-08-05 · conditional · novelty 4.0

L2G, a bilevel training objective that evaluates an inner-loop gradient update on a second disjoint task, improves Prototypical and Relation Networks by one to five accuracy points on mini-ImageNet and tiered-ImageNet.

citing papers explorer

Showing 1 of 1 citing paper.

  • Learning to Generalize to Unseen Tasks with Bilevel Optimization cs.LG · 2019-08-05 · conditional · none · ref 7 · internal anchor

    L2G, a bilevel training objective that evaluates an inner-loop gradient update on a second disjoint task, improves Prototypical and Relation Networks by one to five accuracy points on mini-ImageNet and tiered-ImageNet.