REVIEW 2 cited by
Lifelong Generative Modeling
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Lifelong learning is the problem of learning multiple consecutive tasks in a sequential manner, where knowledge gained from previous tasks is retained and used to aid future learning over the lifetime of the learner. It is essential towards the development of intelligent machines that can adapt to their surroundings. In this work we focus on a lifelong learning approach to unsupervised generative modeling, where we continuously incorporate newly observed distributions into a learned model. We do so through a student-teacher Variational Autoencoder architecture which allows us to learn and preserve all the distributions seen so far, without the need to retain the past data nor the past models. Through the introduction of a novel cross-model regularizer, inspired by a Bayesian update rule, the student model leverages the information learned by the teacher, which acts as a probabilistic knowledge store. The regularizer reduces the effect of catastrophic interference that appears when we learn over sequences of distributions. We validate our model's performance on sequential variants of MNIST, FashionMNIST, PermutedMNIST, SVHN and Celeb-A and demonstrate that our model mitigates the effects of catastrophic interference faced by neural networks in sequential learning scenarios.
Forward citations
Cited by 2 Pith papers
-
Online Continual Learning with Maximally Interfered Retrieval
Selecting replay samples by estimated loss increase after a virtual update improves online continual learning performance over random replay.
-
Incrementally Learning Multiple Diverse Data Domains via Multi-Source Dynamic Expansion Model
MSDEM grows per-task experts on top of multiple frozen ViT backbones with an attention-based fusion and a Gumbel-Softmax graph router, and reports state-of-the-art accuracy on multi-domain continual image classification.
Discussion (0). Continue with ORCID to comment.