Pith. sign in

REVIEW 1 cited by

Facilitating Bayesian Continual Learning by Natural Gradients and Stein Gradients

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1904.10644 v1 pith:MPT5TJBL submitted 2019-04-24 cs.LG cs.AIstat.ML

classification cs.LGcs.AIstat.ML
keywords learningtaskscontinualgradientsknowledgepreviousbayesianmodels
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Continual learning aims to enable machine learning models to learn a general solution space for past and future tasks in a sequential manner. Conventional models tend to forget the knowledge of previous tasks while learning a new task, a phenomenon known as catastrophic forgetting. When using Bayesian models in continual learning, knowledge from previous tasks can be retained in two ways: 1). posterior distributions over the parameters, containing the knowledge gained from inference in previous tasks, which then serve as the priors for the following task; 2). coresets, containing knowledge of data distributions of previous tasks. Here, we show that Bayesian continual learning can be facilitated in terms of these two means through the use of natural gradients and Stein gradients respectively.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Revised Regularization for Efficient Continual Learning through Correlation-Based Parameter Update in Bayesian Neural Networks

    cs.LG 2024-11 reject novelty 4.0 of 10

    The ECL-RR method combines a revised mean/variance regularizer, a parameter-learning-network compression, and a common/distinctive subspace decomposition, and reports top accuracy on four continual learning benchmarks.

Pith tools