Pith. sign in

REVIEW 1 cited by

Distal Interference: Exploring the Limits of Model-Based Continual Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.08255 v1 pith:DSV7MLIF submitted 2024-02-13 cs.LG cs.AIcs.NE

classification cs.LGcs.AIcs.NE
keywords interferencelearningdistalcontinualabel-splinescatastrophictaskscomplexity
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Continual learning is the sequential learning of different tasks by a machine learning model. Continual learning is known to be hindered by catastrophic interference or forgetting, i.e. rapid unlearning of earlier learned tasks when new tasks are learned. Despite their practical success, artificial neural networks (ANNs) are prone to catastrophic interference. This study analyses how gradient descent and overlapping representations between distant input points lead to distal interference and catastrophic interference. Distal interference refers to the phenomenon where training a model on a subset of the domain leads to non-local changes on other subsets of the domain. This study shows that uniformly trainable models without distal interference must be exponentially large. A novel antisymmetric bounded exponential layer B-spline ANN architecture named ABEL-Spline is proposed that can approximate any continuous function, is uniformly trainable, has polynomial computational complexity, and provides some guarantees for distal interference. Experiments are presented to demonstrate the theoretical properties of ABEL-Splines. ABEL-Splines are also evaluated on benchmark regression problems. It is concluded that the weaker distal interference guarantees in ABEL-Splines are insufficient for model-only continual learning. It is conjectured that continual learning with polynomial complexity models requires augmentation of the training data or algorithm.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Multi-Exit Kolmogorov-Arnold Networks: enhancing accuracy and parsimony

    cs.LG 2025-06 conditional novelty 4.0 of 10

    Augmenting Kolmogorov-Arnold Networks with prediction exits at each layer improves accuracy and often yields more parsimonious models, and a differentiable learning-to-exit algorithm automates the choice of exit weights.

Pith tools