Pith. sign in

arXiv preprint arXiv:2109.00267 , year=

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.AI 1

years

2026 1

verdicts

UNVERDICTED 1

representative citing papers

Can Scale Save Us From Plasticity Loss in Large Language Models?

cs.AI · 2026-06-23 · unverdicted · novelty 6.0

Plasticity loss in GPT-style transformers on multilingual tasks persists from 5M to 314M parameters, follows a sublinear scaling law with model size, and occurs in both continual and stationary settings.

citing papers explorer

Showing 1 of 1 citing paper.

  • Can Scale Save Us From Plasticity Loss in Large Language Models? cs.AI · 2026-06-23 · unverdicted · none · ref 108

    Plasticity loss in GPT-style transformers on multilingual tasks persists from 5M to 314M parameters, follows a sublinear scaling law with model size, and occurs in both continual and stationary settings.