pith. machine review for the scientific record. sign in

Language models are unsupervised multitask learners

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CL 1

years

2024 1

verdicts

UNVERDICTED 1

representative citing papers

Massive Activations in Large Language Models

cs.CL · 2024-02-27 · unverdicted · novelty 7.0

Massive activations are constant large values in LLMs that function as indispensable bias terms and concentrate attention probabilities on specific tokens.

citing papers explorer

Showing 1 of 1 citing paper.

  • Massive Activations in Large Language Models cs.CL · 2024-02-27 · unverdicted · none · ref 74

    Massive activations are constant large values in LLMs that function as indispensable bias terms and concentrate attention probabilities on specific tokens.