Pith. sign in

REVIEW 1 cited by

Understanding Token Probability Encoding in Output Embeddings

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.01468 v2 pith:QZBU5RTR submitted 2024-06-03 cs.CL cs.AIcs.LG

Understanding Token Probability Encoding in Output Embeddings

classification cs.CL cs.AIcs.LG
keywords outputembeddingencodingfindprobabilitytokendimensionslanguage
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

In this paper, we investigate the output token probability information in the output embedding of language models. We find an approximate common log-linear encoding of output token probabilities within the output embedding vectors and empirically demonstrate that it is accurate and sparse. As a causality examination, we steer the encoding in output embedding to modify the output probability distribution accurately. Moreover, the sparsity we find in output probability encoding suggests that a large number of dimensions in the output embedding do not contribute to causal language modeling. Therefore, we attempt to delete the output-unrelated dimensions and find more than 30% of the dimensions can be deleted without significant movement in output distribution and sequence generation. Additionally, in the pre-training dynamics of language models, we find that the output embeddings capture the corpus token frequency information in early steps, even before an obvious convergence of parameters starts.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Reading Without a Reader: Large Language Models Collapse Reading and Writing into a Single Entangled Code

    q-bio.NC 2026-07 conditional novelty 5.0

    In decoder-only LLMs, input and output token codes are coupled but sub-ceiling (E≈0.23–0.35 against floor and ceiling anchors), and no output-side score pair can validly dissociate the model's 'reading' from its 'writing'.