Pith. sign in

REVIEW

Uncovering Uncertainty in Transformer Inference

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2412.05768 v1 pith:5ETU5X4Z submitted 2024-12-08 cs.CL cs.AI

classification cs.CLcs.AI
keywords tokenuncertaintycorrectgenerationsincorrectinferenceresidualadditionally
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We explore the Iterative Inference Hypothesis (IIH) within the context of transformer-based language models, aiming to understand how a model's latent representations are progressively refined and whether observable differences are present between correct and incorrect generations. Our findings provide empirical support for the IIH, showing that the nth token embedding in the residual stream follows a trajectory of decreasing loss. Additionally, we observe that the rate at which residual embeddings converge to a stable output representation reflects uncertainty in the token generation process. Finally, we introduce a method utilizing cross-entropy to detect this uncertainty and demonstrate its potential to distinguish between correct and incorrect token generations on a dataset of idioms.

Discussion (0). Sign in to comment.

Pith tools