Pith. sign in

hub

Scalable best-of-n selection for large language models via self-certainty.arXiv preprint arXiv:2502.18581

25 Pith papers cite this work. Polarity classification is still indexing.

25 Pith papers citing it

hub tools

citation-role summary

background 3 other 1

citation-polarity summary

years

2026 21 2025 4

polarities

background 3 unclear 1

representative citing papers

Tracing Uncertainty in Language Model "Reasoning"

cs.LG · 2026-05-08 · unverdicted · novelty 7.0

Uncertainty trace profiles from LM reasoning traces predict correct final answers with AUROC up to 0.807 and enable early error detection using only initial tokens.

Multi-Token Prediction via Self-Distillation

cs.CL · 2026-02-05 · unverdicted · novelty 6.0

Self-distillation turns pretrained autoregressive LMs into multi-token predictors that decode over 3x faster with under 5% accuracy drop on GSM8K.

citing papers explorer

Showing 25 of 25 citing papers.