pith. sign in

Quartet II: Accu- rate LLM pre-training in NVFP4 by improved unbiased gradient estimation

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.LG 5

years

2026 5

roles

background 1

polarities

background 1

representative citing papers

Normalized Architectures are Natively 4-Bit

cs.LG · 2026-05-07 · conditional · novelty 6.0

nGPT's hypersphere constraint makes dot-product signal accumulate constructively under 4-bit quantization while noise averages out, enabling native low-precision training.

citing papers explorer

Showing 5 of 5 citing papers.