Pith. sign in

Title resolution pending

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

QuEST: Stable Training of LLMs with 1-Bit Weights and Activations

cs.LG · 2025-02-07 · conditional · novelty 7.0

A quantization-aware training method with Hadamard normalization and a trust gradient mask trains Llama models stably down to 1-bit weights and activations and makes 4-bit precision Pareto-optimal in accuracy per memory.

citing papers explorer

Showing 1 of 1 citing paper.

  • QuEST: Stable Training of LLMs with 1-Bit Weights and Activations cs.LG · 2025-02-07 · conditional · none · ref 8

    A quantization-aware training method with Hadamard normalization and a trust gradient mask trains Llama models stably down to 1-bit weights and activations and makes 4-bit precision Pareto-optimal in accuracy per memory.