QuEST: Stable Training of LLMs with 1-Bit Weights and Activations
Andrei Panferov, Jiale Chen, Soroush Tabesh, Mahdi Nikdan, Dan Alistarh
February, 2025
Abstract
Quantization-aware training with Hadamard normalization and a trust gradient estimator, enabling stable training down to 1-bit weights and activations.

PhD Candidate in Computer Science
I work on efficient large language models, with a focus on low-precision training, quantization, and scaling laws.