Quartet: Native FP4 Training Can Be Optimal for Large Language Models

Abstract

End-to-end FP4 training of language models, combining low-precision scaling laws with CUDA kernels for Blackwell GPUs.

Publication
NeurIPS 2025
Andrei Panferov
Andrei Panferov
PhD Candidate in Computer Science

I work on efficient large language models, with a focus on low-precision training, quantization, and scaling laws.