paper-conference

Bridging the Gap Between Promise and Performance for Microscaling FP4 Quantization
A study of MXFP4 and NVFP4 inference, introducing Micro-Rotated-GPTQ and GPU kernels tailored to the constraints of microscaling formats.
Apertus: Democratizing Open and Compliant LLMs for Global Language Environments
An open multilingual language model effort with transparent training recipes and a focus on data compliance and language coverage.