Apertus LLM Family Expansion via Distillation and Quantization
Expanding the Apertus model family through distillation and quantization to support a broader range of model sizes and hardware budgets.
A selection of my work on efficient large language models. For the complete list and citation information, see Google Scholar.