Nvidia researchers used model pruning and distillation to create a small language model (SLM) at a fraction of the base cost.
Cibo e viaggi / Food and travel notes by Livio Acerbo

Nvidia researchers used model pruning and distillation to create a small language model (SLM) at a fraction of the base cost.