Model hub — Distillio

← Home

Hand-picked lab flagship models for fine-tuning (safetensors), and ready-to-run GGUF quantizations for local inference (Ollama, LM Studio, llama.cpp). Download links go to the original Hugging Face repos.

Fine-tune models (safetensors)

Local GGUF models

Quantized for Ollama / LM Studio / llama.cpp, ordered by download popularity. Click a repo link to see its exact quant files on Hugging Face.