Skip to main content
DOCUMENTATION • SPECIFICATIONS • TOOLING

Documentation Hub

Complete guides for inspecting Solace distillation datasets, deploying sub-8B quantized weights locally, and running fine-tuning harnesses.

$anvil run huihui-qwen --ctx 262144 --type-k turbo4 --type-v turbo3
STUDENT WEIGHTS1 guide

Sub-8B Quantized Checkpoints

vLLM, llama.cpp, Ollama, and Apple MLX deployment guides for FP8, INT4 AWQ, and GGUF.

HOW-TO RECIPES1 guide

Fine-Tuning & Training Recipes

End-to-end recipes for fine-tuning Llama 3.1 and Qwen 2.5 on consumer and cloud GPUs.