28 repos
Specialized variants of popular language models (Llama, Zephyr, LLaVA) that have been instruction-tuned, aligned, or otherwise fine-tuned for specific tasks and behaviors. These models are distributed as Hugging Face-compatible checkpoints optimized for text generation inference, often produced through supervised fine-tuning (SFT) or alignment techniques. The cluster contains dozens of model variants with different training approaches and dataset combinations, useful for understanding model adaptation, instruction-following, and deployment of customized LLMs.