22 repos
Fine-tuned and distilled variants of large language models optimized for instruction-following and text generation tasks. This cluster centers on creating smaller, more efficient models through distillation techniques applied to base architectures like Mistral and Yi, with a focus on maintaining instruction-following capability across various model sizes (7B to 34B parameters). Repositories here contain model weights, training approaches, and implementations compatible with standard text generation endpoints.