11 repos
Fine-tuned and instruction-aligned large language models, primarily built on open-source base models like Llama 3 and Gemma. This cluster focuses on supervised fine-tuning (SFT) approaches for improving model behavior, reasoning, and alignment with human instructions, often using frameworks like Hugging Face Transformers and text generation inference. The repositories represent both the training methodologies and the resulting model checkpoints designed for improved instruction-following and specialized tasks.