LLM Fine-tuning and Instruction Alignment

11 repos

Fine-tuned and instruction-aligned large language models, primarily built on open-source base models like Llama 3 and Gemma. This cluster focuses on supervised fine-tuning (SFT) approaches for improving model behavior, reasoning, and alignment with human instructions, often using frameworks like Hugging Face Transformers and text generation inference. The repositories represent both the training methodologies and the resulting model checkpoints designed for improved instruction-following and specialized tasks.

generated_from_trainer ·561
tensorboard ·561
text-generation-inference ·561
transformers ·561
endpoints_compatible ·561
text-generation ·561
safetensors ·532
conversational ·298
pytorch ·292
alignment-handbook ·267