Large Language Model Fine-tuning and Inference

16 repos

Text generation models and frameworks for deploying, fine-tuning, and optimizing large language models (LLMs), particularly Mistral-based variants. The cluster centers on transformer-based models packaged with safetensors, model indices, and text-generation-inference infrastructure. Repositories here span model checkpoints, GGUF quantizations, and tooling for efficient inference and adaptation of open-source LLMs.

text-generation ·254
transformers ·253
code ·253
llama-2 ·253
llama ·207
gguf ·169
safetensors ·85
text-generation-inference ·84
model-index ·73
en ·72