Llama Language Models

30 repos

Large language models from Meta's Llama family, including various quantized and optimized versions across different model sizes (1B through 70B parameters). This cluster contains model checkpoints and implementations for running Llama models efficiently, spanning formats like GGML and GPTQ quantization, enabling deployment on resource-constrained hardware. Developers exploring this cluster will find pre-trained model weights, inference optimization techniques, and tooling for serving and fine-tuning open-source large language models.

llama ·3,993
facebook ·3,993
meta ·3,993
transformers ·3,973
text-generation ·3,944
pytorch ·3,483
llama-2 ·3,441
en ·3,186
text-generation-inference ·1,630
safetensors ·1,530