Pre-trained Language Models & Inference

8 repos

Implementations and deployments of large language models (LLMs) including OPT variants, GPT-2 adaptations, and CodeGPT models, with a focus on inference optimization and serving. The cluster centers on frameworks like Transformers and PyTorch for running and fine-tuning these models, plus specialized tools for efficient text generation inference. Repositories here span model checkpoints, language-specific variants (Chinese, Java), and the infrastructure for deploying these models at scale.

vllm ·4,527
mistral-common ·4,147
safetensors ·4,147
mistral ·3,865
pytorch ·671
text-generation ·671
text-generation-inference ·671
transformers ·671
mistral3 ·662
tf ·628