15 repos
Text generation and deployment infrastructure built around PyTorch, with particular focus on adapter-based fine-tuning and inference optimization for models in the Llama family. The cluster centers on practical tools for adapting and serving large language models at various scales, emphasizing compatibility with standard text-generation endpoints. While most repos appear to be model weights or configurations rather than foundational libraries, the shared technical stack and deployment patterns reflect a coherent engineering practice around efficient LLM adaptation and inference.