LLaMA Model Fine-tuning & Inference

9 repos

This cluster focuses on fine-tuned variants of Meta's LLaMA language models, particularly adapted for Chinese language and specialized domains like coding and medical applications. Repositories here cover model weights/deltas, inference optimization, and text generation infrastructure using PyTorch and compatible endpoints. Someone exploring this area would find practical implementations of LLaMA model adaptation, deployment patterns, and domain-specific training approaches.

endpoints_compatible ·36
llama ·36
pytorch ·36
text-generation ·36
text-generation-inference ·36
transformers ·36