23 repos
Fine-tuned and pretrained versions of the Qwen3 language models across multiple sizes (8B, 30B, 32B parameters), optimized for text generation and conversational tasks. The cluster contains model checkpoints and implementations compatible with Hugging Face Transformers, leveraging the safetensors format for efficient model distribution and loading. These represent different training approaches (SFT for supervised fine-tuning, CPT for continued pretraining) targeting various deployment endpoints and inference scenarios.