7 repos
Modified versions of large language models (Qwen and Gemma families) optimized for GGUF format compatibility and conversational endpoints. These repositories represent fine-tuned or quantized model variants designed for deployment flexibility, with emphasis on removing content restrictions and providing balanced inference characteristics. The cluster centers on production-ready model distributions that prioritize format interoperability and inference efficiency across different hardware targets.