31 repos
Quantized large language model weights distributed in GGUF format, optimized for AMD ROCm accelerators. This cluster contains model variants (Qwen, MiMo, Tess) quantized to different precision levels (FP4, Q4_0) and compiled for ROCm-compatible hardware, enabling efficient local inference on AMD GPUs. The repositories represent pre-built model artifacts rather than framework code, with minimal language diversity reflecting their nature as model weight distributions.