14 repos
Quantized language models and text-to-speech systems optimized for efficient inference, primarily centered on GGUF-format model variants. The cluster features multiple quantization levels (Q4, Q8) and language-specific implementations, particularly for French and Spanish, alongside infrastructure for model endpoints and optimization. Repositories here focus on making neural speech synthesis and language models lightweight and deployable across different hardware constraints.