36 repos
Tools and model artifacts for converting large language models to ONNX format for efficient inference across multiple languages. The cluster centers on pre-converted model repositories (LFM2, LFM2.5 series) optimized for deployment, alongside supporting infrastructure for safetensors handling and cross-language (English, French, German, Spanish) model serving. This area addresses the practical challenge of taking large pretrained models and preparing them for production inference with reduced computational requirements.