MLX and Apple Silicon ML Models

39 repos

Quantized machine learning models and inference frameworks optimized for Apple Silicon using the MLX library and Metal acceleration. The cluster centers on efficient implementations of large language models and video generation systems in 4-bit, 8-bit, and bfloat16 precision, enabling on-device inference on Mac hardware. Repositories here demonstrate practical deployment strategies for models like MiniMax-H3 and LongCat video avatars using MLX's native Apple Silicon support.

Python · 1
safetensors ·1,448
apple-silicon ·1,424
mlx ·1,413
en ·1,408
zh ·1,408
quantized ·1,402
4-bit ·1,386
multimodal ·1,376
8-bit ·1,361
conversational ·1,351